обновлено 12 дней назад
Principal Systems Engineer (SRE)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Principal Systems Engineer (SRE) (Azure/Windows/SQL Server): Ensuring the reliability, scalability, and performance of hosted healthcare platforms with an accent on observability, incident response, cloud operations, and automation. Focus on leading complex application, database, and infrastructure troubleshooting, defining SLIs and SLOs, and improving 24x7 service availability for UK-based remote operations.
Location: Remote within the United Kingdom
Company
develops hosted healthcare platforms that help healthcare providers deliver better patient care.
What you will do
- Maintain and improve the reliability, availability, and performance of production healthcare platforms.
- Investigate and resolve complex application, database, infrastructure, and cloud issues.
- Lead incident management, root cause analysis, and post-incident reviews.
- Define and measure service level indicators and objectives, including error budgets.
- Develop monitoring, alerting, scripting, and Infrastructure-as-Code automation.
- Provide technical leadership during major incidents and support deployment and monitoring improvements.
Requirements
- 7+ years of experience supporting enterprise applications, infrastructure, or cloud environments.
- Strong experience with observability and APM tools such as LogicMonitor, AppDynamics, Azure Monitor, Dynatrace, Datadog, or New Relic.
- Deep knowledge of Windows Server, IIS, .NET applications, Windows Clustering, MSMQ, Event Logs, and PerfMon.
- Strong SQL Server experience, including performance tuning, query optimization, blocking analysis, and Always On Availability Groups.
- Experience with Azure and networking fundamentals including DNS, TCP/IP, load balancing, and firewalls.
- Familiarity with ServiceNow or other ITSM platforms and ITIL principles.
Nice to have
- Experience with PowerShell, Python, Terraform, ARM Templates, or Bicep.
- Knowledge of CI/CD pipelines, Azure DevOps, GitHub Actions, Kubernetes, and containerized workloads.
- Experience implementing SLIs, SLOs, and error budgets.
- Background in healthcare technology or patient care environments.
- Bachelor's degree in Computer Science, Information Technology, or Engineering.
Culture & Benefits
- Remote work for candidates within the UK.
- Participation in an on-call rotation supporting a 24x7 healthcare environment.
- Occasional after-hours work for activations, upgrades, and major incidents.
- Competitive compensation and benefits package.
- No travel is required.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
12 дней назад
Senior Site Reliability Engineer (Azure)
100 000 - 140 600CAD
13 дней назад
Lead Site Reliability Engineer (Azure)
113 000 - 142 300CAD
12 дней назад
Engineering Manager, SRE (AI)
12 дней назад
Senior Site Reliability Engineer (Azure/AWS)
130 000 - 160 000$
13 дней назад
Site Reliability Engineer (Azure/Terraform)
105 600 - 145 200$
14 дней назад
DevOps/SRE инженер (AWS/Kubernetes)
4 200 - 6 200€