7 дней назад
Site Reliability Engineer (Azure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Azure/Terraform): Maintaining the Project Delivery Cloud Platform through infrastructure-as-code, observability, CI/CD automation, and production incident response with an accent on cloud reliability, scalability, security, and cost optimization. Focus on designing resilient Azure or AWS environments, managing Kubernetes-based systems, leading root cause analysis, and improving operational procedures.
Location: UK - Newcastle
Company
is a global technology company connecting the physical and digital worlds through positioning, modeling, and data analytics solutions for construction, geospatial, transportation, and related industries.
What you will do
- Develop and maintain Terraform infrastructure-as-code for reliable and scalable cloud environments.
- Implement observability solutions for monitoring, logging, and alerting using tools such as New Relic, DataDog, Sumologic, and Splunk.
- Manage deployments and CI/CD pipelines with Azure DevOps, GitHub, Terraform, and related tools.
- Automate operational workflows and improve system reliability, performance, security, and efficiency.
- Lead incident response, conduct root cause analysis, and create long-term technical solutions.
- Maintain runbooks and documentation, review technical designs, and participate in on-call rotations.
Requirements
- Bachelor’s degree in Computer Engineering or a related field.
- At least 5 years of technical experience owning production infrastructure.
- Experience managing infrastructure hosted on Microsoft Azure or AWS.
- Proficiency in Python, Terraform, and containerization technologies, including Kubernetes.
- Experience with CI/CD tools such as Azure DevOps, Jenkins, Argo CD, Helm, and GitHub.
- Strong knowledge of monitoring, incident management, networking, security, capacity planning, and cost optimization.
Nice to have
- Specialized SRE observability experience with New Relic or DataDog.
- Familiarity with OpenTelemetry, AIOps, MLOps, or SecOps.
Culture & Benefits
- Values-driven environment centered on belonging, growth, and innovation.
- Opportunities to contribute to platform direction and cross-functional technical decisions.
- Emphasis on knowledge sharing, documentation, and continuous improvement.
Hiring process
- Submit an online application through the posting.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
DevOps / SRE Engineer (London)
11 дней назад
DevOps / SRE Engineer (Sheffield)
10 дней назад
Senior Site Reliability Developer (SaaS)
123 250 - 166 750CAD
13 дней назад
Site Reliability Engineer (DataCosmos)
11 дней назад
NOC Engineer / SRE
14 дней назад
Senior Site Reliability Engineer (AI Platform)
69 768 - 96 900€