7 дней назад
Site Reliability Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Azure/Terraform): Building and maintaining reliable cloud infrastructure for Trimble’s Project Delivery Cloud Platform with an accent on infrastructure as code, observability, CI/CD, and production operations. Focus on leading incident response, analyzing root causes, improving system reliability and security, and optimizing capacity and operating costs.
Location: Newcastle, United Kingdom
Company
develops technology that connects the physical and digital worlds, with digital construction solutions for architecture, engineering, construction, and owners.
What you will do
- Develop and maintain infrastructure as code with Terraform for reliable and scalable cloud environments.
- Implement observability, monitoring, logging, and alerting using tools such as New Relic, Datadog, Sumo Logic, and Splunk.
- Manage deployments and CI/CD pipelines with Azure DevOps, GitHub, Terraform, and related tools.
- Automate operational workflows and improve capacity planning and cost optimization.
- Evaluate architectures for reliability, performance, security, and efficiency.
- Lead incident response, perform root cause analysis, maintain runbooks, and participate in on-call rotations.
Requirements
- Bachelor’s degree in Computer Engineering or a related field.
- At least 5 years of technical experience owning production infrastructure.
- Expertise with Microsoft Azure or AWS, Python, Terraform, and containerization.
- Experience with Kubernetes or other containerization technologies.
- Familiarity with CI/CD tools including Azure DevOps, Jenkins, Argo CD, Helm, and GitHub.
- Strong understanding of monitoring, incident management, networking, and security concepts.
Nice to have
- Specialized SRE observability experience with New Relic or Datadog.
- Familiarity with OpenTelemetry, AIOps, MLOps, or SecOps.
Culture & Benefits
- Values-driven culture focused on belonging, growth, and innovation.
- Cross-functional collaboration and knowledge sharing.
- Participation in an on-call rotation for critical incidents.
Hiring process
- Submit an online application through the posting.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
11 дней назад
DevOps / SRE Engineer (London)
11 дней назад
DevOps / SRE Engineer (Sheffield)
10 дней назад
Senior Site Reliability Developer (SaaS)
123 250 - 166 750CAD
11 дней назад
NOC Engineer / SRE
13 дней назад
Site Reliability Engineer (DataCosmos)
14 дней назад
Site Reliability Engineer (AI Platform)
69 768 - 96 900€