Назад
Company hidden
7 дней назад

Site Reliability Engineer

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer (Azure/Terraform): Building and maintaining reliable cloud infrastructure for Trimble’s Project Delivery Cloud Platform with an accent on infrastructure as code, observability, CI/CD, and production operations. Focus on leading incident response, analyzing root causes, improving system reliability and security, and optimizing capacity and operating costs.

Location: Newcastle, United Kingdom

Company

hirify.global develops technology that connects the physical and digital worlds, with digital construction solutions for architecture, engineering, construction, and owners.

What you will do

  • Develop and maintain infrastructure as code with Terraform for reliable and scalable cloud environments.
  • Implement observability, monitoring, logging, and alerting using tools such as New Relic, Datadog, Sumo Logic, and Splunk.
  • Manage deployments and CI/CD pipelines with Azure DevOps, GitHub, Terraform, and related tools.
  • Automate operational workflows and improve capacity planning and cost optimization.
  • Evaluate architectures for reliability, performance, security, and efficiency.
  • Lead incident response, perform root cause analysis, maintain runbooks, and participate in on-call rotations.

Requirements

  • Bachelor’s degree in Computer Engineering or a related field.
  • At least 5 years of technical experience owning production infrastructure.
  • Expertise with Microsoft Azure or AWS, Python, Terraform, and containerization.
  • Experience with Kubernetes or other containerization technologies.
  • Familiarity with CI/CD tools including Azure DevOps, Jenkins, Argo CD, Helm, and GitHub.
  • Strong understanding of monitoring, incident management, networking, and security concepts.

Nice to have

  • Specialized SRE observability experience with New Relic or Datadog.
  • Familiarity with OpenTelemetry, AIOps, MLOps, or SecOps.

Culture & Benefits

  • Values-driven culture focused on belonging, growth, and innovation.
  • Cross-functional collaboration and knowledge sharing.
  • Participation in an on-call rotation for critical incidents.

Hiring process

  • Submit an online application through the posting.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →