Назад
4 дня назад

Engineer II, Site Reliability (Cybersecurity)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
UK
Вакансия из списка Hirify.GlobalВакансия из Hirify RU Global, списка компаний с восточно-европейскими корнями
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Engineer II, Site Reliability (Cybersecurity): Operating and improving mission-critical commercial cloud services across large-scale distributed systems with an accent on Linux administration, automation, monitoring, and performance engineering. Focus on leading incident analysis, troubleshooting infrastructure and hardware issues, and maintaining availability, latency, throughput, and capacity across a 24x7 platform.

Location: Hybrid in London, United Kingdom

Company

Cybersecurity company building an AI-native platform that protects organizations and operates large-scale distributed systems.

What you will do

  • Operate and improve the commercial cloud platform across availability, latency, throughput, monitoring, incident response, and capacity planning.
  • Administer Linux systems across thousands of bare-metal servers and virtual machines.
  • Develop automation, tooling, and software for mission-critical distributed services.
  • Monitor and tune operating system and application performance using telemetry and observability platforms.
  • Lead incident analysis, correlate incidents with systemic problems, and drive remediation.
  • Participate in on-call rotations and troubleshoot server hardware issues in a 24x7 environment.

Requirements

  • Bachelor’s degree or equivalent experience in Computer Science.
  • At least five years of experience in large-scale production environments and two years of software engineering experience.
  • At least two years of experience with C++, Java, Python, or Go.
  • Experience with Linux and infrastructure technologies such as Windows, VMware, Docker, or Kubernetes.
  • Experience with storage technologies, configuration management tools such as Puppet, Chef, or Ansible, and monitoring stacks such as ELK, Prometheus, Grafana, or Zabbix.
  • Strong analytical, communication, documentation, ownership, and incident-response skills, including experience using AI technologies to improve decision-making and workflows.

Culture & Benefits

  • Hybrid work with a vibrant office culture and workplace amenities.
  • Comprehensive physical and mental wellness programs.
  • Competitive vacation and holidays, plus paid parental and adoption leave.
  • Professional development opportunities for employees at every level.
  • Employee networks, geographic groups, and volunteer opportunities.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →