Назад
Company hidden
3 дня назад

Major Incident Manager

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
UK/Singapore/US +4 еще
Релокация
Singapore
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Major Incident Manager (Infrastructure, Security, and Incident Response): Monitoring critical security and infrastructure services, managing SLOs and error budgets, and engineering automated operational and security controls with an accent on infrastructure as code, observability, and DevSecOps. Focus on automating toil, improving incident response and post-mortem actions, and reducing MTTR through preventive engineering.

Location: Singapore; on-site. Candidates must be based in Singapore or be willing to relocate there.

Company

hirify.global develops AI code verification and governance products that help enterprises build reliable, secure, and maintainable software.

What you will do

  • Monitor critical security platforms and core infrastructure, triage high-severity alerts, and manage SLOs and error budgets.
  • Write, review, and test infrastructure and policy as code using tools such as Python, Go, and Terraform.
  • Automate repetitive operational tasks, including security patching, compliance checks, certificate rotation, and infrastructure maintenance.
  • Maintain DevSecOps security tooling in CI/CD pipelines and improve logging, metrics, and tracing capabilities.
  • Participate in the on-call rotation, lead incident response activities, and convert post-mortem findings into preventative engineering work.

Requirements

  • Professional experience with infrastructure as code and configuration management tools such as Terraform, CloudFormation, Ansible, or Puppet.
  • Hands-on experience with AWS, GCP, Azure, or large-scale private cloud infrastructure.
  • Experience defining and measuring SLIs and SLOs for critical services, including error budget management.
  • Experience with observability platforms such as Prometheus, Grafana, ELK/EFK, Datadog, or Splunk.
  • Strong networking knowledge covering TCP/IP, DNS, load balancing, firewalls, and proxies.
  • Experience automating security controls, managing IAM systems, and conducting post-incident reviews.

Culture & Benefits

  • On-site work in Singapore with anchor collaboration days on Mondays, Tuesdays, and Thursdays.
  • Support for relocation to the posted role location.
  • Inclusive work environment with a focus on diversity, equity, and inclusion.
  • Employment is subject to background-check and reference-verification results.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →