Назад
Company hidden
2 месяца назад

Senior Site Reliability Engineer (Cloud Infrastructure)

Формат работы
remote (только Egypt/Jordan)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Egypt/Jordan
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Site Reliability Engineer (Cloud Infrastructure): Enhancing the reliability, scalability, and resilience of distributed messaging platforms across AWS, OCI, OpenStack, and Kubernetes environments with an accent on observability, automation, database performance, and disaster recovery. Focus on managing high-availability production systems, reducing incident response times, and building resilient infrastructure for critical SMS and messaging services.

Location: Remote, with listed locations in Cairo, Egypt or Amman, Jordan. The role requires nighttime availability according to the on-call schedule.

Company

hirify.global is a startup providing CPaaS communication and engagement solutions for more than 5,000 customer-centric companies.

What you will do

  • Own the reliability, uptime, scalability, and resilience of critical production services and distributed messaging platforms.
  • Participate in 24/7 on-call rotations, troubleshoot production incidents, lead post-incident analysis, and improve MTTD and MTTR.
  • Architect and manage infrastructure across AWS, Oracle Cloud Infrastructure, OpenStack, and Kubernetes environments including EKS, OKE, and Rancher RKE2.
  • Operate Kafka, RabbitMQ, Redis, MySQL, and PostgreSQL systems, focusing on availability, performance, backups, and recovery.
  • Build observability, SLO and SLI practices using Prometheus, Grafana, and CloudWatch.
  • Drive infrastructure automation, CI/CD improvements, disaster recovery, business continuity, and security initiatives using Terraform, Helm, Jenkins, Tekton, and GitLab CI/CD.

Requirements

  • Bachelor’s or master’s degree in computer science, engineering, or a related technical field.
  • 8+ years of hands-on production experience in SRE, DevOps, or cloud engineering.
  • Strong expertise in AWS, OCI, OpenStack, Kubernetes, distributed systems, and containerization.
  • Production experience with Kafka, RabbitMQ, Redis, MySQL, and PostgreSQL.
  • Advanced scripting and automation skills in Python, Bash, or Go, plus Terraform, Helm, CI/CD, and Linux administration experience.
  • Nighttime availability during the usual non-working hours of the team is required according to the on-call schedule.

Nice to have

  • Experience with Google Cloud Platform or Microsoft Azure.
  • Experience with AWS CloudFormation, GitHub Actions, ELK Stack, RCA, postmortems, and MTTR reduction.
  • Experience mentoring junior engineers and promoting SRE practices.

Culture & Benefits

  • Collaborative engineering environment focused on ownership, accountability, innovation, and knowledge sharing.
  • Competitive salary and bonus.
  • hirify.global share scheme.
  • 30 holiday days after the first anniversary and a paid birthday day off.
  • Up to 25 days per year working from anywhere in the world.
  • Paid leave and assistance for new parents.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →