Назад
Company hidden
обновлено 21 час назад

Site Reliability Engineer (Kubernetes)

Формат работы
hybrid
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
France
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer (Kubernetes): Building and improving observability, alerting, incident management, and reliability tooling for BlaBlaCar services with an accent on SLOs, SLIs, production troubleshooting, and infrastructure automation. Focus on developing Terraform modules and Go tools, leading high-severity incidents, applying chaos engineering, and using AI to reduce operational toil.

Location: Paris, France; hybrid work with 2–3 days per week in the office. One hiring interview will be onsite.

Company

hirify.global is a community-based travel platform for carpooling and bus travel, serving 27 million members annually across 21 countries.

What you will do

  • Create, maintain, and improve observability and alerting tools and frameworks for software engineering teams.
  • Own the Service Level Objectives framework and support the design and maintenance of service indicators.
  • Define incident management standards and improve processes through post-mortems and chaos engineering.
  • Act as Incident Commander during high-severity production incidents when required.
  • Develop Terraform modules, Go applications, and other tools to automate and improve service reliability.
  • Build operational metrics and incident reporting to drive continuous improvement across teams.

Requirements

  • 3–7 years of experience in SRE, DevOps, or software engineering.
  • Strong knowledge of observability, including metrics, logging, and tracing, with experience using tools such as Datadog.
  • Production troubleshooting and on-call experience in diagnosing and resolving technical issues.
  • Strong communication skills for working with multidisciplinary teams and adapting to different levels of technical expertise.
  • Full working proficiency in English.
  • Ability to take ownership, work with business priorities, and navigate different contexts.

Nice to have

  • Experience with Kubernetes and incident management platforms such as Grafana IRM.
  • Experience with SLOs, SLIs, and OpenTelemetry integration.
  • Programming experience in Go or interest in learning Go; familiarity with object-oriented programming and scripting languages.
  • Experience with web or mobile testing tools and testing software at scale.

Culture & Benefits

  • Collaborative, fast-growing, and innovative environment with small agile teams and strong ownership.
  • Four additional weeks of maternity or paternity leave beyond legal requirements.
  • 50% healthcare coverage, home-office equipment support, and at least 25 days of holiday per year.
  • 50% transportation coverage, local meal plan, and unlimited free carpooling and bus rides.
  • Training, mentorship, internal mobility opportunities, employee stock ownership, and regular team-building events.
  • One paid day off per year to test the product.

Hiring process

  • 45-minute video call with a Talent Acquisition Manager.
  • 60-minute call with the Hiring Manager, followed by a 90-minute system design interview with two team members.
  • 45-minute call with the Head of Foundations; the process typically takes 25–30 days, with offers usually issued within 48 hours.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →