Назад
Company hidden
2 месяца назад

Site Reliability / Production Engineer (SaaS)

2 667€
Формат работы
remote (только Turkey)
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
Turkey
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability / Production Engineer (SaaS): Managing and optimizing live production systems for a B2B Stories platform with an accent on incident response, technical triage, and system reliability. Focus on reducing alert noise, automating operational tasks, and utilizing AI tools for diagnostic and mitigation workflows.

Location: Remote (Must be based in Turkey)

Salary: Up to EUR 32,000 per year

Company

hirify.global is a high-growth B2B SaaS platform that enables companies to integrate Stories into their own platforms to boost user engagement and advertising revenue.

What you will do

  • Respond to live incidents, establish customer impact, and execute proportionate mitigations, rollbacks, or bounded fixes.
  • Perform technical triage using logs, metrics, traces, and AI tools to diagnose system failures.
  • Coordinate the response by escalating to product teams only when deep product knowledge is required.
  • Improve the reliability system by tuning low-value alerts and designing monitoring around real customer outcomes.
  • Create runbooks, automate repeated operational work, and close observability gaps.
  • Analyze material incidents with AI to identify and resolve repeated failure patterns.

Requirements

  • Must be based in Turkey.
  • Ability to work UK-time evening shifts (17:00-01:00) and weekday overnight pager duty (01:00-06:00).
  • Proven ownership of ambiguous live problems and strong operational judgement under uncertainty.
  • Comfortable exploring unfamiliar systems via code, logs, APIs, and command-line tools.
  • AI-native approach to technical work, including hypothesis generation and automation.
  • Experience with live production systems or on-call rotas is strongly preferred.

Nice to have

  • Experience with Azure or Cloudflare.
  • Knowledge of distributed application and API diagnostics.
  • Proficiency with databases, queues, and background-processing systems.
  • Experience with observability, alerting, and incident-management platforms.
  • Application development and safe production debugging skills.

Culture & Benefits

  • Fully remote work arrangement within Turkey.
  • Opportunity to work with a high-growth product used by leading global brands in sports.
  • Access to latest hardware, AI tools, and modern product workflows.
  • Structured shift rota ensuring shared coverage with appropriate rest days.

Hiring process

  • Hiring Manager Conversation (20-30 mins) to discuss the role and working pattern.
  • Paid Take-home Task (60-90 mins) involving a bounded production-incident exercise.
  • Task Review and Interview with the CTO (60-75 mins) to discuss decisions and production judgement.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →