Назад
Company hidden
5 часов назад

TechOps Engineer (AI)

Формат работы
onsite
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
Germany
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
TechOps Engineer (AI): Improving the reliability, scalability, and operability of Talon.One’s production systems with an accent on automation, AI-driven operations, observability, and incident workflows. Focus on building guarded AI agents, automating runbooks and operational tooling, investigating production data, and supporting deployment and database maintenance tasks.

Location: Berlin, Germany

Company

hirify.global builds a scalable promotions engine with a product engineering organization based in Berlin.

What you will do

  • Identify operational toil across R&D and build scripts, automation, and internal tools to eliminate repetitive work.
  • Design and integrate AI agents for operational workflows with guardrails, monitoring, and human oversight.
  • Own and optimize Incident.io workflows, integrations, and automation, and contribute to post-incident reviews.
  • Maintain monitoring, alerting, dashboards, and other observability tooling while investigating logs, metrics, and production data.
  • Turn operational procedures into clear, automated runbooks in partnership with SRE and R&D teams.
  • Support database maintenance, health checks, and release and deployment workflows affecting production reliability.

Requirements

  • 2–4 years of experience in TechOps, DevOps, SRE, Production Engineering, or a similar technical role.
  • Experience operating production systems in a SaaS or cloud environment.
  • Comfort with Linux, command-line tools, logs, monitoring, scripting, and automation.
  • Structured troubleshooting skills and a proactive approach to improving systems, processes, and tooling.
  • Ability to collaborate with engineers across different teams and develop deeper expertise in production reliability.
  • Experience with the listed production stack, including GCP, Kubernetes, Terraform, Helm, Grafana, Datadog, Sentry, PostgreSQL, Go, Python, Bash, and CI/CD tooling.

Nice to have

  • Familiarity with SRE concepts such as SLIs, SLOs, and error budgets.
  • Experience with observability platforms such as Grafana, Datadog, Prometheus, or Sentry.
  • Experience with Kubernetes, GCP, APIs, integrations, CI/CD, or infrastructure automation.

Culture & Benefits

  • International, diverse, collaborative, and flexible work environment.
  • €1,000 annual learning budget and free German language courses.
  • 30 days of annual leave plus paid birthday and moving days.
  • Home office setup budget and monthly home office allowance.
  • Ability to work from abroad worldwide for up to 90 days.
  • Mental health support, Urban Sports Club membership, pension contribution subsidy, public transport subsidy, and BusinessBike access.
  • Dog-friendly Berlin office and family-friendly working arrangements.

Hiring process

  • Apply directly through the application form.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →