Назад
Company hidden
5 часов назад

Site Reliability Engineer (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer (AI) (Linux/Kubernetes/Observability): Designing and operating systems, networks, and services that keep Seekr's AI search platform within SLAs with an accent on reliability, observability, scalability, and incident response. Focus on building automation, load testing applications, monitoring production services, and solving complex operational challenges across hybrid cloud environments.

Location: Hybrid work from offices in Austin, Texas, or Reston, Virginia, United States. Applicants must be U.S. citizens, lawful permanent residents, protected individuals, or eligible to obtain the required U.S. government authorizations.

Company

hirify.global develops explainable and trustworthy AI solutions for enterprises, government, and regulated industries across cloud, on-premises, and edge environments.

What you will do

  • Design, architect, and implement systems, software, networks, and services that keep the hirify.global platform within SLAs.
  • Lead solutions for complex operational, reliability, scalability, and production failure-detection challenges.
  • Collaborate with software engineering teams to build, monitor, troubleshoot, and triage services.
  • Perform application load testing and take ownership of the observability stack.
  • Develop incident response practices and troubleshoot live-system and deployment issues.
  • Automate operational workflows and eliminate inefficiencies.

Requirements

  • 5+ years of experience in Site Reliability Engineering, preferably in SaaS environments.
  • 5+ years of experience with Linux systems and network protocols.
  • Experience with self-hosted monitoring, metrics, and centralized logging tools including ELK, Prometheus, InfluxDB, and Grafana.
  • Proficiency in programming or scripting with Python, Ruby, Bash, or Java.
  • Hands-on experience with Docker, Kubernetes, load testing, Git, and configuration-management automation.
  • Experience with hybrid cloud environments and the ability to maintain SLAs while communicating across technical and business teams.

Nice to have

  • Experience with Puppet and Terraform.
  • Experience monitoring Kubernetes, Elasticsearch, Kafka, and Aerospike.
  • Familiarity with GitLab and ArgoCD.
  • On-premises infrastructure experience.
  • Advanced degree.

Culture & Benefits

  • Flexible hybrid work environment with offices in Austin and Reston.
  • Unlimited paid time off and 14 paid company holidays.
  • RSU equity ownership and role-appropriate total rewards.
  • 401(k) plan with company match.
  • Medical, dental, vision, and life insurance from day one.
  • Paid parental leave and a mission focused on secure, transparent AI.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →