Назад
Company hidden
4 часа назад

Platform Reliability Engineer (Kubernetes)

100 000 - 150 000$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Platform Reliability Engineer (Kubernetes): Operating and improving large-scale distributed systems in production with an accent on reliability engineering, automation, observability, and infrastructure performance. Focus on building CI/CD tooling, managing Kubernetes and Linux environments, leading incident response, and reducing operational toil through SLOs and resilient system design.

Location: 100% remote within the United States

Salary: $100,000–$150,000 annually

Company

hirify.global is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.

What you will do

  • Ensure the availability, performance, and operational excellence of large-scale distributed systems in production.
  • Build automation and tooling using Python, Go, or Java to improve infrastructure and operations.
  • Operate Linux systems, Kubernetes clusters, and container-based workloads at scale.
  • Design and operate CI/CD pipelines for infrastructure and applications.
  • Lead incident response and conduct effective post-incident reviews.
  • Improve platform reliability through observability, SLOs, error budgets, and reduced operational toil.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
  • At least 5 years of SRE, DevOps, or production engineering experience with large-scale distributed systems.
  • Strong programming skills in Python, Go, or Java.
  • Hands-on Linux experience, including networking, performance tuning, and systems-level troubleshooting.
  • Production experience with Kubernetes, containers, observability tools, and CI/CD pipelines.
  • Applicants must be U.S. citizens, Green Card holders, EAD holders, or H-1B transfer candidates; new H-1B visa petitions cannot be sponsored.

Nice to have

  • Experience defining and operationalizing SLOs and error budgets.
  • Exposure to chaos engineering tools such as Chaos Monkey, Gremlin, or Litmus.
  • Experience with AWS, Azure, or GCP.
  • Background in capacity planning, performance engineering, or large-scale load testing.
  • Familiarity with Istio, Linkerd, Consul, or other service mesh technologies.

Culture & Benefits

  • Full-time direct W-2 employment.
  • Opportunity for career growth within an established technology consulting and software development organization.
  • Work focused on cloud, AI, data, enterprise systems, and large-scale distributed platforms.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →