2 часа назад
Systems Reliability Engineer (Kubernetes/Cloud)
100 000 - 150 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Systems Reliability Engineer (Kubernetes/Cloud): Ensuring the availability, performance, and operational excellence of large-scale distributed systems in production with an accent on infrastructure automation, observability, and incident response. Focus on operating Kubernetes and Linux at scale, designing CI/CD pipelines, and improving reliability while reducing operational toil.
Location: 100% remote within the United States
Salary: $100,000–$150,000 annually
Company
is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States.
What you will do
- Ensure the availability, performance, and operational excellence of large-scale distributed systems in production.
- Design, automate, and operate complex infrastructure, services, and production tooling.
- Operate Linux systems, Kubernetes clusters, and container-based workloads at scale.
- Design and maintain CI/CD pipelines for infrastructure and applications.
- Lead incident response and conduct post-incident reviews.
- Improve platform reliability through observability, SLOs, error budgets, and reduced operational toil.
Requirements
- Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
- At least 5 years of SRE, DevOps, or production engineering experience with large-scale distributed systems; the posting lists 6+ years of experience.
- Strong programming skills in Python, Go, or Java.
- Hands-on Linux operations, including networking, performance tuning, and systems troubleshooting.
- Production experience with Kubernetes, containers, observability tooling, and CI/CD pipelines.
- Knowledge of distributed system design, including consistency models, partitioning, and failure semantics, plus strong communication and documentation skills.
Nice to have
- Experience with SLOs, error budgets, chaos engineering, capacity planning, performance engineering, or large-scale load testing.
- Experience with AWS, Azure, or GCP.
- Familiarity with service mesh technologies such as Istio, Linkerd, or Consul.
Culture & Benefits
- Full-time direct W2 employment.
- Remote work within the United States.
- Career growth opportunities within an established organization.
- Applicants must be U.S. citizens, Green Card holders, EAD holders, or H-1B transfer candidates; new H-1B visa petitions cannot be sponsored.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
2 дня назад
Staff Production Engineer (SRE)
140 400 - 372 300$
5 дней назад
Senior Site Reliability Engineer (GovCloud)
117 000 - 209 330$
6 дней назад
Senior Site Reliability Engineer (Kubernetes)
125 000 - 145 000$
2 дня назад
Platform Engineer (AWS)
170 000 - 210 000$
3 дня назад
Senior Site Reliability Engineer (Cloud Infrastructure)
150 000 - 172 000$
Okta
2 дня назад
Staff Site Reliability Engineer (Kubernetes)
194 000 - 267 000$