4 дня назад
Senior Site Reliability Engineer (Kubernetes)
128 000 - 160 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (Kubernetes/Infrastructure): Building and scaling critical Kubernetes infrastructure across public cloud and on-premise environments with an accent on reliability, automation, and multi-cloud operations. Focus on self-healing infrastructure, GitOps delivery, Kubernetes capacity management, service-level objectives, and solving complex platform reliability challenges.
Location: Boston, Massachusetts, United States
Salary: $128,000–$160,000 base salary per year, plus bonus, equity, and applicable benefits.
Company
is a publicly traded technology company and regulated gaming operator headquartered in Boston, using AI to improve customer experiences and operations.
What you will do
- Build and scale critical Kubernetes infrastructure across multiple public clouds and on-premise environments.
- Improve platform stability, performance, scalability, and fault tolerance through self-healing infrastructure and automation.
- Operate the GitOps delivery model with Rancher Fleet, Flux, and Helm.
- Develop Kubernetes scaling and capacity strategies using Karpenter, HPA, KEDA, and predictive scaling.
- Define and monitor service-level objectives and reliability metrics using Datadog and the logging pipeline.
- Contribute to architecture and design discussions, share knowledge, and participate in an on-call rotation.
Requirements
- Bachelor’s degree in Computer Science or a related field, or equivalent education and experience.
- At least 4 years of experience managing distributed cloud and on-premise environments at scale.
- Strong hands-on experience with Amazon Web Services and deep expertise in Kubernetes and container orchestration.
- Software development experience with Go and Python for automation and infrastructure tooling.
- Knowledge of networking, Linux, Docker, containerd, packet-level debugging, and kernel troubleshooting.
- Experience with Infrastructure as Code and configuration management tools.
Nice to have
- Experience with Google Cloud Platform, vSphere, or Nutanix.
Culture & Benefits
- Bonus, equity, and benefits are available as applicable.
- AI is used to support innovation, decision-making, customer experiences, and operational improvements.
- A gaming license issued by the appropriate state agency may be required as a condition of employment; guidance is provided through the process.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Okta
7 дней назад
Staff Site Reliability Engineer, Networking (AWS/FedRAMP)
174 000 - 238 000$
Nscale
6 дней назад
Senior Site Reliability Engineer (AI Infrastructure Operations)
170 000 - 265 000$
Okta
6 дней назад
Staff Site Reliability Engineer, Networking (AWS/Networking)
174 000 - 238 000$
Okta
6 дней назад
Senior Database Reliability Engineer (DBRE, PostgreSQL)
160 000 - 220 000$
10 дней назад
Sr. Site Reliability Engineer (Kubernetes/AWS)
11 дней назад
Sr Staff Site Reliability Engineer (AI)
207 400 - 259 200$