Назад
2 дня назад

Staff Site Reliability Engineer (Kubernetes)

194 000 - 267 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff Site Reliability Engineer (Kubernetes): Building and managing reliable, scalable, and secure Kubernetes platforms on AWS with an accent on high availability, platform automation, dynamic scaling, and service mesh management. Focus on creating production-grade EKS platforms, implementing Helm and Karpenter, securing Istio, optimizing cloud costs, and resolving complex performance and availability incidents.

Location: Hybrid role with locations in Bellevue, Washington; Chicago, Illinois; New York, New York; San Francisco, California; and Washington, DC. In-person onboarding and travel to the San Francisco headquarters or Chicago office is required during the first week. U.S. Person status is required to access federal environments or protected federal data.

Annual base salary: $194,000–$267,000 USD for candidates in California excluding the San Francisco Bay Area, Colorado, Illinois, New York, and Washington; $174,000–$239,000 USD for other applicable locations. Equity, bonus, and benefits may also apply.

Company

Okta provides secure identity infrastructure for workforce access and AI adoption through its Workforce Identity Cloud.

What you will do

  • Design, implement, and maintain highly available, scalable, and fault-tolerant Kubernetes platforms for cloud-native applications.
  • Build, manage, and optimize AWS infrastructure including EKS, ECS, S3, VPCs, RDS, IAM, and related services.
  • Create and maintain Helm charts and automate application and infrastructure deployment through CI/CD pipelines.
  • Implement Karpenter for dynamic Kubernetes scaling and optimize resource usage and cloud costs.
  • Configure and secure Istio service mesh capabilities including traffic management, service discovery, observability, and policy enforcement.
  • Respond to incidents, troubleshoot availability and performance issues, implement security controls, and document operational practices.

Requirements

  • 4+ years of experience with Kubernetes and Helm, 4+ years with Terraform, and 5+ years with AWS.
  • Experience with multi-region cloud environments and cloud-native architectures.
  • Strong expertise in Kubernetes platform creation, networking, storage, high availability, Helm, Karpenter, and Istio.
  • Proficiency with CI/CD pipelines and automation tools such as Jenkins, GitLab, CircleCI, Terraform, Ansible, or Spinnaker.
  • Strong scripting and automation skills in Python, Bash, or Go, plus experience with monitoring, logging, and alerting tools such as Prometheus, Grafana, CloudWatch, or ELK Stack.
  • Ability to submit documentation establishing U.S. Person status upon hire.

Nice to have

  • Knowledge of cloud and Kubernetes security practices, including RBAC, encryption, and compliance frameworks.
  • Familiarity with Docker and containerization.
  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience.
  • CKA, CKAD, or AWS Certified DevOps Engineer certification.

Culture & Benefits

  • Health, dental, and vision insurance.
  • 401(k) and flexible spending account.
  • Paid leave, including PTO and parental leave.
  • Equity and bonus opportunities where applicable.
  • In-person onboarding designed to connect new hires with the organization and mission.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →