Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (FedRAMP) (Kubernetes, Terraform, Go, Python): Building and operating reliable, scalable, and secure cloud services with an accent on automation, observability, and FedRAMP compliance. Focus on designing production infrastructure, troubleshooting Kubernetes and distributed data platforms, improving incident response, and eliminating operational toil through platform engineering.
Location: San Francisco, California; hybrid role. US Person status (US Citizen or Green Card Holder) is required for US FedRAMP projects.
Annual base salary in the San Francisco Bay Area: $147,000–$202,000 USD. Equity, bonus, health, dental and vision insurance, 401(k), flexible spending account, PTO, and parental leave may also apply.
Company
Okta builds identity and access management infrastructure that helps organizations securely operate cloud and AI-enabled services.
What you will do
- Design, build, and operate large-scale cloud infrastructure and customer-facing production services.
- Improve availability, scalability, performance, resilience, and operational readiness through SLIs, SLOs, error budgets, and capacity planning.
- Develop automation, internal platforms, self-service capabilities, and operational guardrails with Go, Python, Terraform, and related technologies.
- Improve deployment safety and operational workflows through CI/CD, GitOps, Kubernetes, Helm, and ArgoCD.
- Lead incident response, post-incident reviews, observability improvements, and long-term operational ownership.
- Collaborate with engineering, architecture, and product teams while mentoring junior and mid-level engineers.
Requirements
- Strong experience operating large-scale production services in AWS and/or GCP, with deep Linux and Kubernetes expertise.
- Strong software engineering skills in Go and/or Python, plus extensive Infrastructure as Code experience with Terraform and Helm.
- Experience troubleshooting Kubernetes networking, storage, scheduling, scaling, workload lifecycle, and distributed data platforms.
- Knowledge of cloud networking, IAM, secrets management, secure infrastructure design, observability, monitoring, and production telemetry.
- Experience with customer-facing systems, incident response, SLA-driven operations, CI/CD, and reliability engineering practices.
- US Person status is required: US citizenship or a Green Card.
Nice to have
- Experience with FedRAMP, SOC 2, HIPAA, or other regulated and security-sensitive cloud environments.
- Experience operating SaaS platforms and Kubernetes-based microservices at scale.
- Experience with AI-assisted operational tooling, GitOps, and ArgoCD.
Culture & Benefits
- Automation-first culture focused on operational excellence, continuous improvement, and engineering quality.
- Participation in a global on-call rotation supporting highly available systems.
- In-person onboarding designed to connect new employees with the organization and its mission.
- Benefits include insurance, 401(k), flexible spending account, paid time off, parental leave, equity, and bonus eligibility where applicable.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
2 дня назад
Senior Software Development Engineer (US Federal)
163 800 - 245 800$
4 дня назад
Senior Site Reliability Engineer (FedRAMP)
140 000 - 185 000$
4 дня назад
Site Reliability Engineer (Kubernetes)
123 000 - 150 000$
5 дней назад
Senior Site Reliability Engineer (GovCloud)
117 000 - 209 330$
3 дня назад
Senior Site Reliability Engineer (Fintech)
160 000 - 200 000$
2 дня назад
Site Reliability Engineer - Enterprise Technology (Kubernetes)
200 000 - 250 000$