1 день назад
Staff Site Reliability Engineer (GCP/Kubernetes)
112 500 - 187 500$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff Site Reliability Engineer (GCP/Kubernetes): Driving reliability strategy and high-risk platform initiatives across GCP, Kubernetes, CI/CD, monitoring, databases, networking, and infrastructure-as-code with an accent on enterprise-scale availability and security. Focus on capacity planning, load testing, incident response, maintenance events, and integrating production-grade AI/ML capabilities into cloud-native platforms.
Location: Chicago, Illinois; hybrid work with regular virtual responsibilities and in-person work at an assigned office for a minimum of two days per week.
Salary: $112,500–$187,500 annually.
Company
provides technology and data solutions that support customers and communities globally.
What you will do
- Drive reliability strategy and contribute to architectural decisions for major platform components.
- Research, test, implement, and continuously improve systems and engineering tooling.
- Lead capacity planning, load testing, security improvements, and high-risk maintenance events.
- Participate in the on-call rotation and provide technical leadership during major incidents.
- Elevate engineering standards through tooling, procedures, communication, and cross-functional collaboration.
- Integrate reliable, production-grade AI/ML capabilities into cloud-native platforms.
Requirements
- 5+ years of experience in cloud architecture, site reliability engineering, platform engineering, or a related field, with enterprise-scale delivery experience.
- Deep hands-on expertise with GCP and Kubernetes for high-volume, highly available workloads with 99.999% reliability targets.
- Strong CI/CD pipeline architecture, observability, monitoring, alerting, Linux, database, networking, and infrastructure-as-code expertise.
- Experience with Terraform, Pulumi, or equivalent tools, plus scripting and automation using Python, Bash, or Go.
- Experience designing and integrating AI/ML solutions, including LLM orchestration, vector databases, model serving infrastructure, and AI observability.
- Ability to work in a hybrid arrangement with at least two days per week in person at an assigned TU office.
Culture & Benefits
- Medical, dental, and vision coverage available from day one, with HSA and FSA options.
- Company-paid life and disability insurance, with optional family and legal coverage.
- Up to 12 weeks of paid parental leave, adoption and fertility assistance, and caregiver support.
- 401(k) with employer match, ESPP, tuition reimbursement, career coaching, and financial wellness resources.
- Flexible time off or paid time off, up to 12 paid holidays, commuter benefits, and paid volunteer time.
- 24/7 access to therapy, coaching, meditation, and broader wellness programs.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Staff Site Reliability Engineer (Fintech)
132 100 - 220 100$
2 дня назад
Staff Site Reliability Engineer (IoT)
102 600 - 193 425$
2 дня назад
Senior DevOps Engineer (GCP/Kubernetes)
5 дней назад
Site Reliability Engineer - Enterprise Technology (Kubernetes)
200 000 - 250 000$
Okta
5 дней назад
Staff Site Reliability Engineer (Kubernetes)
194 000 - 267 000$
7 дней назад
Staff Site Reliability Engineer (AI/Blockchain)
195 000 - 257 500$