Senior Site Reliability Engineer (Kubernetes/Azure)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Senior Site Reliability Engineer (Kubernetes/Azure): Designing and operating resilient cloud platforms, Kubernetes clusters, networking, service mesh components, and automated infrastructure provisioning with an accent on Infrastructure as Code, GitOps, and operational excellence. Focus on building dependency-aware bootstrap orchestration, reducing operational toil, improving incident response and MTTR, and supporting critical systems under security and compliance requirements.
Location: Hybrid position in Austin, Texas, United States; office attendance expectations will be confirmed by the hiring manager. The role is performed in an office setting and requires 5–10% travel.
Salary: $110,700–$171,800 USD per year, with possible bonus, equity, and applicable sales incentive payments.
Company
Global payments technology company facilitating transactions among consumers, merchants, financial institutions, and government entities in more than 200 countries and territories.
What you will do
- Own the lifecycle of cloud infrastructure, Kubernetes clusters, networking, ingress, service discovery, and service mesh components.
- Design resilient platform components using fault isolation, graceful degradation, capacity planning, and saturation control.
- Lead automated infrastructure bootstrap orchestration across cloud, network, and Kubernetes layers.
- Drive Infrastructure as Code and GitOps practices with reproducible, testable, auditable, and reversible changes.
- Participate in on-call escalation, incident response, post-incident reviews, and problem management.
- Collaborate with engineering teams to improve platform adoption, reliability standards, cloud-native practices, and day-2 operations.
Requirements
- 2+ years of relevant experience with a bachelor's degree, or 5+ years of relevant experience.
- English proficiency at B2 level or above is required.
- Strong hands-on experience with public cloud platforms, with Azure mandatory and AWS preferred.
- Production experience operating Kubernetes at scale, container orchestration platforms, cloud networking, IAM and security concepts, and reliability patterns.
- Experience with Terraform or comparable Infrastructure as Code, automation-first workflows, GitOps, and CI/CD pipelines.
- Experience with distributed-systems troubleshooting, observability, service mesh technologies, critical systems, incident management, runbooks, and toil reduction.
Nice to have
- Experience with AWS, Terraform, Istio, App Mesh, or Linkerd.
- Advanced degree such as a master's, MBA, JD, or MD.
- Experience documenting infrastructure procedures and providing first-level support for deployment issues.
Culture & Benefits
- Hybrid work arrangement with office-based collaboration and variable working hours according to departmental needs.
- Medical, dental, and vision coverage.
- 401(k), FSA/HSA, life insurance, paid time off, and wellness program.
- Potential eligibility for bonus and equity.
- Work focused on security, compliance, internal controls, and reliability for mission-critical payment systems.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →