10 дней назад
Global IT Site Reliability Engineer Senior Manager
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Global IT Site Reliability Engineer Senior Manager (Cloud/SRE): Running and improving enterprise reliability systems across cloud infrastructure, observability, automation, identity, security, and network operations with an accent on multi-cloud engineering, telemetry pipelines, and governance. Focus on designing reusable infrastructure patterns, reducing operational toil, leading complex incident response, and embedding Zero Trust and policy-as-code controls across teams.
Location: Hybrid or on-site work model; occasional travel may be required.
Company
is a global management consulting and technology firm delivering business transformation, digital, and technology solutions.
What you will do
- Run and continuously improve reliability engineering systems, including automation, pipelines, observability, and operational tooling.
- Design scalable engineering solutions that reduce operational toil and embed reliability into delivery workflows.
- Develop SRE standards, reusable patterns, infrastructure frameworks, and governance controls across teams.
- Lead complex incident response, systemic remediation, post-incident learning, and operational reviews.
- Mentor engineers and collaborate with engineering, platform, operations, and senior stakeholders.
- Drive improvements in service health, telemetry ingestion, automation coverage, and observability cost management.
Requirements
- 5–8 years of experience in Site Reliability Engineering, Platform Engineering, or related operational engineering disciplines.
- Strong hands-on experience with cloud, automation, observability, CI/CD, Infrastructure-as-Code, and scripting with Python.
- Experience operating cloud infrastructure across at least two of AWS, Azure, GCP, or Alibaba Cloud, including networking, identity, policies, and landing zones.
- Experience with telemetry pipelines, SLIs, SLOs, synthetic checks, alerts, observability platforms, and incident response.
- Experience with identity platforms, secrets management, OIDC, workload identity, dynamic credentials, Zero Trust, least privilege, and security tooling in CI/CD.
- Strong technical communication, stakeholder engagement, mentoring, and cross-team leadership skills; participation in an on-call rotation is required.
Nice to have
- Experience in federated, multi-cloud, or large enterprise environments.
- Familiarity with Docker, Kubernetes, policy-as-code tools, cloud certification, or Alibaba Cloud architecture.
- Experience with AIOps, event correlation, ServiceNow, PagerDuty, FinOps, chargeback, or observability capacity reviews.
Culture & Benefits
- Collaborative work across global consulting, technology, and design teams.
- Senior individual contributor role with mentorship and cross-team influence.
- Hybrid or on-site work model with occasional travel for team or stakeholder engagement.
- Equal opportunity employment and consideration under applicable local laws.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 дней назад
Site Reliability Engineering (SRE) Manager (Azure)
139 700 - 232 900$
11 дней назад
Principal Site Reliability Engineer (Kubernetes)
190 000 - 220 000$
10 дней назад
Staff Site Reliability Engineer (Cybersecurity)
199 750 - 270 000$
14 дней назад
Senior SRE (Site Reliability Engineer) – Modernized Application Operations
145 000 - 170 000$
11 дней назад
Sr. Site Reliability Engineer (AI)
10 дней назад