9 часов назад
Manager, Software Engineering (SRE)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Manager, Software Engineering (SRE) (AWS/Kubernetes): Leading a distributed SRE team responsible for improving the reliability, scalability, operational maturity, and delivery systems of cloud-based services with an accent on incident management, observability, automation, security, and cost optimization. Focus on building operational roadmaps, strengthening CI/CD and platform enablement, reducing toil, and guiding high-availability service operations across engineering stakeholders.
Location: Remote in the United States. Applicants must have legal authorization to work in the jurisdiction where the position is posted and must not require employer sponsorship.
Company
provides cybersecurity products and cloud-based security services.
What you will do
- Lead, coach, hire, and develop engineers working across SRE, platform operations, release engineering, and automation.
- Define and execute the SRE roadmap, improving reliability, availability, scalability, production readiness, and cloud cost efficiency.
- Strengthen incident response, escalation management, observability, runbooks, post-incident reviews, alerting, and root-cause analysis.
- Guide improvements to CI/CD pipelines, release automation, Kubernetes/EKS environments, infrastructure, and shared engineering tooling.
- Partner with engineering, architecture, security, product, compliance, and operations stakeholders on planning, prioritization, and delivery.
- Present operational risks, reliability trends, roadmap progress, and investment needs to technical leaders and executives.
Requirements
- Extensive experience in SRE, DevOps, platform engineering, infrastructure management, cloud infrastructure, or production operations.
- Proven experience leading and managing engineering teams and improving operational reliability and delivery.
- Strong knowledge of AWS, Kubernetes/EKS, Docker, Helm, Linux, Terraform, Ansible, CI/CD, and release automation.
- Experience with scripting and automation using Python, Bash, or Groovy, plus monitoring and observability tools such as Prometheus, Grafana, ELK, or CloudWatch.
- Knowledge of incident management, security, vulnerability management, compliance controls, audit readiness, vendor management, and Agile/Scrum practices.
- Bachelor’s degree in Computer Science, Information Technology, or a related field, or equivalent practical experience.
Nice to have
- Experience with Azure, Java-based services, build systems, or code review workflows.
- Master’s degree or experience with responsible AI-assisted engineering tools.
Culture & Benefits
- Remote-first work environment.
- Permanent employment within the Product Development – Software Engineering department.
- Focus on ownership, collaboration, continuous improvement, and operational excellence.
- Opportunity to work with distributed engineering teams across the United States and Canada.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
Okta
2 дня назад
Manager, Site Reliability Engineering (AWS/Kubernetes)
204 000 - 306 000$
Okta
1 день назад
Manager, Site Reliability Engineering (AWS/Kubernetes)
204 000 - 306 000$
3 дня назад
Sr. Site Reliability Engineer / SWE
6 дней назад
Manager of Site Reliability Engineering (Azure/AWS)
160 000 - 180 000$
14 часов назад
DevOps/SRE Engineer
3 дня назад