4 дня назад
Site Reliability Engineering (SRE) Manager
106 000 - 130 600$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineering (SRE) Manager (Cloud/Automation): Building and improving reliability practices, platforms, and automation tools for complex systems with an accent on observability, incident response, performance, and capacity management. Focus on leading SREs, reducing operational toil through automation, resolving reliability challenges, and applying AI to monitoring and response.
Location: Portland, Maine; Boston, Massachusetts; Dallas, Texas; or Chicago, Illinois
Base pay: $106,000–$130,600 per year
Company
provides platforms and products supporting customer businesses, including B2B payments, expense reimbursements, and virtual card creation.
What you will do
- Lead, mentor, and develop a globally diverse team of Site Reliability Engineers.
- Drive critical operations, incident response, on-call activities, and strategic reliability projects.
- Monitor and manage system health, availability, and performance.
- Develop automation for system provisioning, monitoring, and alerting.
- Improve observability, logging, troubleshooting, and performance optimization.
- Partner with development and product teams to embed reliability, security, and compliance into systems.
Requirements
- 3+ years of experience in system administration, DevOps, or SRE roles.
- 2+ years of experience managing people or teams and 2+ years leading projects.
- Proficiency in scripting and automation with Python, Bash, or Go.
- Experience with monitoring and logging tools such as Grafana, ELK, Splunk, or Prometheus.
- Knowledge of Docker, Kubernetes, CI/CD pipelines, and version control systems.
- Strong problem-solving, communication, and cross-functional collaboration skills.
Nice to have
- Hands-on experience with AWS, Azure, or GCP.
- Familiarity with Terraform, Ansible, or CloudFormation.
- Experience with incident response, SLAs, performance bottlenecks, or eventually consistent systems.
- Experience developing AI-based solutions.
- Knowledge of B2B payments, expense reimbursements, virtual cards, PCI-DSS, or SOX.
Culture & Benefits
- Agile, product-oriented engineering environment focused on innovation, AI, and automation.
- Health, dental, and vision insurance.
- Retirement savings plan, paid time off, life insurance, and disability insurance.
- Health savings account and flexible spending accounts.
- Tuition reimbursement and bonus eligibility for non-sales roles.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Senior SRE (Site Reliability Engineer) – Modernized Application Operations
145 000 - 170 000$
6 дней назад
Site Reliability Engineer
90 000 - 110 000$
10 дней назад
Principal Site Reliability Engineer (AI)
165 000 - 185 000$
6 дней назад
Software Development Engineer, SRE (US Federal)
137 000 - 205 400$
6 дней назад
Site Reliability Engineering (SRE) Manager (Azure)
139 700 - 232 900$
Replit
10 дней назад
Site Reliability Engineer
210 000 - 275 000$