13 дней назад
Manager TAG and Encompass SRE (Site Reliability Engineering)
106 000 - 130 600$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Manager TAG and Encompass SRE (Python/Kubernetes/Observability): Leading an SRE team while building reliable, observable, and efficient systems with an accent on monitoring, automation, incident response, and performance optimization. Focus on reducing operational toil, improving capacity management, developing provisioning and alerting tools, and embedding reliability practices across engineering and product teams.
Location: Portland, ME; Boston, MA; Dallas, TX; or Chicago, IL
Salary: $106,000–$130,600 annually
Company
operates systems supporting B2B payments, expense reimbursements, and virtual card creation.
What you will do
- Lead, mentor, and develop a globally diverse team of Site Reliability Engineers.
- Drive critical operations, incident response, strategic projects, and career development.
- Monitor and manage system health, availability, performance, and capacity.
- Develop automation for system provisioning, monitoring, and alerting.
- Improve observability and logging, and collaborate with development teams on reliability-focused features.
- Participate in on-call rotations and follow IT security and compliance requirements.
Requirements
- 3+ years of experience in system administration, DevOps, or SRE roles.
- 2+ years of experience managing people or teams and leading projects.
- Proficiency in scripting and automation with Python, Bash, or Go.
- Experience with monitoring and logging tools such as Grafana, ELK, Splunk, or Prometheus.
- Knowledge of Docker, Kubernetes, CI/CD pipelines, and version control systems.
- Strong problem-solving, communication, and cross-functional collaboration skills.
Nice to have
- Hands-on experience with AWS, Azure, or GCP.
- Familiarity with Terraform, Ansible, or CloudFormation.
- Experience with incident response, SLAs, performance troubleshooting, or AI-based solutions.
- Experience supporting B2B payments, expense reimbursements, or virtual card systems.
- Knowledge of eventual consistency, PCI-DSS, or SOX compliance requirements.
Culture & Benefits
- Agile, product-oriented engineering environment with a focus on innovation, AI, and automation.
- Health, dental, and vision insurance.
- Retirement savings plan, paid time off, HSA, FSA, life insurance, and disability insurance.
- Tuition reimbursement and eligibility for a quarterly or annual bonus.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 дней назад
Senior SRE (Site Reliability Engineer) – Modernized Application Operations
145 000 - 170 000$
13 дней назад
Staff Site Reliability Engineer (Production Engineer) - Federal
119 000 - 170 000$
14 дней назад
Software Development Engineer, SRE (US Federal)
137 000 - 205 400$
13 дней назад
Site Reliability Engineering (SRE) Manager (Azure)
139 700 - 232 900$
13 дней назад
Quality Engineering Manager (AI/SRE)
116 400 - 194 000$
4 дня назад