обновлено 5 дней назад
Site Reliability Engineer (AWS/Kubernetes)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (AWS/Kubernetes): Designing and maintaining scalable, secure cloud infrastructure for an automotive repair management platform with an accent on observability, automation, high availability, and disaster recovery. Focus on building CI/CD and infrastructure pipelines, resolving critical reliability issues, and improving system performance across development, data, product, and QA environments.
Location: United States. Hybrid and remote work models are available based on proximity to office hubs; travel and attendance at several team and company offsites each year are expected.
Company
provides cloud-based software that helps auto repair shops manage operations, payments, customer engagement, and business growth.
What you will do
- Design and maintain reliable, scalable, and secure cloud infrastructure.
- Build monitoring, alerting, incident response, and system performance optimization practices.
- Automate deployment, testing, infrastructure management, and CI/CD pipelines.
- Implement backup, disaster recovery, failover, and high-availability solutions.
- Apply cloud security and compliance best practices.
- Collaborate with development, data, product, and QA teams while providing technical mentorship to junior DevOps engineers.
Requirements
- 3+ years of experience in DevOps, Site Reliability Engineering, or a related field.
- Hands-on experience with AWS or similar cloud providers and infrastructure as code.
- Experience with Docker, Kubernetes, automation tools, monitoring and observability tools such as Prometheus, Grafana, or ELK.
- Proficiency in Python, Bash, or similar scripting languages and experience designing CI/CD pipelines.
- Strong communication, collaboration, troubleshooting, and critical-incident response skills.
Nice to have
- Experience with Terraform, Prometheus, Grafana, or the ELK stack.
- Exposure to cloud security and compliance practices.
- Experience with Go, Java, or JavaScript.
Culture & Benefits
- Remote and hybrid work options with supported company and team offsites.
- Paid time off, maternity, parental bonding, and medical leave.
- Medical, dental, vision, prescription, mental health, life, and AD&D coverage.
- 401(k) plan with a 100% employer match on contributions up to 6%, plus FSA and HSA options.
- Fitness and wellness reimbursement, a home office setup bonus after one year, and continuing education support.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Senior Site Reliability Engineer (Cloud-Native Infrastructure)
142 800 - 178 500$
11 дней назад
Senior Site Reliability Engineer (Kubernetes)
147 600 - 221 400$
9 дней назад
Sr. Site Reliability Engineer (Kubernetes/AWS)
Okta
6 дней назад
Staff Site Reliability Engineer, Networking (AWS/FedRAMP)
174 000 - 238 000$
Okta
5 дней назад
Staff Site Reliability Engineer, Federal (TS/SCI)
174 000 - 238 000$
Okta
5 дней назад
Staff Site Reliability Engineer, Networking (AWS/Networking)
174 000 - 238 000$