13 дней назад
Site Reliability Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (DevOps/Kubernetes): Building and optimizing reliable production applications and infrastructure automation for a mobile security platform with an accent on incident response, self-healing systems, and automated delivery. Focus on designing resilient services, troubleshooting priority incidents, improving service-level objectives, and supporting 24×7 operations.
Location: Dallas, Texas; hybrid work
Company
develops enterprise mobile threat defense solutions that protect iOS, Android, and Windows devices from cyberattacks and malware using real-time, machine learning-based security.
What you will do
- Design, code, test, and deliver software that automates manual operational work.
- Troubleshoot priority incidents, lead blameless post-mortems, and drive permanent incident resolution.
- Partner with development teams to improve application reliability, scalability, and service-level objectives.
- Design self-healing and resiliency patterns for production systems.
- Automate software upgrades, change management, and release management.
- Participate in 24×7 support coverage and mentor junior developers.
Requirements
- Expertise in at least one technology stack for designing, coding, testing, and delivering software.
- Strong knowledge of infrastructure components, including routers, load balancers, cloud products, containers, compute, storage, and networks.
- Experience with DevOps or application development teams and large-scale software development, preferably using Java, Python, or scripting languages.
- Hands-on experience with Kubernetes, Docker, Docker Swarm-style deployments, continuous delivery tools, Unix, Linux, and Solaris.
- Experience with monitoring, orchestration, configuration management, and infrastructure used in data warehousing or big data environments.
- Excellent debugging, troubleshooting, communication, prioritization, and problem-solving skills for complex and mission-critical systems.
Culture & Benefits
- Work in a curious, collaborative engineering environment with diverse perspectives.
- Take ownership of relevant reliability and infrastructure projects with mentorship and organizational support.
- Support a global-facing engineering organization and participate in 24×7 operational coverage as needed.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →