Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid) (Go/distributed systems): Building and operating foundational libraries, services, and tooling for a cloud-native cybersecurity platform with an accent on reliability engineering, scalability, observability, and multi-cloud architecture. Focus on re-architecting critical systems, eliminating classes of failure, designing resilience automation, and influencing architecture across product engineering teams.
Location: Hybrid role requiring candidates to be based in Midtown Manhattan, NY; Redmond, WA; Sunnyvale, CA; or Austin, TX, with 2–3 days per week in the office.
Salary: $140,000–$215,000 per year, plus potential bonuses, equity grants, health insurance, 401(k), and paid time off.
Company
CrowdStrike is a cybersecurity company providing an AI-native, cloud-based platform for threat detection and response.
What you will do
- Define and execute multi-year reliability roadmaps with engineering leaders across multiple product groups.
- Design and implement shared services, libraries, platforms, and architectural improvements for large-scale distributed systems.
- Build production backend services and automation using modern backend technologies, especially Go.
- Establish observability practices, SLOs, error budgets, performance optimization, and cloud efficiency improvements.
- Lead resilience engineering, chaos experiments, incident response, and follow-up improvements that eliminate recurring failure modes.
- Mentor engineers, influence architectural standards, and drive technical decisions across organizational boundaries.
Requirements
- 10+ years of experience building and operating distributed systems and service-oriented backends at scale.
- 5+ years developing microservices for SaaS products using Go, Java, Scala, Kotlin, Python, or Node.
- Expert proficiency in at least one programming language, with expert-level Go or the ability and willingness to reach that level.
- Deep knowledge of distributed systems, consensus, replication, consistency models, concurrency, scalability, and failure modes.
- Experience making organization-wide architectural decisions, scaling backend systems, and applying engineering best practices.
- Computer Science degree or equivalent experience, plus proven use of AI technologies to improve workflows and business outcomes.
Nice to have
- Experience with Kubernetes, AWS, GCP, OCI, Cassandra, Kafka, Elasticsearch, or OpenSearch.
- Experience operating multi-cloud services and building cloud-agnostic platforms or tooling.
- Background in infrastructure cost optimization, performance engineering, chaos engineering, or SLO/SLI frameworks.
- Experience building internal developer platforms and contributing to open source.
- Experience in cybersecurity or intelligence.
Culture & Benefits
- Hybrid work with flexibility and autonomy in a mission-driven cybersecurity environment.
- Comprehensive physical and mental wellness programs.
- Competitive vacation and holidays, plus paid parental and adoption leave.
- Professional development opportunities, employee networks, geographic groups, and volunteer programs.
- Office amenities and comprehensive U.S. benefits, including health insurance, 401(k), and paid time off.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →