15 часов назад
Site Reliability Engineer, Tech Infra
129 960 - 246 240$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer, Tech Infra (Go/Python/Cloud): Building and operating large-scale, fault-tolerant distributed systems with an accent on automation, scalability, monitoring, and incident response. Focus on designing reliable infrastructure, diagnosing complex production issues, defining SLOs and SLIs, and improving system performance.
Location: Seattle, United States; fully in-person schedule up to 5 days a week
Salary: $129,960–$246,240 annually, with potential additional bonuses, incentives, and restricted stock units.
Company
operates data privacy, cybersecurity, trust and safety, and user data protection programs for TikTok’s U.S. platform and content ecosystem.
What you will do
- Develop and maintain automation procedures that improve system efficiency and reduce manual intervention.
- Design, deploy, and operate robust infrastructure in collaboration with software engineering teams.
- Build systems that scale with increasing web traffic and data volumes.
- Implement monitoring, metrics, and performance testing to identify bottlenecks and maintain system health.
- Participate in on-call rotations, incident management, diagnosis, resolution, and prevention of production issues.
- Define SLOs, SLIs, and SLAs with cross-functional teams and conduct blameless postmortems.
Requirements
- Bachelor’s degree in Computer Science, Information Technology, or a related field and 3+ years of experience.
- Professional experience as a Site Reliability Engineer, Systems Engineer, or similar software engineering professional.
- Programming experience with Go, Python, or other programming languages, with a focus on automation and operational excellence.
- Experience with network architecture, database modeling, cloud systems, and large-scale distributed systems.
- Strong knowledge of Linux operating systems, open-source technologies, debugging, and complex problem-solving.
- Strong communication and cross-functional collaboration skills.
Nice to have
- Knowledge of monitoring tools and methodologies such as Prometheus and Grafana.
- Experience with Docker, Kubernetes, or equivalent container and orchestration platforms.
Culture & Benefits
- Fully in-person work supports rapid alignment, decision-making, team development, and execution.
- Medical, dental, and vision insurance available from the first day.
- 401(k) savings plan with company match, paid parental leave, disability coverage, life insurance, and wellbeing benefits.
- 10 paid holidays, 10 paid sick days, and 17 days of paid personal time, with accruals increasing by tenure.
- Inclusive work environment focused on curiosity, collaboration, continuous learning, and problem-solving.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
19 часов назад
Site Reliability Engineer, Tech Infra - USDS (Cloud Infrastructure)
136 800 - 259 200$
9 часов назад
Site Reliability Engineer, Platform Responsibility - USDS (AI/LLM)
129 960 - 246 240$
7 часов назад
Site Reliability Engineer, Edge Services - USDS
136 800 - 359 720$
1 день назад
Staff Site Reliability Engineer (GCP/Kubernetes)
112 500 - 187 500$
Reddit
5 дней назад
Staff Software Engineer (Observability)
217 000 - 303 900$
6 дней назад
Site Reliability Engineer (AI Infrastructure)
200 000 - 240 000$