1 день назад
Site Reliability Engineer (Tech Infrastructure)
129 960 - 246 240$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer (Tech Infrastructure): Building and operating large-scale, fault-tolerant distributed systems with an accent on automation, scalability, monitoring, and production reliability. Focus on designing robust infrastructure, diagnosing complex incidents, defining SLOs, SLIs, and SLAs, and preventing recurring failures through performance testing and blameless postmortems.
Location: Seattle, United States; fully in-person schedule up to 5 days a week
Salary: $129,960–$246,240 annually, plus potential discretionary bonuses, incentives, and restricted stock units.
Company
operates TikTok applications and focuses on data privacy, cybersecurity, national security, and protection of U.S. user data.
What you will do
- Develop and maintain automation procedures that improve system efficiency and reduce manual intervention.
- Design, deploy, and operate robust infrastructure in collaboration with software engineering teams.
- Build for scalability across web traffic, data growth, and large-scale distributed systems.
- Implement monitoring tools, metrics, and performance tests to identify system health issues and bottlenecks.
- Participate in on-call rotations, incident management, diagnosis, resolution, and prevention of production issues.
- Define SLOs, SLIs, and SLAs with cross-functional teams and conduct blameless postmortems.
Requirements
- Bachelor’s degree in Computer Science, Information Technology, or a related field and 3+ years of experience.
- Professional experience as a Site Reliability Engineer, Systems Engineer, or similar software engineering professional.
- Proficiency in a high-level programming language such as Python, Go, Java, or Shell script.
- Experience with network architecture, database modeling, cloud systems, and large-scale distributed systems.
- Strong understanding of Linux operating systems and open-source technologies.
- Ability to work fully in person in Seattle, with on-site presence up to 5 days per week.
Nice to have
- Experience with Docker, Kubernetes, or equivalent containerization and orchestration platforms.
- Knowledge of monitoring tools and methodologies such as Prometheus and Grafana.
- Strong problem-solving, strategic thinking, debugging, communication, and cross-functional collaboration skills.
Culture & Benefits
- Medical, dental, and vision insurance from day one.
- 401(k) savings plan with company match.
- Paid parental leave, disability coverage, life insurance, and wellbeing benefits.
- 10 paid holidays, 10 paid sick days, and 17 days of paid personal time.
- Inclusive workplace focused on curiosity, openness, collaboration, self-direction, and continuous improvement.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
19 часов назад
Site Reliability Engineer, Tech Infra - USDS (Cloud Infrastructure)
136 800 - 259 200$
9 часов назад
Site Reliability Engineer, Platform Responsibility - USDS (AI/LLM)
129 960 - 246 240$
7 часов назад
Site Reliability Engineer, Edge Services - USDS
136 800 - 359 720$
1 день назад
Staff Site Reliability Engineer (GCP/Kubernetes)
112 500 - 187 500$
3 дня назад
Software Engineer, Infrastructure, AI Labs
145 000 - 195 000$
1 день назад
DevOps Engineer
100 000 - 125 000$