7 часов назад
Site Reliability Engineer, Global E-Commerce - USDS
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Site Reliability Engineer, Global E-Commerce - USDS (E-commerce and SRE): Building and operating large-scale, globally distributed, observable, fault-tolerant e-commerce systems with an accent on service reliability, scalable design, and cloud-native infrastructure. Focus on defining service objectives, automating infrastructure and resiliency, implementing observability across complex service meshes, and managing production incidents.
Location: Sydney, Australia; fully in-person schedule up to 5 days a week
Company
operates data privacy, cybersecurity, trust and safety, and U.S. user data protection programs for TikTok’s U.S. operations.
What you will do
- Own the service level of a critical, revenue-generating global e-commerce platform and its supporting infrastructure.
- Design for service reliability, high scalability, and release management in a cloud-native environment.
- Define service level indicators and data-driven objectives for uptime, latency, and system health.
- Collaborate with engineering and product teams on capacity planning, launch reviews, and transparent service delivery.
- Automate infrastructure, scalability, and service resiliency using infrastructure-as-code practices.
- Implement incident management and post-mortems while participating in on-call rotations.
Requirements
- Good understanding of Unix/Linux operating system internals and networking.
- Experience writing code in Java, Go, Python, or a similar language.
- Experience with algorithms, data structures, complexity analysis, and software design.
- Experience developing tools and APIs to reduce manual interaction with systems and applications.
- Systematic problem-solving skills, effective communication, and a strong sense of drive.
- Ability to work on-site in Sydney, Australia, up to 5 days per week.
Nice to have
- Experience running production-grade web services at scale in a cloud-native environment.
- Experience implementing monitoring, logging, and tracing in complex service meshes.
- Expertise in designing, analyzing, and troubleshooting large-scale distributed systems.
Culture & Benefits
- In-person collaboration focused on speed, alignment, team development, and integrated execution.
- Inclusive workplace that values diverse skills, experiences, and perspectives.
- Culture centered on curiosity, humility, resilience, continuous iteration, and meaningful impact.
- Regular on-call participation as part of production ownership.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
6 дней назад
Site Reliability Engineer (Cloud Banking)
2 дня назад
Principal AI Infrastructure Engineer, Kubernetes
6 дней назад
Senior DevSecOps Engineer (AWS)
3 дня назад
Trading Infrastructure Engineer
5 дней назад
Sr Site Reliability Engineer (Network Automation)
Canva
5 дней назад