обновлено 15 часов назад
Sr Site Reliability Engineer (Network Automation)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Sr Site Reliability Engineer (Network Automation): Design, build, and scale a global network platform across physical datacenters and multi-cloud environments with an accent on network automation, Kubernetes networking, and resilient infrastructure. Focus on troubleshooting complex network issues, automating operational workflows, building safe deployment pipelines, and improving reliability across large-scale distributed systems.
Location: Sydney, Australia
Company
is a global technology company operating an independent digital advertising platform across the open internet.
What you will do
- Design, build, and scale a global network platform across bare-metal datacenters and AWS, Azure, and Alibaba Cloud.
- Support thousands of hosts and develop reliable solutions for petabyte-scale data challenges.
- Troubleshoot complex network issues and maintain high availability and performance.
- Lead root cause analyses and postmortems, converting incidents into operational improvements.
- Build tools, automate workflows, and develop pipelines for safely testing and deploying production changes.
- Participate in a global follow-the-sun on-call rotation.
Requirements
- 6–8 years of hands-on network automation and production infrastructure operations experience.
- Strong software development and networking experience, including TCP/IP, the OSI model, BGP, and OSPF.
- Experience with Kubernetes networking technologies such as Cilium, Calico, and CNIs, plus Kubernetes and Docker troubleshooting.
- Experience with software load balancers such as NGINX Ingress, Envoy, or HAProxy.
- Knowledge of IPv6, SDN, QoS, bandwidth management, and network operating systems including SONiC, Cisco IOS, JunOS, Arista EOS, or Nokia SR Linux/SR OS.
- Proficiency with infrastructure as code, DevOps and SRE principles, Prometheus, Grafana, Python or Go, and production deployment workflows.
Nice to have
- Experience running Kubernetes clusters on bare metal.
- Background in platform engineering and large-scale distributed infrastructure.
- Experience integrating LLMs, MCP, or agentic workflows into engineering processes.
Culture & Benefits
- Software-first approach to network engineering and automation.
- Collaboration across network engineering, application, datacenter, and SRE teams.
- Emphasis on learning quickly, critical thinking, and creative problem solving.
- Inclusive hiring environment and accessibility accommodations.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
13 часов назад
Site Reliability Engineer Staff (Cloud Infrastructure)
15 часов назад
Site Reliability Engineer (Azure)
17 часов назад
Staff Site Reliability Engineer (AWS/Kubernetes)
16 часов назад
Site Reliability Engineer III (Platform)
14 часов назад
Senior Site Reliability Engineer (Data Platform)
16 часов назад