Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
TL;DR
Principal Network Engineer (AI/HPC): Leading the architecture and operation of high-performance network infrastructure for large-scale data centers with an accent on scalability, reliability, and GPU cluster networking. Focus on designing resilient routing and switching systems, driving network automation, and solving complex cross-layer networking challenges.
Location: AMER (Approximately 50% travel required to data centers and offices)
Salary: $220,000 - $280,000 USD
Company
Nscale provides high-performance cloud environments and infrastructure specifically optimized for AI and HPC workloads.
What you will do
- Lead the architecture, design, and deployment of large-scale network infrastructure across data center and cloud environments.
- Define technical strategy for core networking systems, owning decisions on routing, switching, transport, and overlay technologies like VXLAN and EVPN.
- Design highly available and resilient network systems to support mission-critical services at scale.
- Diagnose and resolve complex cross-layer networking issues spanning hardware, software, and infrastructure domains.
- Drive network automation, observability, and operational tooling initiatives to improve system reliability.
- Mentor senior and staff-level engineers to elevate technical standards and guide long-term infrastructure roadmaps.
Requirements
- 10+ years of experience designing and operating large-scale network infrastructure in cloud, hyperscale, ISP, or enterprise environments.
- Expert-level knowledge of networking protocols including BGP, OSPF, IS-IS, MPLS, TCP/IP, and IPv4/IPv6.
- Ability and willingness to travel approximately 50% to data centers and infrastructure sites.
- Willingness to participate in on-call rotations supporting production infrastructure.
- Bachelor’s degree in Computer Science, Engineering, or equivalent practical experience.
- Proven technical leadership across complex, cross-functional engineering initiatives.
Nice to have
- Experience with high-performance networking technologies such as InfiniBand and RoCE.
- Familiarity with GPU cluster networking and AI/HPC infrastructure environments.
- Proficiency in network automation and scripting using Python.
- Experience with network modeling frameworks such as YANG, OpenConfig, and NETCONF.
- Knowledge of cloud networking and hybrid infrastructure environments.
Culture & Benefits
- Competitive benefits package including medical, dental, and vision insurance.
- Flexible paid time off and parental leave.
- Retirement plan participation.
- Strategic leadership role with the ability to influence the future evolution of infrastructure platforms.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →