1 день назад
Senior Network Engineer (AI Infrastructure)
150 000 - 190 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Network Engineer (AI Infrastructure): Building and scaling automated spine-leaf Ethernet networks for AI data centers and GPU clusters with an accent on Cumulus Linux, EVPN/VXLAN, BGP, and high-performance routing. Focus on optimizing east-west traffic, supporting RoCE/RDMA fabrics, designing systems for thousands of nodes, and developing network automation and observability.
Location: Fully remote within the continental United States, with occasional team and company offsites.
Annual base salary: $150,000–$190,000 USD, plus discretionary bonus, equity, and benefits.
Company
develops an end-to-end platform for developing, training, and deploying AI systems, combining developer-focused software with large-scale AI compute.
What you will do
- Design and deploy scalable spine-leaf architectures and high-performance Ethernet fabrics for AI data centers and GPU clusters.
- Build and operate EVPN/VXLAN, BGP, high-speed routing, backbone, DCI, WAN, and edge connectivity solutions.
- Optimize east-west traffic for AI training and inference, including RoCE/RDMA and low-latency transport.
- Develop network provisioning and operations automation using Infrastructure-as-Code.
- Troubleshoot distributed network, performance, and congestion issues while improving observability, telemetry, and operational visibility.
- Collaborate with compute, storage, AI platform, and operations teams to deliver integrated infrastructure solutions.
Requirements
- Must be based in the continental United States; occasional offsites are required.
- Visa sponsorship is not available for this role.
- Hands-on experience with Cumulus NOS, plus experience with SONiC and Junos.
- 5+ years of large-scale data center networking experience, including spine-leaf architectures, L3 fabrics, BGP, EVPN, and VXLAN.
- Experience operating HPC or GPU-dense environments and designing networks for hyperscalers, neoclouds, or high-scale SaaS infrastructure.
- Experience with Python, Ansible, Terraform, or similar automation tools, network observability, telemetry pipelines, and systems scaling to thousands of nodes.
Nice to have
- Familiarity with NVIDIA Spectrum, Quantum, BlueField, or related networking technologies.
- Experience with RDMA, RoCE, or InfiniBand fabrics and multi-region backbone design.
- Exposure to bare-metal provisioning systems or high-growth infrastructure startups.
Culture & Benefits
- Autonomous, high-ownership environment focused on urgency, open communication, continuous improvement, and long-term scalable systems.
- Medical, dental, and vision coverage, retirement support, and paid parental and family leave.
- Unlimited paid time off, company holidays, floating holidays, and a two-week winter company closure.
- Professional development allowance, wellness and work-from-home stipends, and a paid sabbatical after four years.
- Equity through RSUs and flexible schedules for office-based teams.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Senior Technical Consultant (NVIDIA Networking)
180 000 - 210 000$
Lambda
1 день назад
Site Reliability Engineer (AI Infrastructure)
240 000 - 356 000$
14 часов назад
Staff Network Engineer
100 000 - 230 000$
4 дня назад
Senior Network Engineer (Industrial AI Cloud)
1 850€
CrowdStrike
1 день назад
Network Engineer III - Network Deployment Engineer (Remote)
120 000 - 180 000$
4 дня назад