2 часа назад
Network Engineer
250 000 - 320 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Network Engineer (AI Infrastructure): Designing, deploying, and scaling datacenter networks for production-scale AI workloads, distributed systems, and high-performance compute with an accent on modern network fabrics, routing, and operational reliability. Focus on building scalable topology and automation workflows, troubleshooting hardware and distributed-system boundaries, and supporting next-generation AI datacenter infrastructure.
Location: San Francisco, CA; on-site
Salary: $250,000–$320,000 per year plus equity
Company
is building a multi-silicon neocloud that routes AI workloads across heterogeneous hardware for efficient, production-scale inference.
What you will do
- Design, deploy, and scale datacenter network infrastructure for AI workloads, distributed systems, and high-performance compute.
- Lead provisioning, device configuration, connectivity validation, deployment testing, and production turn-up for new infrastructure.
- Develop network topology designs, IPAM, deployment standards, documentation, and infrastructure readiness processes.
- Troubleshoot routing, hardware, connectivity, performance, and distributed-systems issues across physical and software boundaries.
- Partner with infrastructure, systems, deployment, and operations teams to improve reliability, deployment velocity, and scalability.
- Automate provisioning, configuration management, monitoring, deployment validation, and incident response workflows.
Requirements
- Experience designing, deploying, and operating production network infrastructure.
- Strong fundamentals in routing, switching, connectivity, performance, and reliability.
- Experience with spine-leaf or Clos fabrics, backbone or WAN networks, ECMP, BGP, EVPN, VXLAN, and routing policies.
- Knowledge of AI/HPC networking concepts including RoCEv2, InfiniBand, lossless Ethernet, QoS, DSCP, queuing, shaping, LAGs, optical transport, DWDM, coherent optics, and traffic engineering.
- Ability to troubleshoot complex issues across hardware, software, networking, and distributed systems.
- Ability to work independently, take end-to-end ownership, automate systems, and collaborate across engineering and operations teams.
Nice to have
- Experience with AI/HPC, GPU, or large-scale distributed infrastructure.
- Knowledge of AI application traffic patterns, collective operations, and workload colocation strategies.
- Experience with GCP, AWS, Azure, Arista, Cisco, Juniper, NVIDIA networking platforms, or Palo Alto Networks PAN-OS.
- Network automation experience with Python, Ansible, Terraform, or similar tools.
- Familiarity with RDMA, RoCE, InfiniBand, or other high-performance networking environments.
Culture & Benefits
- Join at an early stage and take significant ownership over network infrastructure and architecture.
- Work directly with a small group of highly capable engineers.
- Contribute to production systems serving frontier labs, hyperscalers, and AI-native companies.
- Receive equity in addition to the stated compensation.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →