Назад
Company hidden
19 часов назад

Sr. Software Development Engineer (HPC/ML Networking)

193 300 - 261 500$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Sr. Software Development Engineer (HPC/ML Networking): Developing collective operations and networking solutions to enable AI scaling across multiple accelerators and servers with an accent on low-level C/C++ and Linux kernel optimization. Focus on building high-performance interconnects for the largest AI clusters and models on AWS.

Location: Must be based in Cupertino, CA, USA

Salary: $193,300 - $261,500 annually

Company

Annapurna Labs, an integral part of AWS, develops critical hardware and software components that optimize EC2 infrastructure for AI/ML and HPC workloads.

What you will do

  • Develop collective operations that enable AI to scale across multiple accelerators and servers.
  • Build high-performance networking solutions for Machine Learning (ML) and High-Performance Computing (HPC) workloads on AWS.
  • Collaborate in a mixed-discipline environment with hardware engineers, RTL engineers, scientists, and architects.
  • Design and optimize low-level code using C/C++ and Linux kernels.
  • Mentor junior engineers while working under the guidance of senior and principal-level leadership.

Requirements

  • Location: Must be based in Cupertino, CA, USA
  • 5+ years of professional software development experience (non-internship).
  • Strong proficiency in C/C++ coding.
  • 5+ years of experience leading the design and architecture of scalable and reliable systems.
  • Comprehensive experience with the full software development life cycle, including code reviews and build processes.
  • Proven experience as a mentor or tech lead.

Nice to have

  • Bachelor's degree in Computer Science or equivalent.
  • Experience with ML Communications such as NCCL, NIXL, or NVSHMEM.
  • Knowledge of embedded systems.
  • Experience with high-speed networking or HPC interconnects.

Culture & Benefits

  • Flexible working hours and a core organizational tenet of respecting work-life balance.
  • Opportunity to work on the forefront of AI/ML with the largest clusters and models.
  • Comprehensive health insurance (medical, dental, vision, prescription).
  • 401(k) matching, paid time off, and parental leave.
  • Adoption and surrogacy reimbursement coverage.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →