Назад
Company hidden
1 час назад

System Software Engineer (AI)

100 000 - 180 000CAD
Формат работы
hybrid
Тип работы
fulltime
Английский
b2
Страна
US/Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
System Software Engineer (AI) (C++/Python, GPU Systems): Building communication and execution primitives that enable large-scale AI models to run efficiently across thousands of GPUs with an accent on systems programming, distributed training and inference, and software-hardware performance analysis. Focus on identifying bottlenecks, benchmarking multi-node clusters, and optimizing parallel AI workloads.

Location: Hybrid in Montreal, Canada

Salary: 100,000–180,000 CAD per year, plus equity, benefits, and RRSP

Company

hirify.global is a stealth-mode startup developing foundational infrastructure to improve performance, scalability, and resiliency in large-scale AI data center clusters.

What you will do

  • Design and implement communication and execution primitives connecting high-level AI frameworks with low-level system software.
  • Help shape foundational infrastructure for next-generation AI models.
  • Analyze and resolve performance bottlenecks in distributed training and inference workloads.
  • Investigate the software-hardware interface and conduct performance benchmarking on multi-node clusters.
  • Collaborate across the technology stack on early-stage product development.

Requirements

  • Strong proficiency in C++ and Python.
  • Deep understanding of systems programming fundamentals, including memory management, concurrency, and operating system internals.
  • Proficiency in a Linux development environment.
  • Bachelor’s or Master’s degree in Computer Engineering, Computer Science, or a related field.

Nice to have

  • GPU programming experience with CUDA and performance optimization for parallel architectures.
  • Experience with distributed AI frameworks such as PyTorch, JAX, or DeepSpeed.
  • Experience with inference engines such as vLLM, SGLang, Dynamo, or TRT-LLM.
  • Hands-on experience with large-scale cluster orchestration and telemetry tools.

Culture & Benefits

  • Work alongside experienced builders and operators in an early-stage product environment.
  • Collaboration is based on trust, integrity, respect, unity, and shared responsibility.
  • Benefits include meaningful equity, employee benefits, and RRSP contributions.
  • The company operates across remote and on-site locations in the United States and Canada.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →