Назад
Company hidden
9 часов назад

Compiler Engineer (AI Infrastructure)

180 000 - 400 000$
Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Compiler Engineer (AI Infrastructure): Building compiler and runtime infrastructure that executes large-scale AI inference workloads across heterogeneous hardware with an accent on MLIR transformations, execution planning, scheduling, and memory movement. Focus on optimizing speculative decoding, kernel orchestration, serving infrastructure, and performance-critical systems for production AI workloads.

Location: On-site in San Francisco, CA, United States

Salary: $180,000–$400,000 per year, plus equity

Company

hirify.global is an early-stage AI infrastructure company building a multi-silicon neocloud for efficient inference across heterogeneous hardware.

What you will do

  • Build compiler and runtime infrastructure to improve latency, throughput, and efficiency for large-scale AI inference workloads.
  • Design execution strategies that partition and coordinate workloads across diverse hardware architectures.
  • Develop compiler optimizations across IR transformations, scheduling, memory movement, lowering, and kernel orchestration.
  • Enable new model architectures and serving techniques to run efficiently in production.
  • Work on execution planning, speculative decoding optimization, runtime optimization, and serving infrastructure.
  • Influence the architecture of an execution platform for future AI workload deployment.

Requirements

  • Strong systems and performance engineering fundamentals.
  • Experience building compiler systems, compiler-adjacent infrastructure, or execution and runtime systems.
  • Experience with IR transformations, compiler passes, lowering logic, or code generation systems.
  • Ability to reason about execution behavior, memory systems, scheduling, and hardware efficiency.
  • Strong software engineering skills in C++ and/or Python.
  • Bachelor’s degree in a relevant field, or equivalent education, training, and professional experience.

Nice to have

  • Experience with MLIR, LLVM, XLA, TVM, Triton, or similar compiler and runtime infrastructure.
  • Experience optimizing ML inference or serving workloads.
  • Familiarity with runtime systems, kernel dispatch, launch APIs, or memory allocators.
  • Experience with GPUs, AI accelerators, or heterogeneous hardware systems.
  • Experience profiling and debugging performance-critical systems.
  • Familiarity with scheduling, partitioning, or kernel-level optimizations.

Culture & Benefits

  • Early-stage environment with significant ownership over work.
  • Direct collaboration with a small group of highly capable colleagues.
  • Opportunity to shape the company’s systems, culture, standards, and execution platform.
  • Equity included in the compensation package.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →