Назад
Company hidden
4 часа назад

Member of Technical Staff, Inference (AI)

Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Member of Technical Staff, Inference (AI): Building and optimizing production inference systems for large-scale genome and multimodal biological AI models with an accent on GPU kernels, distributed computing, and performance engineering. Focus on reducing latency for large MoE models, maximizing throughput across heterogeneous hardware, and deploying foundation models reliably across cloud, enterprise, and secure environments.

Location: San Francisco; authorization to work in the United States is required

Company

hirify.global is an AI research lab developing generative biological intelligence and foundation models for genomics, therapeutics, diagnostics, synthetic biology, and biodefense.

What you will do

  • Drive end-to-end performance improvements across model execution, memory management, networking, scheduling, and hardware utilization.
  • Develop and optimize GPU kernels, numerical operators, and inference-serving infrastructure.
  • Build scalable systems for serving, monitoring, benchmarking, and operating foundation models across cloud, enterprise, and secure environments.
  • Deploy biological AI models with pharmaceutical, biotech, research, and government partners.
  • Collaborate with research, infrastructure, and platform teams to bring new model architectures into production.

Requirements

  • Production experience optimizing, deploying, and operating large language models or other foundation models.
  • Deep understanding of GPU architectures and hands-on experience with CUDA, Triton, or equivalent technologies.
  • Ability to diagnose bottlenecks across model architectures, serving systems, networking, memory management, and distributed infrastructure.
  • Strong software engineering skills with proficiency in Python and PyTorch or comparable modern ML frameworks.
  • Experience working with customers, users, or cross-functional stakeholders to deliver production AI systems.
  • Must be authorized to work in the United States; hirify.global participates in E-Verify.

Nice to have

  • Experience with vLLM, TensorRT-LLM, SGLang, DeepSpeed, or similar inference frameworks.
  • Contributions to open-source AI infrastructure, inference frameworks, compilers, or kernel libraries.
  • Experience with distributed systems, cloud infrastructure, and large-scale GPU clusters.
  • Background in biological foundation models, computational biology, genomics, or scientific AI.
  • Experience operating AI systems in regulated, secure, or mission-critical environments.

Culture & Benefits

  • Work on large open biological AI models with applications in therapeutics, diagnostics, synthetic biology, and biodefense.
  • Collaborate with researchers across AI labs, biotechs, pharmaceutical companies, hospitals, government programs, and scientific institutions.
  • Join a multidisciplinary environment spanning distributed systems, model architecture, numerics, AI safety, and biology.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →