Назад
3 дня назад

Staff+ Software Engineer, Inference Velocity (AI)

405 000 - 485 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff+ Software Engineer, Inference Velocity (AI): Building developer productivity infrastructure, toolchains, development environments, and CI/CD systems for inference across GPU, TPU, and Trainium platforms with an accent on accelerator compatibility, workflow efficiency, and engineering velocity. Focus on designing multi-accelerator infrastructure, instrumenting productivity metrics, eliminating development bottlenecks, and aligning shared platform investments across organizations.

Location: Remote-friendly with at least 25% attendance required at an office in San Francisco, Seattle, or New York City, United States. Some roles may require more office time.

Annual salary: $405,000–$485,000 USD

Company

Anthropic builds reliable, interpretable, and steerable AI systems intended to be safe and beneficial for users and society.

What you will do

  • Set the technical direction, architecture, and roadmap for Inference Developer Productivity.
  • Own toolchain management across GPU/CUDA, TPU, and Trainium platforms, including compilers, drivers, libraries, and frameworks.
  • Build development infrastructure such as devbox environments, validation automation, and shared tooling for heterogeneous accelerator hardware.
  • Define productivity metrics, dashboards, and alerting to identify workflow regressions and toolchain failures.
  • Identify and eliminate bottlenecks, toil, and friction across inference engineering workflows.
  • Partner with central Infrastructure, guide build-versus-adopt decisions, and mentor engineers through technical collaboration and reviews.

Requirements

  • 8+ years of software engineering experience, including significant experience as a technical lead or anchor on an infrastructure, platform, or developer productivity team.
  • Deep experience in systems engineering, build and test infrastructure, or ML infrastructure, with hands-on knowledge of toolchains, CI/CD pipelines, and developer workflows.
  • Experience owning toolchains or development environments for ML, HPC, inference, training, or large-scale distributed workloads.
  • Strong expertise in at least one accelerator ecosystem: CUDA/GPU, TPU, or Trainium/AWS Neuron.
  • Experience defining engineering metrics, dashboards, SLOs, or improvement initiatives that measurably increased engineering velocity.
  • Strong written and verbal communication, with the ability to influence technical direction without formal authority.

Nice to have

  • Experience with ML compiler toolchains such as XLA, Triton, or NeuronX.
  • Experience with shared development environments, accelerator-scale CI/CD, Kubernetes-based development, or job scheduling.
  • Prior technical leadership experience on a developer productivity or platform engineering team at a fast-growing AI/ML company.

Culture & Benefits

  • Collaborative work on large-scale research and AI safety initiatives.
  • Competitive compensation, optional equity donation matching, generous vacation, and parental leave.
  • Flexible working hours and office collaboration.
  • Visa sponsorship may be available, with immigration lawyer support, depending on the role and candidate.
  • Bachelor’s degree or equivalent education, training, or professional experience is expected.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →