Назад
2 дня назад

Staff + Sr. Software Engineer, Scaling (AI)

320 000 - 485 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff + Sr. Software Engineer, Scaling (AI): Building and scaling distributed inference systems that serve Claude to millions of users, with an accent on intelligent routing, fleet-wide orchestration, and compute efficiency across accelerators and cloud providers. Focus on designing resilient autoscaling and deployment systems, optimizing high-performance inference, and integrating new AI hardware and model architectures.

Location: San Francisco, California or Seattle, Washington; hybrid policy requires working from an Anthropic office at least 25% of the time

Annual salary: $320,000–$485,000 USD

Company

Anthropic is a public benefit corporation building reliable, interpretable, and steerable AI systems.

What you will do

  • Design, build, and maintain distributed systems serving Claude to millions of users.
  • Develop intelligent request routing, load balancing, traffic management, and resilient real-time systems.
  • Optimize compute efficiency and fleet costs through autoscaling and workload orchestration across multiple cloud providers.
  • Build and operate production-grade deployment pipelines for releasing new models.
  • Provide high-performance inference infrastructure for next-generation model research.
  • Integrate new AI accelerator platforms and support new model architectures.

Requirements

  • Significant software engineering experience, particularly with distributed systems.
  • Bachelor’s degree or equivalent education, training, or professional experience in a relevant field.
  • Ability to work flexibly, take ownership beyond a defined job description, and focus on impact.
  • Interest in machine learning systems and infrastructure and their societal implications.
  • Ability to work from the San Francisco or Seattle office at least 25% of the time.

Nice to have

  • Experience with high-performance, large-scale distributed systems and machine learning systems at scale.
  • Experience with load balancing, request routing, or traffic management.
  • Familiarity with LLM inference optimization, batching, and caching strategies.
  • Experience with Kubernetes and AWS, GCP, or Azure.
  • Proficiency in Python or Rust.

Culture & Benefits

  • Collaborative research environment focused on trustworthy and steerable AI.
  • Flexible working hours and a collaborative office environment.
  • Generous vacation and parental leave.
  • Competitive compensation and optional equity donation matching.
  • Visa sponsorship is available, subject to role and candidate eligibility.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →