Назад
1 день назад

Staff + Sr. Software Engineer, Cloud Inference (AI)

320 000 - 485 000$
Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff + Sr. Software Engineer, Cloud Inference (AI): Building and operating backend services, infrastructure, and deployment systems that serve Claude across AWS, GCP, Azure, and other cloud platforms with an accent on large-scale distributed systems, multi-cloud abstractions, and cost-effective inference. Focus on designing capacity management, autoscaling, workload routing, and observability solutions while solving complex reliability, performance, and security challenges across heterogeneous cloud environments.

Location: San Francisco, CA or Seattle, WA. Hybrid policy: staff are expected to work from one of the offices at least 25% of the time, with some roles requiring more office attendance.

Annual salary: $320,000–$485,000 USD

Company

Anthropic builds reliable, interpretable, and steerable AI systems designed to be safe and beneficial for users and society.

What you will do

  • Design, build, and own backend services and infrastructure serving Claude across AWS, GCP, Azure, and future cloud service providers.
  • Work with inference, product API, systems, and security teams and cloud partners to launch serving stacks, resolve operational issues, and influence provider roadmaps.
  • Build CI/CD validation and deployment pipelines that release model versions to millions of users without regressions.
  • Develop cross-provider interfaces and tooling that reduce platform complexity and improve cost-effective inference management.
  • Contribute to capacity planning, autoscaling, workload routing, and accelerator and region selection.
  • Analyze observability data to identify performance bottlenecks, cost anomalies, and production regressions.

Requirements

  • Significant software engineering experience with high-performance, large-scale distributed systems serving millions of users.
  • Experience building or operating services on AWS, GCP, or Azure, with exposure to Kubernetes, infrastructure as code, or container orchestration.
  • Experience collaborating with internal teams and external partners to deliver impact.
  • Ability to learn new technologies, hardware platforms, and cloud provider ecosystems quickly.
  • Autonomous ownership of problems from design through operations.
  • Bachelor’s degree or equivalent education, training, or professional experience in a relevant field.

Nice to have

  • Experience scaling infrastructure or products across multiple cloud platforms.
  • Experience with capacity management, cost optimization, or resource planning across heterogeneous environments.
  • Understanding of multi-region deployments, geographic routing, and global traffic management.
  • Proficiency in Python or Rust.

Culture & Benefits

  • Collaborative work across research, engineering, policy, and business disciplines.
  • Flexible working hours, generous vacation and parental leave, and competitive benefits.
  • Optional equity donation matching.
  • Visa sponsorship may be available, with immigration lawyer support, depending on the role and candidate.
  • Office-based collaboration with frequent research discussions and emphasis on communication.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →