3 дня назад
Staff+ Software Engineer, Inference Velocity (AI)
405 000 - 485 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Staff+ Software Engineer, Inference Velocity (AI): Building developer productivity infrastructure, toolchains, development environments, and CI/CD systems for inference across GPU, TPU, and Trainium platforms with an accent on accelerator compatibility, workflow efficiency, and engineering velocity. Focus on designing multi-accelerator infrastructure, instrumenting productivity metrics, eliminating development bottlenecks, and aligning shared platform investments across organizations.
Location: Remote-friendly with at least 25% attendance required at an office in San Francisco, Seattle, or New York City, United States. Some roles may require more office time.
Annual salary: $405,000–$485,000 USD
Company
Anthropic builds reliable, interpretable, and steerable AI systems intended to be safe and beneficial for users and society.
What you will do
- Set the technical direction, architecture, and roadmap for Inference Developer Productivity.
- Own toolchain management across GPU/CUDA, TPU, and Trainium platforms, including compilers, drivers, libraries, and frameworks.
- Build development infrastructure such as devbox environments, validation automation, and shared tooling for heterogeneous accelerator hardware.
- Define productivity metrics, dashboards, and alerting to identify workflow regressions and toolchain failures.
- Identify and eliminate bottlenecks, toil, and friction across inference engineering workflows.
- Partner with central Infrastructure, guide build-versus-adopt decisions, and mentor engineers through technical collaboration and reviews.
Requirements
- 8+ years of software engineering experience, including significant experience as a technical lead or anchor on an infrastructure, platform, or developer productivity team.
- Deep experience in systems engineering, build and test infrastructure, or ML infrastructure, with hands-on knowledge of toolchains, CI/CD pipelines, and developer workflows.
- Experience owning toolchains or development environments for ML, HPC, inference, training, or large-scale distributed workloads.
- Strong expertise in at least one accelerator ecosystem: CUDA/GPU, TPU, or Trainium/AWS Neuron.
- Experience defining engineering metrics, dashboards, SLOs, or improvement initiatives that measurably increased engineering velocity.
- Strong written and verbal communication, with the ability to influence technical direction without formal authority.
Nice to have
- Experience with ML compiler toolchains such as XLA, Triton, or NeuronX.
- Experience with shared development environments, accelerator-scale CI/CD, Kubernetes-based development, or job scheduling.
- Prior technical leadership experience on a developer productivity or platform engineering team at a fast-growing AI/ML company.
Culture & Benefits
- Collaborative work on large-scale research and AI safety initiatives.
- Competitive compensation, optional equity donation matching, generous vacation, and parental leave.
- Flexible working hours and office collaboration.
- Visa sponsorship may be available, with immigration lawyer support, depending on the role and candidate.
- Bachelor’s degree or equivalent education, training, or professional experience is expected.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
9 дней назад
AI Inference Engineer
250 000 - 300 000$
Baseten
3 дня назад
Software Engineer (AI)
180 000 - 360 000$
Baseten
3 дня назад
Dedicated Inference Engineer (AI)
165 000 - 330 000$
9 дней назад
Staff AI Engineer
205 000 - 307 000$
5 дней назад
Senior Software Engineer (AI Infrastructure)
260 000 - 285 000$
Baseten
3 дня назад
Forward Deployed Engineers (AI)
200 000 - 400 000$