Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Infrastructure Engineer (AI): Building resilient, scalable infrastructure for a high-traffic enterprise generative AI platform with an accent on Kubernetes, cloud platforms, infrastructure automation, and service reliability. Focus on leading incident response, designing fault-tolerant systems, defining SLOs and error budgets, and integrating AI-assisted infrastructure workflows.
Location: Hybrid, based out of the New York City, San Francisco, Seattle, or London hubs
Salary: $155.4K–$273.7K base compensation in San Francisco and New York City; $139.8K–$246.3K in other US locations, plus equity.
Company
Writer provides an end-to-end enterprise AI platform for building and deploying AI agents grounded in company data and powered by enterprise-grade large language models.
What you will do
- Build and operate resilient, performant, and highly available infrastructure for enterprise AI workflows.
- Design scalable, fault-tolerant systems across AWS, GCP, and Azure using Kubernetes, Helm, Terraform, and related cloud tooling.
- Automate infrastructure management and operational tasks with Python or Go, reducing manual on-call work and toil.
- Lead incident response, post-mortems, and root-cause analysis while applying findings to system architecture.
- Own SLOs, error budgets, on-call operations, reliability, performance, and efficiency for core services.
- Collaborate with product, security, and engineering teams on reliable system design from conception through launch.
Requirements
- 5+ years of experience in infrastructure engineering, DevOps, or a similar role operating large-scale, high-availability production systems.
- Production containerization experience with Kubernetes, plus Helm and Terraform or Pulumi on a major cloud platform; AWS experience is preferred.
- Strong proficiency in Python or Go for automation and tooling.
- Daily experience using AI-assisted or agentic development workflows is required.
- Experience with monitoring and logging stacks such as Prometheus, Grafana, and ELK or equivalent tools.
- Demonstrated ownership of at least one infrastructure build from zero to production, including measurable outcomes.
Nice to have
- Software engineering background with experience designing and shipping production services, libraries, or internal frameworks in Python, Go, or a comparable language.
Culture & Benefits
- Full-time employment with competitive compensation, equity, and a 401(k).
- Medical, dental, and vision coverage for employees and families.
- Generous paid time off, company holidays, and 16 weeks of paid parental leave.
- Fertility and family planning support, early-detection cancer testing, FSA options, and eligible HSA contributions.
- Annual wellness and learning-and-development stipends.
- Company-wide and team off-sites.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 часов назад
Member of Technical Staff, Infrastructure Engineer (AI)
175 000 - 240 000$
3 дня назад
Infrastructure Engineer (AI)
165 000 - 200 000$
5 дней назад
Platform Engineer (AI)
140 000 - 180 000$
4 дня назад
Site Reliability Engineer (AI Platform)
15 часов назад
Senior Software Engineer (Infrastructure)
170 000 - 220 000$
2 дня назад
Site Reliability Engineer (AWS)
180 000 - 220 000$