5 часов назад
Infrastructure and Reliability Engineer (Kubernetes/Terraform)
196 000 - 235 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Infrastructure and Reliability Engineer (Kubernetes/Terraform): Building and maintaining the infrastructure for a large-scale data replication platform with an accent on Kubernetes, cloud resources, CI/CD, observability, and reliability. Focus on designing AI-augmented release tooling, automating incident response and rollbacks, and setting infrastructure standards across AWS and GCP.
Location: San Francisco, CA; onsite 4 days per week
Salary: $196K–$235K base salary, plus equity
Company
builds an open data and action layer that gives AI agents authenticated access to business data across hundreds of sources.
What you will do
- Own infrastructure for the Data Replication platform, including Kubernetes clusters, CI/CD pipelines, secrets management, networking, and AWS and GCP resources.
- Partner with product engineers to integrate product features reliably with infrastructure.
- Improve observability, alerting, anomaly detection, and on-call operations.
- Build AI-augmented release and internal tooling, including canary deployments, progressive rollouts, automated release qualification, and rollback automation.
- Set infrastructure standards through self-service tooling, runbooks, and engineering coaching.
Requirements
- 7+ years of experience in infrastructure, platform engineering, SRE, or DevOps.
- Production ownership of Kubernetes, Helm, and Terraform.
- Deep experience with observability stacks such as Prometheus, Grafana, or Datadog.
- Experience owning CI/CD pipelines and developer tooling.
- Ability and willingness to read backend code, understand system failures, and instrument services correctly.
- Fluency with LLMs and agentic frameworks for automation, debugging, and reducing toil.
Nice to have
- Experience with data pipelines, replication systems, or ETL/ELT platforms.
- Experience with control plane/data plane architectures or internal developer platforms.
- Experience with , CDKs, or connector-based architectures.
Culture & Benefits
- Flexible PTO with an encouraged minimum of 25 days off annually.
- 16 weeks of fully paid parental leave for all parents.
- Medical, dental, and vision coverage for employees and dependents.
- 401(k) retirement plan, commuter benefits, and monthly internet reimbursement.
- Professional development budget, conference sponsorship, and book reimbursement.
- Collaborative in-person culture with breakfast and lunch at the San Francisco office.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
LangChain
4 часа назад
Database Infra Engineer (AI)
180 000 - 230 000$
5 часов назад
Infrastructure Engineer (AI)
200 000 - 400 000$
7 часов назад
Infrastructure Engineer (AI)
160 000 - 245 000$
8 часов назад
Infrastructure Engineer (AI)
150 000 - 300 000CAD
11 часов назад
DevOps Engineer (AI)
85 000 - 180 000$
5 дней назад
Infrastructure Engineer (AI)
200 000 - 350 000$