2 дня назад
Cloud DevOps Engineer (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Cloud DevOps Engineer (AI): Building and operating production Kubernetes infrastructure, infrastructure-as-code environments, CI/CD pipelines, observability systems, and platforms for AI workloads with an accent on reliability, security, cost management, and scaling. Focus on provisioning and debugging clusters, automating deployments, investigating incidents, hardening cloud systems, and balancing inference cost and latency.
Location: Remote across Latin America, aligned to the client's working day.
Company
is a San Francisco-based software development company that builds and operates production AI systems and provides nearshore AI engineering teams.
What you will do
- Provision, upgrade, configure, and debug production Kubernetes clusters, including networking and workload configuration.
- Build reproducible infrastructure with Terraform or equivalent infrastructure-as-code tools and detect configuration drift.
- Develop and maintain CI/CD pipelines, container image workflows, and reliable production deployments.
- Operate monitoring and alerting, investigate incidents, identify root causes, and implement preventive changes.
- Support infrastructure for AI inference workloads, balancing scaling, latency, and operating costs.
- Harden Linux, Kubernetes, containers, and service meshes while supporting developers and QA with tooling and documentation.
Requirements
- 5+ years of experience as a DevOps, SRE, or systems engineer operating production infrastructure.
- Strong experience with Linux administration, networking, Git, Bash, and Python or Go.
- Production Kubernetes experience covering provisioning, upgrades, and troubleshooting.
- Hands-on experience with Terraform or equivalent, CI/CD platforms such as GitHub Actions or GitLab, and container images.
- Cloud deployment experience on AWS, Azure, or GCP, including managed Kubernetes services, monitoring, incident response, and infrastructure cost management.
- Clear written and spoken English at C1 level or above, with a bachelor's degree in Computer Science, a related field, or equivalent professional experience.
Nice to have
- Service mesh and traffic management experience with Istio, Linkerd, or equivalent tools.
- Helm, Kustomize, or equivalent environment templating experience.
- Production database operations with PostgreSQL, MongoDB, RDS, DynamoDB, or equivalent systems.
- Experience running or scaling inference workloads and evaluating hosted APIs against self-managed models.
- SOC 2 or HIPAA delivery experience, cloud certifications, open-source infrastructure contributions, or published technical writing.
Culture & Benefits
- 100% remote-first work anywhere in Latin America.
- Work aligned to client working hours and access to AI-assisted coding tools such as Claude Code, Codex, Cursor, and GitHub Copilot.
- Paid time off and U.S. holidays.
- AI training, certifications, and mentored career development.
- Profit sharing, U.S. dollar remuneration, and maternity coverage.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
3 дня назад
Senior DevOps Engineer (Cloud)
155 000 - 175 000$
8 дней назад
Senior DevOps Lead (Cloud)
3 дня назад
DevOps Engineer (AI)
90 000 - 130 000$
2 дня назад
Senior Cloud DevOps Engineer (AI)
6 дней назад
Senior DevOps / Platform Reliability Engineer (AI Platform)
4 дня назад
Senior DevSecOps Engineer (AWS)
120 000 - 150 000$