Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Sr. IT Site Reliability Software Engineer (Python/Terraform): Building resilient, automated cloud infrastructure and observability platforms for internal IT services with an accent on reliability engineering, infrastructure as code, CI/CD, and secure service defaults. Focus on designing scalable AWS/Azure infrastructure, creating logging and monitoring foundations, building AI automation tooling, and leading incident response and preventive engineering work.
Location: Costa Rica
Company
Databricks is a data and AI company providing a unified platform for data, analytics, AI applications, and agents.
What you will do
- Design and deploy production-grade cloud infrastructure using Terraform or Pulumi.
- Optimize system performance, architecture, scalability, availability, and latency for critical IT services.
- Build CI/CD pipelines with GitHub Actions and manage hosted and self-hosted runners.
- Create infrastructure with logging, metrics, tracing, alerts, and security enabled by default.
- Build internal AI plugins and automation scripts for developer workflows.
- Lead incident response, on-call troubleshooting, dashboards, and blameless post-mortems.
Requirements
- 5+ years of production-level software engineering experience.
- Strong Python proficiency is required.
- Expertise in Terraform or Pulumi, including modules and state management.
- Hands-on experience with AWS, Azure, or GCP, plus Kubernetes and Docker.
- Understanding of logging, metrics, tracing, and observability tools such as Datadog, Prometheus, or ELK.
- Experience with Kafka or messaging queues, GitHub Actions, and GitHub Runners; ability to execute ambiguous projects independently.
Culture & Benefits
- Product-led IT organization focused on scalable internal services and infrastructure.
- Cross-functional collaboration with Security, Engineering, and Support teams.
- Shared on-call rotation and blameless post-mortems.
- Comprehensive regional benefits and employee perks.
- Commitment to diversity, inclusion, and equal employment opportunity.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
8 дней назад
Sr. Site Reliability Engineer (Kubernetes/AWS)
10 дней назад
Senior Software Engineer, Platform Reliability (AWS/Kubernetes)
Wheely
3 дня назад
Site Reliability Engineer (AWS/Kubernetes)
5 000€
11 дней назад
Senior Site Reliability Engineer (FinTech)
11 дней назад
Sr. Manager, Site Reliability
Okta
10 дней назад
Senior TDI Site Reliability Engineer (AWS)
165 000 - 225 600$