3 дня назад
Senior Site Reliability Engineer
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (GCP/Kubernetes): Building infrastructure automation, observability, and DevSecOps capabilities for an AI-powered fintech platform with an accent on Google Cloud, Kubernetes, security, and service reliability. Focus on designing provisioning workflows, defining SLOs and error budgets, optimizing cloud resources, and responding to production incidents.
Location: Mexico; remote work currently available, with a planned transition toward hybrid work at hubs in Mexico City, Guadalajara, Bogotá, Buenos Aires, Rio de Janeiro, and São Paulo
Salary: Competitive salary; initial stock options grant and annual performance bonus
Company
is building an AI-powered financial ecosystem for Latin immigrants in the U.S., using WhatsApp, blockchain, and stablecoins to enable cross-border remittances.
What you will do
- Manage and optimize infrastructure on Google Cloud Platform and Google Kubernetes Engine.
- Automate infrastructure provisioning and configuration with Terraform, Helm, Go, Python, and Bash.
- Build and improve monitoring and alerting with OpenTelemetry, Prometheus, Grafana, and logging frameworks.
- Participate in on-call rotations, incident response, and post-mortem analyses.
- Define and track SLOs, SLIs, and error budgets to improve service health and performance.
- Implement cloud security practices, collaborate across Engineering, Security, and Product, and optimize GKE costs and resource usage.
Requirements
- 4+ years of experience as an SRE or Platform Engineer.
- Hands-on experience with GCP, GKE, Kubernetes, Docker, Linux, Terraform, Helm, and GitHub Actions.
- Programming or scripting proficiency in Go, Python, or Bash.
- Knowledge of Kubernetes architecture, deployments, networking, troubleshooting, TCP/IP, DNS, and load balancing.
- Experience with PostgreSQL or distributed databases, observability, incident management, on-call operations, and post-mortems.
- Understanding of FinOps, cloud security, SLOs, SLIs, and error budgets, with strong problem-solving and collaboration skills.
Nice to have
- GCP Professional DevOps Engineer or Cloud Architect certification.
- Certified Kubernetes Administrator certification.
- Experience with FinOps, cloud security, regulated industries, PagerDuty, or similar tools.
Culture & Benefits
- Remote work environment with planned hybrid collaboration at regional hubs.
- Health, dental, and vision plans.
- Unlimited paid time off and paid parental leave.
- Continuous learning opportunities and growth in an entrepreneurial environment.
- Stock options and annual performance bonus.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Sr. Site Reliability Engineer (Kubernetes/AWS)
5 дней назад
Site Reliability Engineer (AI)
Okta
6 дней назад
Staff Site Reliability Engineer (Splunk)
194 000 - 267 000$
Okta
7 дней назад
Staff Site Reliability Engineer (Splunk)
174 000 - 239 000$
4 дня назад
Staff Site Reliability Engineer (AI)
252 000 - 308 000$
5 дней назад