4 дня назад
Senior Site Reliability Engineer (AI/Cloud)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Senior Site Reliability Engineer (AI/Cloud): Building automated systems and observability solutions for reliable, scalable SaaS products with an accent on cloud infrastructure, AI-powered operations, and continuous delivery. Focus on optimizing performance, automating operational runbooks, managing Kubernetes environments, and leading 24/7 production incident triage and root-cause analysis.
Location: Remote in the United States
Company
delivers software and service solutions for financial reporting and capital markets transactions.
What you will do
- Build and maintain reliable, secure, scalable infrastructure for SaaS products.
- Implement SRE practices, monitoring, alerting, SLOs, and SLAs across applications and cloud environments.
- Apply AI for intelligent observability, incident prediction, automated remediation, and CloudOps operations.
- Automate operational runbooks, infrastructure management, and CI/CD pipelines.
- Optimize application performance and manage Kubernetes clusters, networking, firewalls, and load balancing.
- Participate in 24/7/365 on-call rotations and lead production incident triage and root-cause analysis.
Requirements
- 5+ years of experience designing, securing, monitoring, and maintaining cloud infrastructure in Azure or AWS.
- 5+ years of software development experience with a modern language such as C# .NET or Java.
- 5+ years of experience with automated deployments, Infrastructure as Code, and CI/CD tools such as Harness, Azure DevOps, Ansible, Jenkins, or Terraform.
- Experience with performance and availability monitoring using tools such as New Relic, Dynatrace, DataDog, or AppDynamics.
- Experience with PowerShell, Python, or Bash automation in Windows or Linux production environments.
- Experience with databases, Kubernetes, cloud networking, security, and public client-facing revenue-generating systems.
Nice to have
- Certifications or training in AI, Cloud AI services, or AIOps platforms.
- Experience with post-deployment verification testing and database performance tools.
- Bachelor’s degree in Computer Science or equivalent work experience.
Culture & Benefits
- Values-driven environment focused on collaboration, accountability, and continuous learning.
- Flexible workplace with comprehensive employee benefits.
- Opportunities for professional growth and exposure to emerging technologies.
- Competitive compensation and support for employee well-being.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
5 дней назад
Site Reliability Engineer, Tech Lead (AI)
5 дней назад
Senior DevOps / Site Reliability Engineer (SRE) (Cybersecurity)
165 000 - 215 000$
9 дней назад
Staff Site Reliability Engineer (AI)
252 000 - 308 000$
7 дней назад
Senior Site Reliability Engineer - Workflow Automation (Airflow)
10 дней назад
Staff Site Reliability Engineer (Kubernetes)
9 дней назад
Senior Site Reliability Engineer (Azure/AWS)
130 000 - 160 000$