обновлено 6 дней назад
SW Engineer - Developer Systems Reliability Engineering (AI-AIOps)
88 000 - 136 900$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
SW Engineer - Developer Systems Reliability Engineering (AI-AIOps) (Cloud Platform/SRE): Building and advancing reliability engineering practices for Visa’s Cloud Platform with an accent on observability, automation, and AI-AIOps. Focus on managing SLAs and SLOs, operationalizing monitoring across IaaS, PaaS, and container environments, and improving incident response through automation and operational intelligence.
Location: Austin, Texas, United States; hybrid work with at least 3 days in the office. The role includes 5–10% travel and shift-based, on-call, weekend, and 24/7/365 support coverage.
Salary: $88,000–$136,900 USD per year, with possible bonus, equity, and sales incentive payments where applicable.
Company
is a global payments technology company facilitating transactions between consumers, merchants, financial institutions, and government entities.
What you will do
- Maintain platform SLAs and SLOs by improving operational processes, standardization, and automation.
- Own observability, monitoring, alerting, dashboards, telemetry, SLIs, and operational workflows across IaaS, PaaS, and container environments.
- Deliver automation and AI-AIOps initiatives to reduce toil, improve reliability, and accelerate incident response.
- Partner with development teams on release and service transition reviews, defining monitoring, runbooks, incident response, capacity, and production-readiness requirements.
- Collaborate with infrastructure and operations teams to maintain and enhance the Cloud Platform.
- Analyze recurring technical issues, identify patterns, and support internal stakeholders across complex platform challenges.
Requirements
- Bachelor’s degree or at least 3 years of relevant work experience.
- At least 2 years of experience in a Platform, SRE, or Production Engineering group supporting high-availability platforms or applications.
- Experience with CI/CD tools such as Jenkins, GitHub, Bitbucket, ArgoCD, Artifactory, Azure DevOps, or similar.
- Experience with observability tools such as Grafana, Prometheus, Splunk, Datadog, New Relic, Dynatrace, or Sentry.
- Experience supporting relational and non-relational databases, Linux or Windows systems, distributed computing, and container infrastructure.
- Beginner-level programming or scripting in at least three relevant technologies, including Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, or CloudFormation.
Nice to have
- Experience applying Claude, ChatGPT, GitHub Copilot, or similar AI-enabled tools to automation, observability, incident response, or operational workflows.
- Experience with YAML, JSON, HTML, XML, virtualization, and container-first transformation.
- Experience managing distributed container platforms, including deployment and release management, provisioning, capacity management, and workload management.
Culture & Benefits
- Hands-on work supporting a large-scale cloud platform used by internal engineering stakeholders.
- Medical, dental, and vision insurance.
- 401(k), FSA/HSA, life insurance, paid time off, and wellness program.
- Work hours vary according to department needs.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
10 дней назад
Software Engineer, AI Infrastructure - USDS (AI)
122 574 - 259 200$
5 дней назад
Staff Software Engineer (Reliability Engineering & SDLC Governance)
165 000 - 260 500$
9 дней назад
Senior Systems Reliability Engineer (Java)
109 000 - 150 000$
5 дней назад
DevOps & SRE Engineer (Kubernetes)
100 000 - 150 000$
6 дней назад
Software Engineer (Cloud Infrastructure/SRE)
147 900 - 220 000$
7 дней назад
Software Engineer - Cloud Platform / Reliability (AWS/Kubernetes)
113 050 - 168 300$