Назад
Company hidden
обновлено 6 дней назад

SW Engineer - Developer Systems Reliability Engineering (AI-AIOps)

88 000 - 136 900$
Формат работы
hybrid
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
SW Engineer - Developer Systems Reliability Engineering (AI-AIOps) (Cloud Platform/SRE): Building and advancing reliability engineering practices for Visa’s Cloud Platform with an accent on observability, automation, and AI-AIOps. Focus on managing SLAs and SLOs, operationalizing monitoring across IaaS, PaaS, and container environments, and improving incident response through automation and operational intelligence.

Location: Austin, Texas, United States; hybrid work with at least 3 days in the office. The role includes 5–10% travel and shift-based, on-call, weekend, and 24/7/365 support coverage.

Salary: $88,000–$136,900 USD per year, with possible bonus, equity, and sales incentive payments where applicable.

Company

hirify.global is a global payments technology company facilitating transactions between consumers, merchants, financial institutions, and government entities.

What you will do

  • Maintain platform SLAs and SLOs by improving operational processes, standardization, and automation.
  • Own observability, monitoring, alerting, dashboards, telemetry, SLIs, and operational workflows across IaaS, PaaS, and container environments.
  • Deliver automation and AI-AIOps initiatives to reduce toil, improve reliability, and accelerate incident response.
  • Partner with development teams on release and service transition reviews, defining monitoring, runbooks, incident response, capacity, and production-readiness requirements.
  • Collaborate with infrastructure and operations teams to maintain and enhance the hirify.global Cloud Platform.
  • Analyze recurring technical issues, identify patterns, and support internal stakeholders across complex platform challenges.

Requirements

  • Bachelor’s degree or at least 3 years of relevant work experience.
  • At least 2 years of experience in a Platform, SRE, or Production Engineering group supporting high-availability platforms or applications.
  • Experience with CI/CD tools such as Jenkins, GitHub, Bitbucket, ArgoCD, Artifactory, Azure DevOps, or similar.
  • Experience with observability tools such as Grafana, Prometheus, Splunk, Datadog, New Relic, Dynatrace, or Sentry.
  • Experience supporting relational and non-relational databases, Linux or Windows systems, distributed computing, and container infrastructure.
  • Beginner-level programming or scripting in at least three relevant technologies, including Python, Java, Go, PowerShell, JavaScript, Terraform, Ansible, Helm, Chef, or CloudFormation.

Nice to have

  • Experience applying Claude, ChatGPT, GitHub Copilot, or similar AI-enabled tools to automation, observability, incident response, or operational workflows.
  • Experience with YAML, JSON, HTML, XML, virtualization, and container-first transformation.
  • Experience managing distributed container platforms, including deployment and release management, provisioning, capacity management, and workload management.

Culture & Benefits

  • Hands-on work supporting a large-scale cloud platform used by internal engineering stakeholders.
  • Medical, dental, and vision insurance.
  • 401(k), FSA/HSA, life insurance, paid time off, and wellness program.
  • Work hours vary according to department needs.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →