Назад
Company hidden
3 дня назад

Sr. SRE - Site Reliability Engineer

150 000 - 162 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Sr. SRE - Site Reliability Engineer (SRE/Cloud Operations): Ensuring the reliability, availability, and operational readiness of more than 300 mission-critical banking applications with an accent on distributed systems, observability, automation, and incident management. Focus on designing resilient platforms, leading service recovery and root-cause analysis, coordinating infrastructure modernization, and reducing operational risk in a regulated financial environment.

Location: On site in Austin, Texas, United States

Salary: $150,000–$162,000 per year, plus bonus or incentive opportunities

Company

hirify.global delivers technology solutions supporting investing, banking, financial planning, and highly regulated financial services.

What you will do

  • Advance platform reliability, resiliency, availability, and operational readiness across complex distributed systems.
  • Support more than 300 mission-critical banking applications and production environments.
  • Strengthen observability, monitoring, alert management, and automation to improve stability and reduce operational overhead.
  • Lead incident response, root-cause analysis, service recovery, disaster recovery planning, and infrastructure coordination.
  • Collaborate with software engineering, infrastructure, security, architecture, vendor, and business teams.
  • Participate in an on-call rotation approximately once every 5–6 weeks.

Requirements

  • 6+ years of experience supporting enterprise technology platforms, automation, monitoring, SDLC practices, and operational improvement in large-scale environments.
  • Experience with high-availability distributed systems, production operations, platform reliability, incident response, root-cause analysis, and service recovery.
  • Experience administering and troubleshooting Windows Server and Linux, including performance tuning and operational support.
  • Experience with cloud applications and infrastructure, IP networking, databases, messaging systems, and observability platforms.
  • Scripting or development experience with PowerShell, Python, Java, .NET, or Bash.
  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field. Must be authorized to work full-time in the United States without employer sponsorship.

Nice to have

  • 8+ years supporting mission-critical platforms in financial services or other highly regulated industries.
  • Experience scaling SRE practices, including SLOs, post-incident reviews, observability, and reliability metrics.
  • Experience with infrastructure modernization, vulnerability remediation, security operations, regulatory compliance, and disaster recovery.
  • Experience with Google Cloud Platform, Tanzu Application Service/Cloud Foundry, Actimize, or similar platforms.
  • Experience mentoring engineers and establishing operational standards.

Culture & Benefits

  • In-office collaboration with a regular in-person work model.
  • 401(k) with company match and employee stock purchase plan.
  • Paid vacation, volunteering time, and a 28-day sabbatical after five years for eligible positions.
  • Paid parental leave and family-building benefits.
  • Health, dental, and vision insurance, plus tuition reimbursement.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →