Назад
Company hidden
7 дней назад

Production Site Reliability Engineer (Java)

120 000 - 155 000$
Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Production Site Reliability Engineer (Java): Protecting the stability, performance, and resiliency of Schwab’s Order Management System with an accent on production incident response, distributed Java applications, database performance, and observability. Focus on restoring critical systems, automating operational improvements, reducing recurring incidents, and strengthening production readiness during night shifts and weekend support.

Location: Onsite in Austin, TX; Southlake, TX; or Omaha, NE. The role requires working onsite four days per week, primarily on night shifts from 2pm–10pm CST, with weekend and rotating on-call support as needed.

Salary: USD $120,000–$155,000 per year, plus bonus or incentive opportunities.

Company

hirify.global is a financial services company developing technology that supports clients’ financial futures.

What you will do

  • Protect the stability, performance, and resiliency of the Order Management System.
  • Assess and resolve complex production incidents and rapidly restore critical systems.
  • Collaborate with engineering, infrastructure, database, and vendor teams to minimize business impact.
  • Improve operational excellence through automation, observability, production readiness, and continuous improvement.
  • Lead incident response, identify root causes, and implement permanent corrective actions.
  • Shape operational practices, mentor engineers, and support medium-to-high-risk production releases.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field, or equivalent practical experience.
  • 5+ years of experience in production support, SRE, software operations, or a related technology support role.
  • Advanced troubleshooting experience with Java-based distributed applications and SQL-backed enterprise systems.
  • Strong knowledge of Oracle Database, SQL analysis, database troubleshooting, and performance optimization.
  • Experience with AppDynamics, Splunk, Grafana, InfluxDB, Control-M, Linux administration, and RHEL environments.
  • Availability for night shifts, weekends, and rotating on-call support; experience with Python or Shell scripting and high-severity incident management.

Nice to have

  • Experience improving availability, reliability, and mean time to resolution.
  • Experience creating monitoring strategies, operational runbooks, and escalation procedures.
  • Technical leadership, mentoring, or coaching experience within production support or SRE teams.
  • Knowledge of change management, risk management, security, and compliance in regulated environments.

Culture & Benefits

  • Collaborative environment focused on innovation, accountability, knowledge sharing, and continuous learning.
  • 401(k) with company match and employee stock purchase plan.
  • Health, dental, and vision insurance.
  • Paid vacation, volunteering time, parental leave, and family-building benefits.
  • Tuition reimbursement and a 28-day sabbatical after five years of service for eligible positions.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →