Назад
Company hidden
5 дней назад

Site Reliability Engineer, ITOF - ITOC

Формат работы
onsite
Тип работы
fulltime
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer, ITOF - ITOC (SRE/DevOps): Improving the reliability and operational consistency of IT products and solutions with an accent on proactive incident prevention, monitoring, automation, and production support. Focus on conducting SRE assessments, closing maturity gaps, resolving critical performance issues, and coaching Agile teams on reliability practices.

Location: Houston, Texas, United States of America; on-call support during U.S. business hours. Relocation will not be considered. Visa sponsorship is not available, and candidates with time-limited visa status will not be considered.

Company

hirify.global is hiring for an enterprise IT reliability engineering role supporting products and services in production environments.

What you will do

  • Drive proactive incident prevention through service performance analysis, process improvement, and operational gap identification.
  • Troubleshoot and resolve production and critical application, infrastructure, and system performance issues.
  • Provide on-call operational support during U.S. business hours and act as a technical resource during major incidents.
  • Conduct SRE technical assessments, identify maturity gaps, and develop continuous improvement plans.
  • Improve reliability through monitoring, logging, automation, and operational excellence initiatives.
  • Partner with the SRE Community of Practice to define standards, best practices, and capability development.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent experience.
  • Experience supporting IT products and services in production environments.
  • Experience troubleshooting and resolving complex application, infrastructure, or system performance issues.
  • Experience working in Agile and DevOps environments.
  • Knowledge of monitoring, logging, automation, and operational support practices.
  • Ability to participate in on-call support and major incident response activities.

Nice to have

  • SRE, Cloud, DevOps, Agile, or IT Service Management certifications.
  • Experience conducting reliability assessments and improving operational maturity.
  • Experience with automation scripting and operational workflow optimization.
  • Knowledge of SRE principles, observability, incident management, and performance engineering.
  • Experience coaching teams and influencing adoption of reliability engineering practices.

Culture & Benefits

  • Collaboration with cross-functional technology teams and communities of practice.
  • Participation in enterprise-wide reliability standards and capability development.
  • Equal opportunity employment and reasonable accommodations for qualified individuals with disabilities.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →