Назад
Company hidden
4 часа назад

Staff Senior Reliability Engineer (Robotics)

125 000 - 171 600$
Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Staff Senior Reliability Engineer (Robotics): Leading end-to-end root cause investigations for complex production incidents across software, infrastructure, industrial controls, and production SOPs with an accent on customer communication, data analysis, and verified corrective action. Focus on identifying recurring failure patterns, driving continuous service improvement through automation and tooling, and influencing cross-functional teams without direct authority.

Location: USA remote with travel

Salary: $125,000–$171,600 per year, plus benefits.

Company

hirify.global develops AI-powered robotic and software platforms for warehouse automation and supply chain operations.

What you will do

  • Lead high-impact root cause analysis investigations for complex production incidents across software, infrastructure, industrial controls, and production procedures.
  • Chair structured, blameless RCA reviews aligned with ITIL Problem Management.
  • Act as the customer-facing technical lead for RCA discussions, updates, and formal deliverables within SLA timelines.
  • Analyze logs, telemetry, and incident trends to identify recurring failure patterns and recommend strategic improvements.
  • Present data-backed findings, risks, and recommendations to senior leadership and customer stakeholders.
  • Drive Continuous Service Improvement through automation, tooling, reporting, and team mentoring.

Requirements

  • At least 8 years of experience supporting complex, business-critical production environments with a focus on reliability.
  • At least 5 years leading technical RCA, post-incident reviews, or ITIL-aligned Problem Management.
  • Hands-on troubleshooting and data analysis across distributed systems, on-premises infrastructure, custom software, logs, telemetry, and incident datasets.
  • Regular customer interaction and the ability to build trust with executive, technical, and non-technical stakeholders.
  • Ability to manage multiple high-priority investigations and influence cross-functional teams without direct authority.
  • Bachelor’s degree in a technical field or equivalent practical experience.

Nice to have

  • Experience with warehouse automation, robotics, industrial controls, or large-scale physical production environments.
  • Experience with Kubernetes, VMware, Linux, Grafana, Prometheus, Zabbix, GitLab, and AI-assisted troubleshooting.
  • Experience with Power BI, Tableau, reporting automation, or dashboard development.
  • Deep knowledge of ITSM/ITIL practices, including Incident Management, Problem Management, RCA, and Continuous Service Improvement.

Culture & Benefits

  • Full-time employment with medical, dental, vision, and disability benefits.
  • 401(k), paid time off, and other benefits are included in the compensation package.
  • Collaborative workplace focused on innovation, inclusion, and diverse perspectives.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →