Назад
Company hidden
2 дня назад

Senior / Staff Software Engineer (Observability / SRE) (AI)

148 000 - 249 000$
Формат работы
remote (только USA)/hybrid
Тип работы
fulltime
Грейд
senior/lead
Английский
b2
Страна
US/Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/

TL;DR

Senior / Staff Software Engineer (Observability / SRE) (AI): Design and lead the architecture and development of hirify.global’s monitoring and observability stack for cloud and on-prem environments with an accent on performance optimization and system reliability. Focus on analyzing end-to-end performance across hardware and Linux kernel using eBPF and profiling tools, and building CI/CD-based automation for telemetry and anomaly detection.

Location: Remote (US & Canada) or Hybrid in Toronto, ON; Pittsburgh, PA; or Dallas, TX

Salary: $148,000 - $249,000 USD

Company

hirify.global is a leader in Physical AI, developing technology for commercial autonomous trucks and robotaxis.

What you will do

  • Lead the architecture and development of the monitoring and observability stack for cloud and on-prem environments.
  • Develop benchmarks and integrate stress, chaos, and regression tests to validate hardware and platform choices.
  • Optimize end-to-end performance across hardware, firmware, and Linux kernel using eBPF and profiling tools.
  • Build automation and observability tooling using Go, Python, Java, and Kubernetes for CI/CD telemetry.
  • Collaborate with client teams to support application-specific observability requirements.
  • Mentor other engineers and communicate technical insights to stakeholders and leadership.

Requirements

  • 5+ years of software or systems engineering experience with end-to-end ownership of complex projects.
  • Proficiency in Python, Rust, or C/C++ and strong system design skills.
  • Deep knowledge of Linux internals (CPU scheduling, memory, I/O, networking) and perf tooling.
  • Experience with Kubernetes, microservices, and distributed systems.
  • Must be based in the US or Canada.
  • Proven ability to write design documents and lead cross-functional efforts.

Nice to have

  • Experience with OpenTelemetry and Grafana OSS.
  • Performance tuning for databases, streaming/batch platforms, or ML platforms (PyTorch).
  • Experience with GPU/xPU or Arm performance optimization.
  • Familiarity with Prometheus and distributed tracing.

Culture & Benefits

  • Competitive compensation and equity awards.
  • Comprehensive Medical, Dental, and Vision coverage for full-time employees.
  • Unlimited Vacation and flexible working hours.
  • Work from home support.
  • Daily catered meals, snacks, and drinks for office-based employees.
  • Regular team-building activities and social events.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →