Назад
Company hidden
2 дня назад

Site Reliability Engineer, USG (Aerospace)

Формат работы
remote (только USA)
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Site Reliability Engineer, USG (Aerospace) (Observability): Building and scaling a centralized observability platform for satellite networks, ground stations, and distributed aerospace systems with an accent on metrics, logging, tracing, and service reliability. Focus on defining SLOs and SLIs, automating Kubernetes and multi-cloud infrastructure, and leading monitoring, alerting, and incident response.

Location: Remote (United States)

Company

hirify.global develops laser communications technology and temporospatial software-defined networking platforms for aerospace, satellite, airborne, cislunar, and deep-space communications.

What you will do

  • Design and build a centralized observability platform covering metrics, logging, and distributed tracing.
  • Define and manage SLOs, SLIs, and error budgets for core products and high-availability distributed systems.
  • Partner with software engineers to establish observability practices, templates, documentation, and OpenTelemetry instrumentation.
  • Automate deployment, scaling, and management of the observability stack with Terraform and GitOps workflows such as ArgoCD.
  • Provide visibility into Kubernetes clusters and GCP and AWS environments.
  • Lead monitoring, alerting, incident response, and blameless post-mortems, including on-call support.

Requirements

  • Active Top Secret (TS/SCI) security clearance required.
  • 4+ years of experience in SRE or platform engineering, focused on observability for large-scale distributed compute or network systems.
  • Hands-on expertise with observability platforms including Prometheus, Grafana, Loki or ELK, OpenTelemetry, Tempo or Jaeger, and related tools.
  • Production experience with GCP and Kubernetes, plus Infrastructure as Code and GitOps principles.
  • Proficiency in a systems programming language, with Go and Python preferred for debugging and tooling.
  • Applicants must meet U.S. export-control eligibility requirements, including qualifying as a U.S. person or being eligible for the required export authorization.

Nice to have

  • Experience operating multi-cloud environments with GCP and AWS.
  • GitLab CI experience and knowledge of service mesh technologies such as Istio or Linkerd.
  • Experience instrumenting Go and C++ applications or tuning JVM observability for Java applications.
  • An active Secret clearance or higher.

Culture & Benefits

  • Work on aerospace communications and critical national security programs.
  • Professional development and advancement opportunities.
  • Collaborative, supportive, and inclusive workplace.
  • Flexible working arrangements, including hybrid remote and in-office schedules.
  • Competitive compensation with equity options, 401(k), health, dental, vision, and life insurance, plus paid time off.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →