Назад
2 часа назад

Senior Software Engineer, Infrastructure Engineering (Go)

182 000 - 242 000$
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Software Engineer, Infrastructure Engineering (Go): Building and operating reliable services for bare-metal infrastructure with an accent on incident response, observability, automation, and hardware lifecycle management. Focus on designing scalable operational tooling, creating CI/CD pipelines, and improving service resilience through root cause analysis, dashboards, alerts, and disaster recovery.

Location: New York, NY, United States

Salary: $182,000–$242,000 annually, plus discretionary bonus, equity awards, and benefits.

Company

CoreWeave provides cloud infrastructure, technology, and tools for AI labs, startups, and enterprises.

What you will do

  • Lead incident response, service recovery, root cause analysis, post-incident reviews, and long-term reliability improvements.
  • Develop and maintain incident response playbooks, communicate during incidents, and reduce recurring support queries and production incidents.
  • Improve core service performance, robustness, production supportability, and disaster recovery at scale.
  • Own observability and service health using Prometheus and Grafana; build dashboards, alerts, and automation for incident detection and recovery.
  • Build Go- or Redfish-based infrastructure services, CI/CD pipelines, and automation for the server hardware lifecycle.
  • Partner with Fleet Operations to create scalable self-service tooling and processes, and participate in the production on-call rotation.

Requirements

  • 7+ years of experience in cloud operations, site reliability engineering, or related technical roles.
  • Proficiency with Go or Python and prior experience with Prometheus and Grafana.
  • Understanding of cloud platforms such as Kubernetes, AWS, and GCP, including deployment of containerized applications with Kubernetes.
  • Familiarity with incident management practices, ITIL, and SRE best practices.
  • Production on-call experience, strong analytical and problem-solving skills, and excellent documentation abilities.
  • Eligibility to access U.S. export-controlled information is required, including qualifying U.S. person status or eligibility for applicable export authorization.

Culture & Benefits

  • Medical, dental, and vision insurance fully paid by CoreWeave, plus life and disability insurance.
  • Flexible Spending Account, Health Savings Account, 401(k) with employer match, ESPP participation, and tuition reimbursement.
  • Flexible PTO, paid parental leave, family-forming support, childcare support, and mental wellness benefits.
  • Catered lunch at office and data center locations.
  • Casual work environment focused on innovative disruption.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →