Назад
Company hidden
1 день назад

AI Infrastructure Operations Engineer (SiteOps)

Тип работы
fulltime
Грейд
junior
Английский
b2
Страна
US/Canada
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Infrastructure Operations Engineer (SiteOps) (AI infrastructure/hardware operations): Deploying, bringing up, monitoring, and troubleshooting Cerebras AI clusters in data center environments with an accent on hardware validation, telemetry, Linux systems, and networking fundamentals. Focus on executing power-on and readiness procedures, identifying incidents early, collecting diagnostic data, and escalating issues through established workflows.

Location: US and Canada offices

Company

Develops large-scale AI accelerator systems and infrastructure designed to deliver high-speed model training and inference.

What you will do

  • Assist with deployment and bring-up of CS-X systems, cluster servers, and networking hardware.
  • Execute power-on sequencing, readiness checks, and hardware validation tests.
  • Monitor hardware telemetry, alerts, and operational dashboards.
  • Perform first-line troubleshooting, collect logs and observations, and escalate incidents through established workflows.
  • Participate in incident response under senior engineer guidance and provide feedback on tooling and process gaps.

Requirements

  • Bachelor’s degree in a relevant engineering field or equivalent experience.
  • 0–3 years of experience in hardware operations, systems engineering, or data center environments.
  • Basic familiarity with server hardware, networking fundamentals, and Linux systems.
  • Ability to consistently execute defined hardware bring-up, validation, monitoring, and escalation procedures.

Nice to have

  • Internship or early-career experience in a data center or hardware laboratory environment.
  • Exposure to monitoring or telemetry systems.
  • Comfort working in data centers.

Culture & Benefits

  • Work on an AI platform designed to extend beyond GPU-based computing.
  • Opportunity to contribute to AI research and open-source research initiatives.
  • Exposure to one of the fastest AI supercomputers in the world.
  • Startup vitality combined with job stability.
  • Non-corporate culture focused on individual beliefs, learning, growth, and support.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →