Назад
Company hidden
10 дней назад

Software Engineer (AI Infrastructure)

Формат работы
onsite
Тип работы
fulltime
Грейд
middle
Английский
b2
Страна
Ireland
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Software Engineer (AI Infrastructure): Building Conductor, a self-driving control plane for keeping tens of thousands of AI accelerators running efficiently, with an accent on distributed systems, energy-aware scheduling, and autonomous remediation. Focus on designing greenfield services, predicting hardware failures from noisy telemetry, and coordinating workload recovery across physical infrastructure.

Location: Dublin, Ireland; on-site

Company

hirify.global builds vertically integrated AI infrastructure spanning energy, data center construction, hardware, and cloud services.

What you will do

  • Design, build, and operate greenfield distributed services and control planes for AI infrastructure.
  • Develop closed-loop remediation systems that drain, checkpoint, replace, and resume live AI workloads automatically.
  • Build Conductor, a self-driving control plane that optimizes power, cost, and useful compute in real time.
  • Develop unified observability pipelines, energy-aware compute scheduling, straggler detection, and self-qualifying hardware pipelines.
  • Partner with staff and principal engineers to review architectures and establish safety rails for autonomous operations.
  • Collaborate with hardware engineering, energy management, data center construction, and customer support.

Requirements

  • At least 2 years of experience building, shipping, and operating backend or infrastructure systems in production.
  • Strong foundations in distributed systems, including state reconciliation, retries, idempotency, consistency tradeoffs, and autonomous automation.
  • Strong engineering fundamentals in Go, Rust, C++, or another modern systems language.
  • Ability to work through ambiguous problems, create clear designs, and deliver features to production.
  • Bachelor's degree or equivalent practical experience in Computer Science, Engineering, or a related technical field.
  • Ability to work on-site in Dublin, Ireland.

Nice to have

  • Experience with GPU health telemetry, NVLink, InfiniBand, RoCE, or hardware thermal and power behavior.
  • Experience designing observability and telemetry pipelines across compute, storage, and networking layers.
  • Familiarity with Kubernetes internals, Slurm, or other distributed orchestrators.
  • Experience applying statistical methods, anomaly detection, or time-series forecasting to operational data.
  • Knowledge of zero-trust architectures, policy-based access systems, or automated multi-tenant audit streaming.

Culture & Benefits

  • Work with experts across energy, manufacturing, data center construction, and cloud services.
  • Full-time employment with pension contributions.
  • Private health and dental insurance.
  • Income protection and life assurance.
  • Compensation may be paid as a salary or hourly rate, determined by experience, skills, education, internal equity, and market data.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →