Назад
Company hidden
10 часов назад

Distinguished Engineer (AI Infrastructure)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Distinguished Engineer (AI Infrastructure): Defining and implementing the multi-year technical vision for high-performance AI infrastructure across cloud services, on-premise edge appliances, and distributed control planes with an accent on observability, fleet orchestration, and autonomous infrastructure agents. Focus on building production prototypes and systems, making high-stakes architecture decisions, and delivering multi-tenant platforms with reliable deployment, telemetry, and rollback capabilities.

Location: US - Headquarters; on-site

Company

hirify.global builds high-performance infrastructure for demanding artificial intelligence workloads across silicon, systems, and networking.

What you will do

  • Set and implement the multi-year technical vision for the platform through competency analysis, working prototypes, production code, and cross-functional delivery.
  • Design distributed control planes connecting cloud services with on-premise edge appliances through mTLS gRPC streams.
  • Build large-scale observability and telemetry pipelines using OpenTelemetry, Kafka, time-series storage, real-time fabric views, packet-event analysis, and fleet health aggregation.
  • Develop autonomous infrastructure agents for onboarding, upgrades, drift remediation, failure recovery, and tool-based control-plane interactions.
  • Lead fleet orchestration, multi-tenant SaaS architecture, device lifecycle management, and intent compilation with verification and rollback.
  • Partner with Product, Customer Engineering, and engineering teams while reviewing critical code and communicating architectural decisions to technical and executive audiences.

Requirements

  • 18–20+ years of experience building and operating distributed systems for enterprise customers across cloud and on-premise environments.
  • Deep proficiency in Go or a comparable systems language, with strong distributed systems fundamentals.
  • Experience with observability platforms, microservices, gRPC, Protocol Buffers, OpenAPI, streaming systems, and multiple storage paradigms.
  • Track record building multi-tenant platforms with tenant isolation, RBAC, identity federation, and hybrid cloud/on-premise deployment models.
  • Experience with Kubernetes and Helm, strong API versioning and backward-compatibility practices, and high-stakes build-versus-integrate decisions.
  • Hands-on ability to build prototypes, write production code, review critical pull requests, and communicate architecture through writing and diagrams.

Nice to have

  • AI/ML agent architectures for infrastructure operations, including LangGraph, AutoGen, MCP tool use, and human-in-the-loop workflows.
  • Network automation, datacenter networking, OpenConfig, gNMI, or large-scale fabric management.
  • Hub-and-spoke, edge computing, or control-plane/data-plane separation architectures.
  • OpenTelemetry contributor experience, device enrollment systems, open-source contributions, or published distributed-systems work.

Culture & Benefits

  • Work in a talent-dense, high-performing engineering environment focused on ownership, technical rigor, and speed.
  • Build foundational infrastructure with immediate impact on demanding AI workloads.
  • Access an accessible hiring process with accommodations available upon request.
  • Work for a company backed by more than $300M in funding and experiencing rapid global adoption.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →