Назад
Company hidden
обновлено 9 часов назад

Senior Systems Engineer (Workers AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Senior Systems Engineer (Workers AI) (Rust/distributed systems): Design and build AI inference infrastructure across a global network with an accent on heterogeneous accelerators, model scheduling, cold starts, KV cache management, and request routing. Focus on optimizing high-performance distributed systems, expanding observability, improving reliability, and leading complex platform projects through deployment.

Location: Austin, Texas, United States (hybrid/in-office). Applicants progressing to the offer stage may be asked to attend an in-person interview at a company office or hub.

Company

Large-scale Internet infrastructure and security company operating a global network that protects and accelerates websites and Internet applications.

What you will do

  • Design, build, and maintain core components of a serverless AI inference platform.
  • Optimize model scheduling, accelerator utilization, load balancing, caching, and request routing to reduce latency.
  • Improve platform reliability and resilience by identifying and mitigating systemic risks.
  • Expand metrics, logging, tracing, and alerting to detect and resolve production issues proactively.
  • Lead cross-functional technical projects from architecture and design through deployment and operationalization.
  • Mentor junior engineers and contribute to a collaborative engineering culture.

Requirements

  • Proven systems engineering experience focused on distributed, high-performance systems.
  • Expert proficiency in Rust, particularly in asynchronous environments.
  • Hands-on understanding of networking and application protocols, including TCP, HTTP, and WebSocket.
  • Experience with scaling and performance optimization in distributed environments, including load balancing and caching.
  • Ability to work in the Austin, Texas hybrid/in-office arrangement.

Nice to have

  • Experience with Kubernetes and/or Nomad.
  • Familiarity with large-scale inference serving for LLMs and diffusion models.

Culture & Benefits

  • Work on infrastructure for real-time voice, open LLMs, and customer-deployed models across a global network.
  • Collaborate with AI/ML engineers, hardware partners, and product teams.
  • Contribute to projects focused on protecting and improving the open Internet.
  • Employment may require authorization to receive technology controlled under U.S. export control laws without sponsorship for an export license.

Hiring process

  • Applicants who reach the offer stage may be asked to attend an in-person interview at a company office or hub.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →