Назад
обновлено 22 дня назад

AI Infrastructure Engineer (Sandbox Platform)

Формат работы
remote
Тип работы
fulltime
Грейд
senior
Английский
b2
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
AI Infrastructure Engineer (Sandbox Platform) (Secure Code Execution): Building a high-performance agent sandboxing platform, client library, and APIs for secure code execution across containerized and virtualized environments with an accent on isolation, reproducibility, and developer experience. Focus on optimizing cold-start latency and resource utilization, debugging production systems, and designing architecture across kernel and orchestration layers.

Company

Scale AI develops reliable AI systems, high-quality training data, and full-stack technologies for AI applications used by enterprises, governments, and leading technology companies.

What you will do

  • Design and build a secure sandboxing platform, client library, and API surface for code execution in containerized and virtualized environments.
  • Ensure strong isolation, security, and reproducibility across user sessions and workloads.
  • Optimize cold-start latency, memory footprint, and resource utilization at scale.
  • Debug production issues, monitor systems, respond to incidents, and implement preventive fixes through root cause analysis.
  • Partner with internal platform users to understand needs, debug issues, and develop supporting tooling.
  • Lead architecture reviews and own projects from design through deployment while contributing to the sandboxing product roadmap.

Requirements

  • 4+ years of experience building high-performance systems software and maintaining libraries, SDKs, or developer-facing APIs.
  • Deep understanding of Linux internals, including process isolation, memory management, cgroups, and namespaces.
  • Experience with containerization and virtualization technologies such as Docker, Firecracker, gVisor, QEMU, or Kata Containers.
  • Proficiency in a systems programming language such as Go, Rust, or C/C++.
  • Experience with API design, error propagation, documentation, and developer experience for well-crafted libraries.
  • Ability to work across infrastructure layers, from kernel modules to orchestration frameworks such as Kubernetes, and navigate production performance and security tradeoffs.

Nice to have

  • Founder or early-engineer experience at an infrastructure-focused startup with end-to-end product ownership.
  • Familiarity with LLM agents and agent frameworks such as OpenHands, Agent2Agent, or MCP.
  • Experience running secure workloads in multi-tenant or untrusted environments, including FaaS, CI sandboxes, or remote notebooks.
  • Exposure to snapshotting and restore technologies such as CRIU, VM snapshots, or overlayfs.
  • Open-source contributions to systems or developer-tools projects and production on-call experience.

Culture & Benefits

  • Benefits may include comprehensive health, dental, and vision coverage.
  • Retirement benefits, a learning and development stipend, and generous paid time off are available for eligible roles.
  • Eligible roles may include equity compensation and a commuter stipend.
  • Inclusive and equal-opportunity workplace with reasonable accommodation support for applicants with disabilities.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →