Назад

Не получаете ответ?

Telegram-вакансии старше 7 дней могут быть уже неактуальны.

7 дней назад

AI Platform Engineer

Формат работы
hybrid
Тип работы
fulltime
Английский
b2
Страна
CR
vacancy_detail.hirify_telegram_tooltipВакансия из Telegram канала -

Мэтч & Сопровод

Покажет вашу совместимость и напишет письмо

Описание вакансии

TL;DR
AI Platform Engineer (GPU/LLM Inference): Building and operating a dedicated multi-GPU inference platform across operating systems, storage, networking, observability, and model-serving with an accent on open-weight models, benchmarking, and platform reliability. Focus on deploying and tuning models, designing routing and quota systems, optimizing GPU workloads, and supporting engineers integrating internal models into applications and developer tooling.

AI Platform Engineer

Company

SatoshiLabs Group

Conditions

6 days ago

Skills

Candidate Availability

Required and preferred rules are kept separate and reflect the wording in the original posting.

About the Role

You will own a dedicated multi-GPU inference platform across its operating system, storage, networking, observability, model-serving stack, and user access. You will deploy and benchmark open-weight models, manage routing and quotas, document the platform, support integrations, and teach engineers how to use it effectively.

Requirements

  • Have coding and Infrastructure as Code skills
  • Know Ansible, Terraform, Docker, or Kubernetes
  • Have Linux systems administration experience covering networking, storage, containers, systemd, and troubleshooting
  • Understand the open-weight model ecosystem
  • Communicate in English
  • Demonstrate a service-oriented approach to supporting engineers and documenting systems

Responsibilities

  • Collaborate with IT to manage the GPU server lifecycle across operating systems, storage, and networking
  • Implement observability and alerting for GPU utilization, thermals, memory, latency, and throughput
  • Build and operate the model inference stack
  • Deploy, upgrade, and tune open-weight models
  • Evaluate and benchmark new models for quality, throughput, and latency
  • Recommend which models to run
  • Manage quotas, routing, and cost and usage reporting across teams
  • Optimize cloud model usage when needed
  • Support developers integrating internal models into IDE assistants, agents, CI pipelines, and applications
  • Maintain API keys, endpoints, and documentation
  • Write self-service onboarding guides
  • Run workshops and office hours
  • Design queueing and workload-prioritization approaches

Benefits

  • Option to be paid in bitcoin
  • Flexible working hours
  • Professional development budget for training, courses, and workshops
  • Regular company events and social gatherings
  • Office gym
  • Massages
  • Football table
  • Billiards
  • PlayStation
  • 3D printer
  • Free on-site parking
  • MultiSport card
  • Company mobile phone tariff
  • Yoga classes
  • Fitness classes

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →

Текст вакансии взят без изменений

Источник -