Назад
Company hidden
3 дня назад

Applied ML Engineer (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
middle/senior
Английский
c1
Страна
Japan
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Applied ML Engineer (AI): Building and deploying efficient foundation models for enterprise customers with an accent on inference optimization, model serving, and performance profiling. Focus on bridging the gap between research prototypes and production-ready systems in constrained environments.

Location: Must be based in Japan (Hybrid, Tokyo office attendance 1-2 times per week)

Company

hirify.global is an MIT CSAIL spin-out building general-purpose, efficient AI systems for deployment across diverse hardware, from data centers to edge devices.

What you will do

  • Own end-to-end applied ML projects for Japanese enterprise customers, from discovery to production deployment.
  • Integrate, profile, and optimize model inference to meet strict requirements for latency, throughput, and reliability.
  • Develop software infrastructure including data pipelines, evaluation systems, and serving components.
  • Perform fine-tuning and post-training using techniques like SFT and preference optimization.
  • Design task-specific evaluations and conduct systematic error analysis to iterate on model performance.
  • Collaborate with global research and engineering teams to influence the product roadmap based on field insights.

Requirements

  • Must be based in Japan and able to work from the Tokyo office 1-2 times per week.
  • Strong engineering skills with experience shipping production-quality ML systems.
  • Hands-on experience deploying modern language or multimodal models beyond proof-of-concept.
  • Proficiency in model serving, performance profiling, and inference optimization.
  • Professional proficiency in English for global technical collaboration.
  • Ability to lead technical discussions with customers and translate ambiguous needs into sound plans.

Nice to have

  • Working proficiency in Japanese.
  • Experience with LLM post-training methods.
  • Familiarity with deployment frameworks like vLLM, SGLang, llama.cpp, or ONNX Runtime.
  • Experience with quantization or hardware-aware optimization for edge platforms.

Culture & Benefits

  • High-impact work on foundation models in a rapidly growing market.
  • Collaborative, growth-driven environment with direct access to core research teams.
  • Flexible working hours with a primarily remote work policy.
  • Competitive salary and equity in a unicorn-stage company.
  • Unlimited paid time off and standard benefits for employees in Japan.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →