Назад
2 дня назад

Forward Deployed Engineer (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
Japan
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Forward Deployed Engineer (AI): Building and operating customer-facing AI systems across hardware and software integration, deployment, validation, automation, and production operations with an accent on accelerator architecture, Kubernetes, and LLM inference. Focus on debugging complete inference stacks, establishing reliable operating procedures, and driving enterprise deployments from requirements through production readiness.

Location: Hybrid, based out of Tokyo, Japan

Company

Tenstorrent develops AI computing platforms combining high-performance RISC-V CPUs, software, compilers, networking, and semiconductor technologies.

What you will do

  • Own technical outcomes for customer AI systems from integration and validation through production readiness and ongoing operations.
  • Work directly with customer engineering and operations teams to understand environments, constraints, and desired outcomes.
  • Integrate hardware and software, deploy systems, automate workflows, and debug issues across the inference stack.
  • Contribute production code and establish reliable acceptance criteria, validation plans, and operating procedures.
  • Drive customer systems through resolution and communicate technical trade-offs to customer leadership and core engineering teams.

Requirements

  • 5+ years of relevant experience in applied engineering, machine learning engineering, MLOps, platform engineering, infrastructure engineering, or site reliability engineering.
  • Strong software engineering skills and proficiency in Python; C++ experience is a plus.
  • Experience with Kubernetes and Linux administration at multi-node, HPC, or AI-cluster scale.
  • Experience with Helm-based deployments, observability, infrastructure automation, CI/CD, release engineering, and LLM inference frameworks such as vLLM, SGLang, or Mooncake.
  • Experience turning ambiguous requirements into production-ready implementations, acceptance criteria, validation plans, and operating procedures.
  • Native or business-level Japanese and business-level English are required.

Nice to have

  • C++ experience.

Culture & Benefits

  • Competitive compensation package and benefits.
  • Collaborative environment focused on curiosity and solving complex technical problems.
  • Opportunity to learn how AI hardware and software co-design translates silicon capabilities into latency and throughput gains.
  • Opportunity to operate disaggregated inference systems on Kubernetes while balancing performance, reliability, maintainability, and cost.
  • Equal opportunity employment.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →