Назад
Company hidden
7 дней назад

ML Inference Performance Engineer (AI)

Формат работы
hybrid
Тип работы
fulltime
Грейд
junior
Английский
b2
Страна
France/UK/Netherlands +3 еще
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
ML Inference Performance Engineer (AI): Evaluating and benchmarking machine-learning inference systems across model export, compiler toolchains, runtimes, and accelerator hardware with an accent on reproducible measurement, pipeline consistency, and performance tooling. Focus on investigating hardware and software bottlenecks, analysing host-device overhead, and turning benchmark results into engineering and roadmap decisions.

Location: Hybrid work from the office in Eindhoven, Netherlands

Company

hirify.global develops the Metis AI Platform and next-generation AI acceleration technology.

What you will do

  • Develop and improve benchmarking tools for throughput, latency, power, and accuracy across device-level, host-transaction, and end-to-end scenarios.
  • Define reproducible evaluation procedures, standardize results, and maintain performance visualizations and dashboards.
  • Evaluate AI accelerator products, SDKs, toolchains, model support, flexibility, and limitations across vendors.
  • Analyse full computer vision inference pipelines, including host-device transaction overhead and end-to-end performance using equivalent configurations.
  • Set up and maintain lab hosts across multiple hardware platforms and support new evaluation hardware.
  • Prepare structured reports that inform engineering and product roadmap decisions.

Requirements

  • Currently enrolled in the final years of a Bachelor's programme or in a Master's programme in Computer Engineering, Electrical Engineering, Computer Science, or a related field; the role may also be completed as a Master's thesis project.
  • Python development experience and knowledge of C/C++.
  • Experience with end-to-end computer vision pipelines and benchmarking concepts such as performance and latency.
  • Experience with inference tools, APIs, or SDKs such as TensorRT, plus familiarity with quantization, ONNX, PyTorch, and other deep learning model concepts.
  • Experience with agentic AI, Git, LLM benchmarking, Linux, Bash scripting, Docker, and embedded hosts.
  • Proficient written and verbal English communication skills are required, together with clear documentation and good organizational skills.

Nice to have

  • GStreamer knowledge.
  • Basic GUI design experience.

Culture & Benefits

  • International organization with employees working remotely from multiple countries and offices across Europe.
  • Open culture supporting creativity and continuous innovation.
  • Collaborative ownership with freedom and responsibility.
  • Attractive compensation package with a pension plan and employee insurance.
  • Option to receive company shares.
  • Inclusive equal-opportunity environment.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →