Назад
Company hidden
6 дней назад

Audio AI Engineer (Robotics)

Формат работы
onsite
Тип работы
fulltime
Грейд
senior
Английский
c1
Страна
Germany
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Audio AI Engineer (Robotics): Building real-time audio pipelines, speech interfaces, and acoustic perception models for humanoid robots with an accent on conversational AI, audio machine learning, and embedded audio DSP. Focus on designing low-latency voice interaction systems, synchronizing audio with vision and proprioception, and deploying models on on-robot compute in noisy real-world environments.

Location: On-site in Metzingen / Riederich, Germany

Company

Develops humanoid robots with AI-based perception, interaction, and autonomous action capabilities.

What you will do

  • Build and own edge-to-cloud voice interaction pipelines covering automatic speech recognition, text-to-speech, wake-word detection, voice activity detection, and natural language understanding.
  • Design, train, integrate, and evaluate audio encoders and acoustic event recognition models for speech, ambient sounds, and contact sounds.
  • Architect real-time audio processing from microphone acquisition through denoising, beamforming, source separation, and model input.
  • Optimize audio systems and model inference for on-robot compute, low latency, noisy environments, and multi-speaker scenarios.
  • Drive data collection and evaluation infrastructure based on recordings from real deployments.
  • Help select and integrate microphone arrays, contact and tactile microphones, ADC frontends, calibration, and mounting requirements with AI and hardware teams.

Requirements

  • Master’s or PhD in Computer Science, Electrical Engineering, Computational Linguistics, or a related field.
  • At least 3 years of professional experience in audio-related AI engineering.
  • Demonstrated impact through publications, shipped systems, or both.
  • Depth in conversational AI, audio ML modeling, or embedded audio DSP, with breadth across the other areas.
  • Strong Python programming skills; C/C++ experience is beneficial for real-time and on-device work.
  • Excellent English skills required; German is optional.

Nice to have

  • Experience with ROS or robotics middleware.
  • Experience with agentic frameworks and LLM tool use.
  • Experience with audio simulation, room acoustics, or spatial audio.
  • Hands-on experience setting up recording equipment for ML data collection, including microphone selection, placement, and calibration.

Culture & Benefits

  • Collaborate with AI, hardware, and robotics teams.
  • Contribute to real-world robot deployments and explore new technical approaches.
  • Work across conversational AI, audio modeling, embedded DSP, and sensor integration.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →