Назад
Company hidden
17 часов назад

Audio Software Engineer (Real-Time Audio)

170 000 - 400 000$
Тип работы
fulltime
Грейд
senior
Английский
b2
Страна
US
Вакансия из списка Hirify.GlobalВакансия из Hirify Global, списка международных tech-компаний
Для мэтча и отклика нужен Plus

Мэтч & Сопровод

Для мэтча с этой вакансией нужен Plus

Описание вакансии

Текст:
/
TL;DR
Audio Software Engineer (Real-Time Audio): Building and shipping real-time audio systems for a full-duplex voice agent with an accent on WebRTC processing, browser audio pipelines, and production DSP. Focus on implementing echo cancellation, noise suppression, voice activity detection, low-latency streaming, and natural turn-taking across C++/Rust WebAssembly and TypeScript clients.

Location: San Jose, United States

Salary: $170,000–$400,000 annually

Company

Artificial intelligence company developing multimodal agentic systems and next-generation hardware that interact with people through speech, text, vision, and persistent memory.

What you will do

  • Own audio quality in the live client, including echo, self-interruption, dropouts, and clipping.
  • Build and tune browser audio pipelines with the Web Audio API, AudioWorklet, and MediaStream constraints.
  • Develop the end-to-end WebRTC audio path covering echo cancellation, noise suppression, and voice activity detection.
  • Ship production DSP as C++ or Rust compiled to WebAssembly and as TypeScript in the audio pipeline.
  • Tune endpointing, interruption handling, turn-taking, latency, and streaming artifacts.
  • Manage audio features from prototyping through production and collaborate with designers, platform engineers, and the speech team.

Requirements

  • 5+ years of software engineering experience and experience shipping real-time audio to products used by real users.
  • Hands-on experience with WebRTC, echo cancellation, noise suppression, and voice activity detection.
  • Strong DSP fundamentals, including adaptive filtering, STFT, resampling, and gain control.
  • Production experience with C/C++ or Rust DSP and browser delivery through WebAssembly.
  • Working knowledge of the Web Audio API, AudioWorklet, MediaStream constraints, latency, buffering, and sample rates.
  • Ability to own features end-to-end in a shared production codebase.

Nice to have

  • Experience in voice, speech, or video-conferencing products.
  • Audio ML or on-device inference experience, including RNNoise, DeepFilterNet, Silero VAD, ONNX Runtime, or Core ML.
  • Familiarity with WebRTC internals, AEC3, Opus, LiveKit, or Pipecat.
  • Experience with React, TypeScript, target-speaker isolation, diarization, barge-in, or conversational turn detection.

Culture & Benefits

  • Full-time position with a US base salary range of $170,000–$400,000 annually.
  • Total compensation may include additional components and benefits depending on the role.
  • Work focuses on production systems operating in homes, cars, and noisy environments.

Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →