17 часов назад
Audio Software Engineer (Real-Time Audio)
170 000 - 400 000$
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Audio Software Engineer (Real-Time Audio): Building and shipping real-time audio systems for a full-duplex voice agent with an accent on WebRTC processing, browser audio pipelines, and production DSP. Focus on implementing echo cancellation, noise suppression, voice activity detection, low-latency streaming, and natural turn-taking across C++/Rust WebAssembly and TypeScript clients.
Location: San Jose, United States
Salary: $170,000–$400,000 annually
Company
Artificial intelligence company developing multimodal agentic systems and next-generation hardware that interact with people through speech, text, vision, and persistent memory.
What you will do
- Own audio quality in the live client, including echo, self-interruption, dropouts, and clipping.
- Build and tune browser audio pipelines with the Web Audio API, AudioWorklet, and MediaStream constraints.
- Develop the end-to-end WebRTC audio path covering echo cancellation, noise suppression, and voice activity detection.
- Ship production DSP as C++ or Rust compiled to WebAssembly and as TypeScript in the audio pipeline.
- Tune endpointing, interruption handling, turn-taking, latency, and streaming artifacts.
- Manage audio features from prototyping through production and collaborate with designers, platform engineers, and the speech team.
Requirements
- 5+ years of software engineering experience and experience shipping real-time audio to products used by real users.
- Hands-on experience with WebRTC, echo cancellation, noise suppression, and voice activity detection.
- Strong DSP fundamentals, including adaptive filtering, STFT, resampling, and gain control.
- Production experience with C/C++ or Rust DSP and browser delivery through WebAssembly.
- Working knowledge of the Web Audio API, AudioWorklet, MediaStream constraints, latency, buffering, and sample rates.
- Ability to own features end-to-end in a shared production codebase.
Nice to have
- Experience in voice, speech, or video-conferencing products.
- Audio ML or on-device inference experience, including RNNoise, DeepFilterNet, Silero VAD, ONNX Runtime, or Core ML.
- Familiarity with WebRTC internals, AEC3, Opus, LiveKit, or Pipecat.
- Experience with React, TypeScript, target-speaker isolation, diarization, barge-in, or conversational turn detection.
Culture & Benefits
- Full-time position with a US base salary range of $170,000–$400,000 annually.
- Total compensation may include additional components and benefits depending on the role.
- Work focuses on production systems operating in homes, cars, and noisy environments.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
2 дня назад
Senior Software Engineer, DVIP
137 100 - 188 300$
Razer
17 часов назад
Senior Audio Video (AV) Engineer (C++)
29 минут назад
Senior Compiler Engineer
13 часов назад
Knowledge Engineer (AI)
110 000 - 170 000$
9 часов назад
General Application (AI)
75 000 - 300 000$
3 дня назад
General Application (AI)
119 000 - 144 000$