4 дня назад
Research Engineer - Audio & Speech Models (AI)
Мэтч & Сопровод
Для мэтча с этой вакансией нужен Plus
Описание вакансии
Текст:
TL;DR
Research Engineer - Audio & Speech Models (AI): Building open-source audio autoencoders, ASR, TTS, SSL, and speech-to-speech models with an accent on large-scale model training, audio data processing, and novel architecture design. Focus on optimizing multi-node GPU training, running rigorous ablations, and developing highly parallel data processing pipelines.
Location: San Francisco, United States; on-site
Company
Develops open-source AI models for audio and speech applications.
What you will do
- Build the next generation of open-source audio autoencoders, ASR, TTS, SSL, and speech-to-speech models.
- Run large-scale audio model training and optimize the training stack.
- Collect, process, and evaluate audio datasets.
- Design and improve model architectures and training methodologies through ablations.
- Develop research projects from conception through experimentation and write-up.
- Contribute to large-scale engineering implementation and existing codebases.
Requirements
- Expertise training audio-domain models, including text-to-speech, ASR, speech-to-speech, or speech-emotion-recognition models.
- Experience training audio autoencoders and working with audio signal processing.
- Experience with diffusion models, consistency models, or GANs.
- Experience training on large-scale, multi-node GPU clusters.
- Strong experimental methodology and understanding of highly parallel data processing pipelines.
- Proficiency with PyTorch and Python, plus published machine learning research and a postgraduate scientific degree.
Culture & Benefits
- Research and engineering excellence are valued equally, with a methodical approach to ambitious goals.
- Encouragement to pursue new ideas and move quickly from experimentation to impact.
- Medical, dental, vision, and FSA plans.
- Competitive compensation, 401(k), unlimited PTO, and company holidays.
- Relocation and immigration support may be available on a case-by-case basis.
- In-office snacks and meals in a collaborative, high-energy environment.
Будьте осторожны: если работодатель просит войти в их систему, используя iCloud/Google, прислать код/пароль, запустить код/ПО, не делайте этого - это мошенники. Обязательно жмите "Пожаловаться" или пишите в поддержку. Подробнее в гайде →
Похожие вакансии
4 дня назад
Machine Learning Engineer (AI)
200 000 - 300 000$
4 дня назад
Research Scientist (Audio AI)
210 000 - 360 000$
Deepgram
3 дня назад
Data Scientist (AI)
178 000 - 220 000$
Hippocratic AI
1 день назад
Research Scientist in Speech Technologies (AI)
4 дня назад
AI Research Scientist
4 дня назад
Senior Software Engineer (Edge AI)
160 000 - 190 000$