Salary not listed
Research Staff, Voice AI Foundations [UK - Remote]
AIRemote — London, UK
Published on 2026-10-10
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.
About this role
As a member of the Research Staff at Deepgram, you will be involved in pioneering the development of Latent Space Models for voice AI. The role focuses on addressing fundamental challenges in audio data processing and aims to innovate new paradigms for audio AI. This position is ideal for researchers who want to tackle complex problems and have a strong foundation in statistical learning theory and deep learning.
About the company
Deepgram builds speech-to-text and voice AI APIs that convert audio into text and power voice applications for developers and enterprises.
Stack
Latent Space Modelsneural audio codecsgenerative modelsmultimodal systems
What you'll do
- Develop neural audio codecs for low bit-rate compression
- Pioneer steerable generative models for synthesizing diverse human speech
- Create embedding systems to enable precise control over audio aspects
- Generate synthetic audio data to scale voice interaction capabilities
- Design model architectures for efficient training and real-time inference
What we're looking for
- Strong foundation in statistical learning theory
- Expertise in foundation model architectures
- Ability to bridge theory and practice
- Experience building data pipelines for large datasets
- Track record of designing controlled experiments
Nice to have
- Knowledge of optimizing models for real-world deployment
- History of open-source contributions or research publications
