Salary not listed

Research Staff, Voice AI Foundations [UK - Remote]

AIRemote — London, UK
Published on 2026-10-10
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

As a member of the Research Staff at Deepgram, you will be involved in pioneering the development of Latent Space Models for voice AI. The role focuses on addressing fundamental challenges in audio data processing and aims to innovate new paradigms for audio AI. This position is ideal for researchers who want to tackle complex problems and have a strong foundation in statistical learning theory and deep learning.

About the company

Deepgram builds speech-to-text and voice AI APIs that convert audio into text and power voice applications for developers and enterprises.

Stack

Latent Space Modelsneural audio codecsgenerative modelsmultimodal systems

What you'll do

  • Develop neural audio codecs for low bit-rate compression
  • Pioneer steerable generative models for synthesizing diverse human speech
  • Create embedding systems to enable precise control over audio aspects
  • Generate synthetic audio data to scale voice interaction capabilities
  • Design model architectures for efficient training and real-time inference

What we're looking for

  • Strong foundation in statistical learning theory
  • Expertise in foundation model architectures
  • Ability to bridge theory and practice
  • Experience building data pipelines for large datasets
  • Track record of designing controlled experiments

Nice to have

  • Knowledge of optimizing models for real-world deployment
  • History of open-source contributions or research publications
View original job post