Salary not listed

Research Staff, Voice AI Foundations [Remote]

AIRemote — Australia- Melbourne
Published on 2026-10-10
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

As a Research Staff member at Deepgram, you will develop Latent Space Models to address challenges in voice AI related to data, scale, and cost. The role involves pioneering neural audio codecs and generative models for synthesizing human speech across diverse contexts.

About the company

Deepgram builds speech-to-text and voice AI APIs that convert audio into text and power voice applications for developers and enterprises.

What you'll do

  • Build next-generation neural audio codecs for high fidelity reconstruction.
  • Pioneer steerable generative models for synthesizing human speech.
  • Develop embedding systems for detailed control over audio factors.
  • Leverage latent recombination for generating synthetic audio data at scale.

What we're looking for

  • Strong mathematical foundation in statistical learning theory.
  • Expertise in foundation model architectures.
  • Ability to bridge theoretical and practical applications.
  • Experience in processing and curating large datasets.

Nice to have

  • Open-source contributions in speech/language AI.
  • Research publications advancing AI technologies.
View original job post