$150,000 – $250,000
Listed on Deepgram’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
You'll lead research into latent space models for voice AI, tackling fundamental challenges in audio compression, generative speech synthesis, and data efficiency. This role suits researchers who thrive on unsolved problems, move rapidly between theory and implementation, and want to pioneer entirely new approaches to making voice AI accessible at scale.
Our summary, not Deepgram’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Strong mathematical foundation in statistical learning theory, self-supervised and multimodal learning
- Deep expertise in foundation model architectures and multi-modal scaling
- Ability to bridge theory and practice, deriving and implementing novel mathematical formulations
- Experience building data pipelines for massive datasets with quality curation
- Track record of designing controlled experiments validating architectural innovations
- Experience optimizing models for real-world deployment with hardware efficiency
- History of open-source contributions or research publications in speech/language AI