$213,000 – $328,300
Listed on Deepgram’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role leads Deepgram's text-to-speech research program from strategy through production, owning the technical roadmap and building a team of researchers to advance speech-generation quality. It suits experienced speech-synthesis researchers comfortable with hands-on technical leadership, fast-moving environments, and shipping models to production at scale.
Our summary, not Deepgram’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Deep expertise in TTS, speech generation, or audio generative modeling with a track record of training large-scale neural models
- Command of modern speech-generation stack and open problems in naturalness, expressiveness, controllability, robustness, voice consistency, and inference cost
- History of setting research direction under uncertainty: prioritizing experiments, allocating compute and researcher time, and killing unsuccessful approaches
- Experience leading researchers and research engineers through technical leaders while remaining technically influential
- AI as default mode of work, not occasional tool, with a specific view of its limitations in speech research
- Ability to make complex technical tradeoffs legible to product, engineering, and executive audiences
Nice to have
- TTS or generative-audio models deployed at production scale
- Built or substantially scaled a high-performing AI research organization
- Sophisticated evaluation systems for generative speech
- Recognized external contributions in speech synthesis, neural audio codecs, speech language models, or multimodal models
- Experience in fast-moving startup or research environments