$165,000 – $220,000
Listed on Deepgram’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role sits between research and engineering to unlock value from conversational audio data at massive scale. You'll design active-learning systems, characterize what makes speech data challenging across languages and conditions, and run experiments that translate data insights into model improvements.
Our summary, not Deepgram’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Hands-on experience building data pipelines or solving model-facing data problems in ML or applied research
- Strong Python and data tooling with ability to write analysis and automation code yourself
- Experience with data selection, characterization, or active learning problems
- Familiarity reasoning about model output quality, confidence, and error modes from speech or NLP systems
- Track record of converting messy, ambiguous data into measurable model or product gains
- Building reusable systems and pipelines rather than one-off analyses
- Clear communication of complex technical findings to non-technical audiences
- Active, demonstrated use of AI tools in your own work
Nice to have
- Direct experience with automatic speech recognition, text-to-speech, or audio data
- Work with ensemble labeling, pseudo-labeling, or LLM-assisted annotation
- Familiarity with data provenance, PII handling, or GDPR-compliant pipelines
- Building custom or fine-tuned models for specific customers or domains
- Direct collaboration with research and engineering teams on shared infrastructure