# Applied ML Engineer

Hiring organization: [Deepgram](https://career.thegoodapps.co/organizations/deepgram)

Canonical page: https://career.thegoodapps.co/jobs/58b6eb76-7754-437d-b9de-85e6ab3c41c9

Listed on Deepgram's own careers site. Applications go to them directly.

- Seniority: Senior
- Location: USA | Remote
- Remote: yes
- Salary: 150000 – 220000 USD per year

## Summary

This role bridges research and production at Deepgram's Voice AI platform, owning the pipeline that turns speech models from research notebooks into reliable, scaled services. It suits engineers who thrive at the intersection of ML systems and infrastructure, comfortable optimizing for both researcher productivity and production performance.

_Our summary, not Deepgram's wording._

## Skills named

Model Deployment, Python, PyTorch

## Required

- Strong Python proficiency and production-quality ML code writing
- Experience shipping ML models from prototype to production at scale
- Understanding of modern deep learning stack and large model training/evaluation/serving
- ML pipeline and tooling experience (training orchestration, evaluation, packaging, deployment, model CI/CD)
- Knowledge of serving optimization (latency, throughput, batching, resource efficiency)
- Comfort with distributed systems and GPU compute environments

## Nice to have

- Experience specifically with research-to-production handoff and systems
- Background in speech, audio, or real-time/streaming ML
- Experience building automated model evaluation and release-gating systems with regression detection
- Familiarity with hybrid on-premise GPU clusters and cloud infrastructure with workload orchestration
- Experience with inference optimization techniques (quantization, distillation, compilation, runtime tuning)
- Track record building internal platforms or developer-facing tooling that improved model shipping

Apply on Deepgram's site: https://jobs.ashbyhq.com/Deepgram/94ae2781-a85f-493a-86c1-ff85a9289355/application
