# Senior Technical Program Manager (Engineering) - AI Tooling & Systems

Hiring organization: [Deepgram](https://career.thegoodapps.co/organizations/deepgram)

Canonical page: https://career.thegoodapps.co/jobs/73e12fee-bbee-409a-8b09-9b621af5a3ba

Listed on Deepgram's own careers site. Applications go to them directly.

- Seniority: Senior
- Location: USA | Remote
- Remote: yes
- Salary: 152000 – 208000 USD per year

## Summary

This role leads the design and delivery of ML infrastructure, model serving systems, and AI tooling that enable Deepgram's research and engineering teams to build and deploy voice AI models at scale. You'll coordinate across research, engineering, and product to translate ML requirements into production systems while optimizing for cost, latency, and developer velocity.

_Our summary, not Deepgram's wording._

## Skills named

AWS SageMaker, CUDA, Hugging Face, MLflow, PyTorch

## Required

- 5+ years of program management or technical leadership in ML infrastructure, ML platforms, or AI tooling
- Strong technical acumen in ML systems with hands-on experience as an ML engineer, systems engineer, or ML infrastructure engineer
- Experience coordinating cross-functional ML programs from training through evaluation, serving, and monitoring
- Ability to translate ML and research requirements into robust, scalable infrastructure
- Comfortable navigating complex technical tradeoffs around accuracy, latency, and cost
- Excellent communication with both technical and non-technical stakeholders
- Experience in high-growth or startup environments

## Nice to have

- Hands-on experience with model serving frameworks like vLLM, TensorRT, or TorchServe
- Experience optimizing LLM or speech/audio model inference through quantization, distillation, or KV-cache optimization
- Familiarity with ML experiment tracking and versioning tools
- Background with feature stores, vector databases, or real-time ML systems
- Knowledge of cost optimization for GPU and ML workloads
- Experience with multi-region model serving or edge deployment
- Hands-on experience with PyTorch, CUDA, Hugging Face, or cloud ML platforms

Apply on Deepgram's site: https://jobs.ashbyhq.com/Deepgram/1c34f6ba-6998-447c-9485-d4cf56db42de/application
