$160,000 – $239,000
Listed on Deepgram’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role oversees the complete infrastructure architecture for a voice AI platform's production and research systems, managing GPU clusters and multi-cloud deployments at massive scale. It suits senior infrastructure engineers with deep experience in Kubernetes, storage systems, and GPU environments who want to make high-impact architectural decisions across real-time inference and large-scale ML training.
Our summary, not Deepgram’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 7+ years infrastructure engineering or systems architecture experience
- Multi-cloud architecture design spanning AWS and another major cloud provider or on-premises
- Deep storage system expertise including block, object, and file storage
- Kubernetes compute orchestration experience
- Hands-on GPU infrastructure experience
- Capacity planning and scaling experience for high-growth environments
- Ability to communicate architectural decisions to technical and non-technical audiences
- Networking fundamentals knowledge as it relates to infrastructure
Nice to have
- ML training workload infrastructure architecture experience
- Cost optimization and FinOps background
- Bare metal infrastructure operations experience
- Network architecture design expertise
- Infrastructure modeling and simulation for capacity planning
- Slurm or Ray job scheduling systems familiarity
- Power, cooling, and physical infrastructure knowledge for GPU deployments