$150,000 – $350,000
Listed on CharacterAI’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role supports machine learning research and products by building and maintaining GPU infrastructure, cluster diagnostics tools, and experiment management systems. It suits engineers with deep experience in ML operations who want to optimize hardware utilization and solve large-scale training and serving challenges.
Our summary, not CharacterAI’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 4+ years supporting ML infrastructure
- Experience diagnosing ML infrastructure problems and failures
- Cloud platform experience (Compute Engine, Kubernetes, Cloud Storage)
- GPU experience
Nice to have
- Large GPU cluster experience
- High-performance computing and networking experience
- Large language model training experience
- GPU kernel development experience