Skip to main content
CareerApp

Machine Learning Infrastructure Engineer

CharacterAI

Redwood City, CA · Remote · Senior

$150,000 – $350,000

Listed on CharacterAI’s own careers site. You apply with them directly — we never stand between you and the employer.

What this role is

This role supports machine learning research and products by building and maintaining GPU infrastructure, cluster diagnostics tools, and experiment management systems. It suits engineers with deep experience in ML operations who want to optimize hardware utilization and solve large-scale training and serving challenges.

Our summary, not CharacterAI’s wording. The full posting is on their site.

Skills this role names

Log in to see which of these are already on your profile.

What they ask for

Required

  • 4+ years supporting ML infrastructure
  • Experience diagnosing ML infrastructure problems and failures
  • Cloud platform experience (Compute Engine, Kubernetes, Cloud Storage)
  • GPU experience

Nice to have

  • Large GPU cluster experience
  • High-performance computing and networking experience
  • Large language model training experience
  • GPU kernel development experience

Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.