$245,000 – $295,000
Listed on Crusoe’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
Lead a team building core platform systems that transform large-scale compute infrastructure into reliable and efficient capacity for Crusoe's AI workloads. This role suits infrastructure engineers with 10+ years of experience and 3+ years in leadership who thrive on hands-on technical management and operating distributed systems at scale.
Our summary, not Crusoe’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 10+ years infrastructure or systems software development experience
- 3+ years engineering leadership experience
- Deep expertise building large-scale infrastructure platforms with resource pooling and allocation
- Strong background with Kubernetes and cloud platforms (GCP, AWS, or Azure)
- Experience with distributed state management and control systems
- Experience with efficiency, capacity, or performance engineering
- Player-coach management approach with hands-on technical credibility
- Track record hiring and growing infrastructure engineers
- Ability to operate in fast-moving, ambiguous environments
Nice to have
- Experience operating Kubernetes on bare-metal infrastructure and managed cloud services (GKE, EKS, AKS)
- Familiarity with operational challenges of GPU clusters, AI training, and inference workloads
- Knowledge of platform security concepts (secure boot, measured boot, TPMs, hardware attestation)
- Experience with capacity forecasting, demand modeling, or allocation optimization at scale
- Hands-on background with telemetry and observability platforms (Prometheus, OpenTelemetry, Grafana)
- Prior infrastructure platform building at hyperscalers or cloud providers with internal engineers as primary customers
- Familiarity with hardware-software co-design