$250,000 – $300,000
Listed on Crusoe’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role automates deployment and testing for Crusoe's AI infrastructure, managing CI/CD pipelines and orchestration of large-scale GPU clusters across multiple datacenters. It suits deeply experienced infrastructure engineers who want to work on low-level systems automation at the cutting edge of AI compute.
Our summary, not Crusoe’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 12+ years of infrastructure or deployment automation experience
- Bachelor's or Master's degree in Computer Science, Electrical Engineering, or related field
- Experience building automated integration testing for AI Cloud environments
- Working knowledge of Kubernetes, Docker, Terraform, and Postgres
- Deep knowledge of CI/CD pipelines and Gitlab
- Experience with at least one configuration management system (Ansible, Puppet, Chef, or SaltStack)
- Advanced proficiency in Python and/or Bash
- Knowledge of Linux kernel internals including PCIe topology, VFIO, and memory management
- Familiarity with NVIDIA (CUDA/NCCL) and/or AMD (ROCm/RCCL) stacks in multi-node context
- Strong understanding of RDMA, RoCE, and InfiniBand protocols in virtualized systems
Nice to have
- Experience with MNNVL (Multi-Node NVLink) or specialized AI fabric architectures
- Familiarity with hardware-level debugging tools and performance profilers such as NVIDIA Nsight or AMD Omniperf
- Knowledge of containerized GPU orchestration with Kubernetes device plugins