Listed on Krea AI’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role involves building and operating the distributed systems and infrastructure that power Krea's AI research and model training at scale, including GPU clusters, data pipelines, and custom orchestration systems. It suits engineers with strong mental models of how distributed systems work, who want to solve hard infrastructure problems in AI without necessarily requiring prior ML experience.
Our summary, not Krea AI’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Strong intuition for distributed systems
- Mental model of how systems interact and function under different conditions
Nice to have
- Python
- PyArrow
- DuckDB
- SQL
- massive relational databases
- PyTorch
- Pandas
- NumPy
- Kubernetes
- designing and implementing large-scale ETL systems
- fundamental knowledge of containerization, operating systems, file-systems, and networking
- distributed systems design
- distributed training systems (NCCL, InfiniBand, RDMA)
- streaming and event processing systems (Kafka, Pulsar, or similar)
- PyTorch internals, custom dataloaders, and training infrastructure