# Engineer, Supercomputing & Distributed Systems

Hiring organization: [Krea AI](https://career.thegoodapps.co/organizations/krea-ai)

Canonical page: https://career.thegoodapps.co/jobs/20934905-203f-499b-84ff-92a6728999f9

Listed on Krea AI's own careers site. Applications go to them directly.

- Employment type: full time
- Location: San Francisco, CA

## Summary

This role involves building and operating the distributed systems and infrastructure that power Krea's AI research and model training at scale, including GPU clusters, data pipelines, and custom orchestration systems. It suits engineers with strong mental models of how distributed systems work, who want to solve hard infrastructure problems in AI without necessarily requiring prior ML experience.

_Our summary, not Krea AI's wording._

## Skills named

Apache Kafka, Containerization, DuckDB, Kubernetes, NumPy, Pandas, Python, PyTorch, SQL

## Required

- Strong intuition for distributed systems
- Mental model of how systems interact and function under different conditions

## Nice to have

- Python
- PyArrow
- DuckDB
- SQL
- massive relational databases
- PyTorch
- Pandas
- NumPy
- Kubernetes
- designing and implementing large-scale ETL systems
- fundamental knowledge of containerization, operating systems, file-systems, and networking
- distributed systems design
- distributed training systems (NCCL, InfiniBand, RDMA)
- streaming and event processing systems (Kafka, Pulsar, or similar)
- PyTorch internals, custom dataloaders, and training infrastructure

Apply on Krea AI's site: https://jobs.ashbyhq.com/krea/ebe94024-eef6-4306-a019-10072ad0f4c9/application
