# Member of Technical Staff - ML Performance

Hiring organization: [Modal](https://career.thegoodapps.co/organizations/modal)

Canonical page: https://career.thegoodapps.co/jobs/69d23637-88f3-4a89-a3ba-36eb9e15a7bb

Listed on Modal's own careers site. Applications go to them directly.

- Seniority: Senior
- Location: New York, NY
- Salary: 200000 – 350000 USD per year

## Summary

Modal is seeking an engineer to optimize machine learning inference and fine-tuning workloads on their GPU infrastructure platform. This role suits someone with deep experience in ML systems performance who enjoys debugging GPU bottlenecks and shipping optimizations at scale.

_Our summary, not Modal's wording._

## Skills named

CUDA, PyTorch

## Required

- 5+ years writing high-performance code
- Experience with PyTorch and ML inference frameworks
- Knowledge of Nvidia GPU architecture and CUDA
- Demonstrated ML performance engineering work (GPU occupancy tuning, algorithmic optimization, or overhead elimination)

## Nice to have

- Linux kernel and operating system internals knowledge
- Container and file system experience

Apply on Modal's site: https://jobs.ashbyhq.com/modal/af17da5e-23ca-4802-854d-5f0546e1ed32/application
