$180,000 – $250,000
Listed on Modal’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
Modal is seeking a Forward Deployed ML Engineer to work directly with leading AI companies optimizing production workloads on their GPU infrastructure platform. The role combines hands-on technical work in inference and training optimization with customer-facing partnership and open-source contribution.
Our summary, not Modal’s wording. The full posting is on their site.
What they ask for
Required
- 2+ years professional ML engineering experience
- Hands-on experience with inference optimization, model training, GPU programming, or ML infrastructure
- Familiarity with serving and training toolchains
- Strong technical communicator
- Genuine interest in working directly with customers
- Willing to work in-person in New York City, San Francisco, or Stockholm
Nice to have
- Side projects, open-source contributions, or published work in ML or systems performance