$150,000 – $350,000
Listed on Modal’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
Modal is hiring a researcher to lead inference optimizations for their LLM serving platform, focusing on techniques like speculative decoding and quantization that reduce cost and latency. This role suits someone with a background shipping inference systems or research who can independently drive projects from conception through deployment.
Our summary, not Modal’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Research or systems background in LLM inference with published or shipped work
- Fluency across the LLM serving stack from kernels to schedulers
- Track record of shipping research or systems that others build upon
- Ability to independently drive research bets from idea through results
- In-person work in NYC or San Francisco office