$350,000 – $850,000
Listed on Menlo Ventures Portfolio’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role involves keeping Anthropic's large-scale language model training pipelines running reliably and efficiently, requiring someone comfortable operating production systems during launches while also contributing research improvements to the training stack. It suits engineers who enjoy both debugging complex infrastructure problems and designing experiments to optimize model performance.
Our summary, not Menlo Ventures Portfolio’s wording. The full posting is on their site.
What they ask for
Required
- Hands-on experience training large language models, or deep expertise with JAX, TPU, PyTorch, or large-scale distributed systems
- Comfort being on-call for production systems and working extended hours during launches
- Ability to debug complex, ambiguous problems across multiple layers of the stack
- Strong communication and cross-timezone collaboration skills
- Bachelor's degree or equivalent education, training, and/or experience
Nice to have
- Previous experience training LLMs or extensive work with JAX/TPU, PyTorch, or other ML frameworks at scale
- Open-source LLM framework contributions (e.g., open_lm, llm-foundry, mesh-transformer-jax)
- Published research on model training, scaling laws, or ML systems
- Production ML systems or observability tools experience
- Background as a systems engineer, quant, or similar role requiring technical depth and operational excellence