$350,000 – $850,000
Listed on Menlo Ventures Portfolio’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role owns the complete pipeline of creating training data and reinforcement learning environments to improve visual reasoning in large language models, partnering across teams to translate research into real-world capabilities. You'll combine applied research with hands-on data work, managing vendor relationships, designing reward systems, and running experiments to measure how data strategy improvements enhance multimodal model performance.
Our summary, not Menlo Ventures Portfolio’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 7+ years of machine learning, computer vision, or software engineering experience
- Experience with reinforcement learning or reward design for vision-language models
- Familiarity with architecture and training of large vision-language models
- Ability to manage technical vendor relationships
- Results-oriented with flexibility and impact focus
Nice to have
- Experience designing evals or benchmarks for LLMs or vision-language models
- Background in large-scale pretraining and reinforcement learning
- Deep learning research on images, video, or other modalities
- Experience developing complex agentic systems with LLMs
- Large-scale ETL and data pipeline development