Skip to main content
CareerApp

Skill

Reinforcement Learning

Data Science, Analytics and AI/ML

Reinforcement learning is a branch of machine learning in which an agent learns to make decisions by taking actions in an environment and receiving rewards or penalties, aiming to maximize cumulative reward over time. It underlies applications like game-playing AI (e.g., AlphaGo), robotics control, recommendation systems, and, more recently, fine-tuning large language models via techniques like RLHF. Researchers and engineers use frameworks such as OpenAI Gym, Stable Baselines, and RLlib to develop and test these algorithms.

See who is hiring

Reinforcement Learning in the job market

Last checked September 13, 2026

Open roles
86

on Career App right now

Employers
29

hiring for it

Median pay
$285,000

from 71 roles

Disclose pay
83%

of these roles

Open roles requiring Reinforcement Learning (86)

Research Staff, LLMs

Deepgram

San Francisco, CA · $150,000 – $250,000 · Remote

Deepgram seeks an experienced researcher to advance large language model architectures and training at scale, focusing on solving fundamental challenges in voice AI. This role suits researchers with deep expertise in transformers, distributed training, and LLM optimization who thrive in fast-moving environments and want to apply cutting-edge techniques to production systems.

Listed on Deepgram’s careers site · Apply there ↗

American International Group

Ph.D. Research Data Scientist, GenAI

American International Group

Atlanta, GA

This role leads research and development of generative AI and agentic systems within an insurance company, building production-ready AI solutions that automate complex business processes. It suits Ph.D.-level researchers with hands-on machine learning expertise and 2+ years of industry experience who want to translate cutting-edge AI research into real-world insurance applications.

Listed on American International Group’s careers site · Apply there ↗

Target

Sr Data Scientist - Supply Chain Optimization (Middle Mile)

Target

Full-time · $98,000 – $211,000

Target is seeking a senior data scientist to develop and deploy optimization and machine learning algorithms that improve supply chain decisions across inventory, transportation, and distribution. This role suits experienced practitioners with strong quantitative backgrounds who want to apply advanced analytics to large-scale operational problems.

Listed on Target’s careers site · Apply there ↗

Principal AI Researcher

Workday

Full-time · $228,000 – $342,000

This role leads Workday's newly formed AI Research team in advancing large language models and autonomous agents for enterprise applications. It suits seasoned AI researchers with publications or shipped products who want to define cutting-edge research directions while bridging theory and production-scale impact.

Listed on Workday’s careers site · Apply there ↗

Zillow

Principal Machine Learning Engineer, Agentic AI

Zillow

Full-time · Remote-USA · $194,200 – $326,600

This role leads the development of AI agent systems at Zillow, building multimodal technologies that power autonomous assistants for real estate. You'll prototype and deploy advanced reasoning models, mentor other engineers, and drive innovation in agentic AI while shipping to millions of users.

Listed on Zillow’s careers site · Apply there ↗

AI Research Scientist, Agentic Systems (Remote)

CrowdStrike

Full-time · USA - Remote · $120,000 – $180,000

This role focuses on building AI agents and training large language models to automate cybersecurity workflows, working with real-world threat data and security operations processes. It suits researchers with deep machine learning expertise who want to apply post-training techniques like reinforcement learning to autonomous systems in a domain with unique, large-scale data.

Listed on CrowdStrike’s careers site · Apply there ↗

Machine Learning Engineer, Next-Generation Recommendation Systems

Unity

$127,400 – $191,200

This role designs and deploys machine learning systems that power ad recommendations across billions of game players, focusing on next-generation techniques like large language models and reinforcement learning. It's suited to recent PhD graduates with strong research foundations who want to move cutting-edge AI ideas into production at massive scale.

Listed on Unity’s careers site · Apply there ↗

AI Robotics Researcher Intern (Dexterous Manipulation)

NIO

Internship · San Jose, CA · $38 – $46 / hour

NIO seeks an AI robotics research intern to develop frameworks for teaching robots dexterous manipulation by learning from human demonstrations and large-scale data, with a focus on translating human skills into real-world robotic behaviors through simulation and hardware deployment.

Listed on NIO’s careers site · Apply there ↗

Expedia Group

Senior Machine Learning Scientist - CRM Marketing

Expedia Group

Full-time · Seattle, WA · $173,000 – $242,500

This role leads machine learning systems that power personalized CRM campaigns at Expedia, determining which customers to target, when to reach them, and what offers to make. It suits experienced ML scientists who want to own production systems end-to-end, work across business and engineering teams, and influence retention and growth strategies for a global travel company.

Listed on Expedia Group’s careers site · Apply there ↗

Intuit

Manager 2, AI Science

Intuit

Mountain View, CA · $237,500 – $321,500

Lead a team of AI scientists building credit risk and fraud detection models for Intuit's consumer lending and banking products, protecting hundreds of thousands of customers while staying hands-on with complex modeling problems. This role suits experienced ML leaders who want to own high-impact risk infrastructure across multiple business verticals and grow both the team and themselves.

Listed on Intuit’s careers site · Apply there ↗

View all 86 on Jobs

Asked for alongside Reinforcement Learning

Measured from the 86 open roles that name Reinforcement Learning — not from a curated list.

What employers state they pay

Median

$285,000

Middle half

$236,000 – $361,000

From 71 of 86 open roles — the 83% that publish a range, annualized to USD. Only some employers state pay, and the ones that do skew to California and New York, where the law requires it.

Who employers are hiring

  • Entry11%
  • Mid8%
  • Senior43%
  • Lead / Staff / Principal33%
  • Management4%

From 72 roles whose posting states a level.

Where Reinforcement Learning roles are

Among the 71 of 86 open roles we could place on a map.

Employers hiring for Reinforcement Learning

29 in total, most open roles first.

Related skills

Curated neighbors in the taxonomy, whether or not employers ask for them together.

Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.