Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.
27 open listings · Page 1 of 2
Roles posted by organizations here, alongside roles we found on employers’ own careers sites. Crawled roles say so on the card and send you to the employer to apply. Only show organizations on Career App
Anyscale
San Francisco, CA · $200,000 – $240,000 · Remote
Anyscale seeks a Senior Site Reliability Engineer to build and maintain the control and data plane infrastructure powering their Ray-based distributed AI platform. This role is ideal for someone with deep expertise in Kubernetes, cloud-native systems, and distributed computing who wants to tackle large-scale infrastructure challenges.
Listed on Anyscale’s careers site · Apply there ↗
XPeng
Full-time · Santa Clara, CA · $125,580 – $212,520
This role manages the technical program for XPeng's autonomous-driving training data infrastructure, overseeing data collection, processing, and delivery at scale across multiple engineering teams. It suits someone with substantial experience in cross-functional program management who understands large-scale data systems and can coordinate complex initiatives across distributed teams.
Listed on XPeng’s careers site · Apply there ↗
Quora
Remote - Multiple Locations, United States, Canada · $97,600 – $139,000
This role involves building and maintaining the ML serving infrastructure that powers Quora's production systems, working with distributed systems, GPU optimization, and Kubernetes. It's designed for recent graduates who want to learn large-scale ML infrastructure from experienced engineers while shipping production code quickly.
Listed on Quora’s careers site · Apply there ↗
Terran Orbital
Full-time · Irvine, CA · $70,000 – $110,000
This role creates visuals and 3D animations for a satellite manufacturer's marketing, proposals, and internal communications. It suits designers who excel at translating complex aerospace engineering into compelling multimedia and can work independently on fast-paced projects.
Listed on Terran Orbital’s careers site · Apply there ↗
Crusoe
Full-time · New York, NY · $155,000 – $200,000
This role involves helping enterprise customers successfully deploy AI and machine learning workloads on Crusoe's GPU infrastructure, moving from technical discovery through proof-of-concept to production launch. It's suited to cloud infrastructure engineers who want customer-facing technical responsibility and hands-on ownership of AI infrastructure challenges.
Listed on Crusoe’s careers site · Apply there ↗
Notion
Full-time · San Francisco, CA · $180,000 – $201,000 · Remote
Notion seeks a Software Engineer to build foundational AI platform systems that enable product teams to ship AI features safely and quickly at scale. This role suits engineers experienced with LLM or ML infrastructure who enjoy solving complex reliability, latency, and cost challenges across distributed systems.
Listed on Notion’s careers site · Apply there ↗
Motional
Full-time · Boston, MA · $200,000 – $275,000 · Remote
This Principal Engineer role oversees the development of an off-board AI evaluation framework that uses large language models and multimodal data to automatically assess autonomous vehicle safety, performance, and driving intelligence from historical logs. The position suits experienced software engineers or AI/ML specialists with deep knowledge of autonomous vehicles who want to bridge data science and real-world vehicle deployment at an established, well-funded company.
Listed on Motional’s careers site · Apply there ↗
Intuit
Atlanta, GA
An AI scientist position focused on developing and deploying machine learning models to power marketing and personalization features. This role suits someone with advanced training in a technical field who wants hands-on experience across the full ML lifecycle, from research through production, while collaborating closely with product and engineering teams.
Listed on Intuit’s careers site · Apply there ↗
Confido
Full-time · New York, NY · $210,000 – $300,000
Own the ML infrastructure layer at Confido, building the pipelines, serving systems, and cloud foundation that turn AI models into reliable, cost-efficient production systems at scale. This role suits someone who thrives at the intersection of software engineering and infrastructure, ready to give an AI/ML team a smooth path from research to production in a fast-growing startup.
Listed on Confido’s careers site · Apply there ↗
Pika Labs
Palo Alto, CA · $185,000 – $400,000
Pika is hiring a senior data engineer to design and operate large-scale data pipelines that feed multimodal AI model training, handling datasets across text, image, audio, and video. The role suits experienced engineers who have built production ML data infrastructure and want to own the full lifecycle of data quality and curation for cutting-edge generative models.
Listed on Pika Labs’s careers site · Apply there ↗
Deepgram
USA | Remote · $160,000 – $239,000
This role oversees the complete infrastructure architecture for a voice AI platform's production and research systems, managing GPU clusters and multi-cloud deployments at massive scale. It suits senior infrastructure engineers with deep experience in Kubernetes, storage systems, and GPU environments who want to make high-impact architectural decisions across real-time inference and large-scale ML training.
Listed on Deepgram’s careers site · Apply there ↗
Full-time · Remote - United States · $230,000 – $322,000
Reddit is seeking a Staff Machine Learning Engineer to lead the development of large-scale ML systems powering recommendations, search, and content understanding across their consumer platform. This role suits experienced ML practitioners who enjoy building complex systems from research through production and want to influence product direction while mentoring other engineers.
Listed on Reddit’s careers site · Apply there ↗
Eight Sleep
Full-time · San Francisco, CA
This role involves building machine learning systems that personalize sleep experiences through predictive models, foundation model applications, and behavior analysis. It suits engineers who enjoy owning projects end-to-end from research through production deployment and want to work on consumer health technology.
Listed on Eight Sleep’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
Cohere seeks a senior engineer to build and maintain the training framework powering their large-scale language models, working across distributed systems, HPC infrastructure, and tooling. This role suits someone with deep expertise in distributed training systems who wants ownership over critical ML infrastructure components.
Listed on Cohere’s careers site · Apply there ↗
Udio
Full-time · New York, NY · $180,000 – $220,000
This role involves building the data infrastructure that powers a generative audio company's research, focusing on ingesting and unifying large datasets from multiple external sources. You'll design systems for entity resolution, deduplication, and data enrichment at scale, working closely with ML researchers to prepare training-ready datasets.
Listed on Udio’s careers site · Apply there ↗
Runway
Remote · $270,000 – $370,000
This role owns the data strategy for AI world models, designing datasets and running experiments to understand how training data shapes model capabilities across creative and robotics applications. It suits experienced machine learning engineers who want to bridge data science and research at a company building large-scale generative systems.
Listed on Runway’s careers site · Apply there ↗
Anyscale
San Francisco, CA · $315,000 – $375,000
Lead Anyscale's technical support organization, guiding a team of engineers who help customers run production AI workloads on the platform while translating their feedback into product improvements. This role suits experienced engineering leaders comfortable with both hands-on technical problem-solving and building operational systems that scale.
Listed on Anyscale’s careers site · Apply there ↗
Motional
Full-time · Boston, MA · $123,000 – $163,500 · Remote
This role involves building infrastructure and tools that allow machine learning researchers and engineers to develop autonomous driving models more effectively. You'd optimize training systems, scale cloud platforms, and create AI tooling to remove friction from the ML development process.
Listed on Motional’s careers site · Apply there ↗
Crusoe
Denver, CO · $175,000 – $250,000
Crusoe seeks a senior solutions engineer to help enterprise customers deploy AI/ML workloads on their GPU infrastructure, serving as the technical bridge between customers and the engineering team. The role combines hands-on infrastructure work—building and optimizing Kubernetes-based systems—with customer engagement, technical storytelling, and feedback that shapes the platform.
Listed on Crusoe’s careers site · Apply there ↗
Full-time · Remote - United States · $230,000 – $322,000
Reddit seeks a Staff ML Systems Engineer to design and build the infrastructure that powers their recommendation and content discovery systems at massive scale. You'll architect MLOps platforms, optimize distributed training pipelines, and work with graph structures containing billions of nodes while collaborating with ML teams across the company.
Listed on Reddit’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
Cohere is hiring a research engineer to develop and scale machine learning infrastructure for large language model post-training, with a focus on distributed reinforcement learning methods. This role suits someone who combines strong software engineering discipline with deep interest in optimizing model training at scale.
Listed on Cohere’s careers site · Apply there ↗
Anyscale
San Francisco, CA · $170,000 – $199,000 · Remote
This role combines technical support with machine learning expertise, helping customers successfully deploy and scale AI applications on Anyscale's distributed computing platform. It suits experienced ML practitioners who enjoy troubleshooting complex systems, mentoring others, and collaborating across engineering teams to solve customer challenges.
Listed on Anyscale’s careers site · Apply there ↗
Remote - United States · $216,700 – $303,400
Reddit seeks a senior engineer to design and build the infrastructure platforms that enable the company's machine learning teams to train, optimize, and deploy recommendation and content discovery models at scale. This role suits experienced ML infrastructure engineers who enjoy solving complex systems problems in distributed computing environments and collaborating closely with model development teams.
Listed on Reddit’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
A technical role focused on advancing large language model post-training, performance optimization, and shipping production systems at scale. Suited for software engineers and researchers who want to bridge machine learning research with production infrastructure on some of the world's largest compute clusters.
Listed on Cohere’s careers site · Apply there ↗
Anyscale
San Francisco, CA · $170,000 – $245,000 · Remote
Anyscale is hiring an engineer to optimize large-scale LLM inference systems, building infrastructure that developers can use to run machine learning models efficiently from laptop to cluster. This role suits someone with distributed systems expertise who wants to work on high-performance AI infrastructure and contribute to open-source projects like Ray and vLLM.
Listed on Anyscale’s careers site · Apply there ↗