Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.
206 open listings · Page 8 of 9
Roles posted by organizations here, alongside roles we found on employers’ own careers sites. Crawled roles say so on the card and send you to the employer to apply. Only show organizations on Career App
Motional
Full-time · Boston, MA · $159,000 – $207,000
This role focuses on optimizing machine learning model deployment and performance on embedded compute platforms for autonomous vehicles. You'll work across the full ML software stack, collaborating with deep learning teams to improve GPU/NPU utilization and develop next-generation AI compute systems.
Listed on Motional’s careers site · Apply there ↗
Menlo Ventures Portfolio
Full-time · New York, NY · $350,000 – $850,000 · Remote
Anthropic is hiring a research engineer to work on large language model pre-training, combining cutting-edge ML research with practical system engineering. The role suits someone who wants to contribute to safe AI development while working on problems spanning model architecture, training infrastructure, and algorithmic innovation.
Listed on Menlo Ventures Portfolio’s careers site · Apply there ↗
XPeng
Full-time · Santa Clara, CA · $174,720 – $295,680
This role involves building and optimizing large-scale vision-language-action foundation models that form the core of XPENG's autonomous driving systems. You'll design multi-modal architectures, lead pretraining strategies on massive fleet data, and collaborate across research and infrastructure teams to deploy intelligent models for next-generation vehicles.
Listed on XPeng’s careers site · Apply there ↗
NVIDIA
Santa Clara, CA · $320,000 – $488,750
NVIDIA seeks a distinguished architect to design next-generation communication technologies and platforms for deep learning and HPC at massive scale, working across GPU, networking, and software teams to push the boundaries of what's possible in data center performance.
Listed on NVIDIA’s careers site · Apply there ↗
Remote - United States · $216,700 – $303,400
Reddit seeks a senior engineer to design and build the infrastructure platforms that enable the company's machine learning teams to train, optimize, and deploy recommendation and content discovery models at scale. This role suits experienced ML infrastructure engineers who enjoy solving complex systems problems in distributed computing environments and collaborating closely with model development teams.
Listed on Reddit’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
This is a research-focused engineering role on Cohere's team developing code-generating AI models and autonomous agent systems for enterprise use. You'll blend research and engineering, implementing novel approaches to train and deploy scalable code models while collaborating with researchers to push the boundaries of what these systems can do.
Listed on Cohere’s careers site · Apply there ↗
Motional
Boston, MA · $240,000 – $330,000 · Remote
A leadership role on the machine learning team building behavior prediction systems for autonomous vehicles. This suits experienced ML engineers and technical leaders ready to manage a team while developing production models for self-driving cars.
Listed on Motional’s careers site · Apply there ↗
Menlo Ventures Portfolio
San Francisco, CA · $350,000 – $850,000
This role involves building infrastructure systems to train, evaluate, and deploy AI models at Anthropic, working across the full stack from data pipelines to performance optimization. It suits experienced infrastructure engineers comfortable with distributed systems, containerization, and large-scale ML workloads who want to contribute to AI safety research.
Listed on Menlo Ventures Portfolio’s careers site · Apply there ↗
XPeng
Full-time · Santa Clara, CA · $215,280 – $364,320
XPeng is seeking a machine learning engineer to develop vision-language-action foundation models for autonomous driving systems, working with massive datasets and distributed computing infrastructure. This role suits researchers and practitioners with deep learning expertise who want to shape next-generation self-driving technology.
Listed on XPeng’s careers site · Apply there ↗
NVIDIA
$152,000 – $241,500
This role analyzes large-scale GPU workload and infrastructure data to identify optimization opportunities, working across teams to turn telemetry into actionable insights and tools. It suits someone with deep data analysis experience who can move fluidly between exploration, visualization, and machine learning to solve production problems at scale.
Listed on NVIDIA’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
Cohere is hiring a research engineer to develop and scale machine learning infrastructure for large language model post-training, with a focus on distributed reinforcement learning methods. This role suits someone who combines strong software engineering discipline with deep interest in optimizing model training at scale.
Listed on Cohere’s careers site · Apply there ↗
Motional
Full-time · Boston, MA · $146,000 – $225,000 · Remote
This role involves developing and deploying machine learning models for autonomous vehicle perception and prediction, working alongside research scientists to advance self-driving technology. It suits experienced ML engineers with deep learning expertise who want to tackle real-world challenges in autonomous driving at a company scaling toward commercial deployment.
Listed on Motional’s careers site · Apply there ↗
NVIDIA
$184,000 – $287,500
NVIDIA seeks a senior engineer to develop and optimize JAX-based components for its AI platform, working on core infrastructure that powers deep learning research and real-world applications. This role combines systems-level performance optimization with collaborative work across research and product teams to advance numerical computing tools.
Listed on NVIDIA’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
A technical role focused on advancing large language model post-training, performance optimization, and shipping production systems at scale. Suited for software engineers and researchers who want to bridge machine learning research with production infrastructure on some of the world's largest compute clusters.
Listed on Cohere’s careers site · Apply there ↗
Motional
Boston, MA · $240,000 – $330,000 · Remote
Lead a team of machine learning engineers building prediction and motion planning models for autonomous vehicles, translating research into production systems deployed on self-driving fleets. This role suits experienced technical leaders who combine deep expertise in ML and autonomous driving with proven ability to manage and grow engineering teams.
Listed on Motional’s careers site · Apply there ↗
NVIDIA
$184,000 – $356,500
This role involves integrating CUDA features and distributed runtime technologies into AI frameworks like PyTorch and vLLM, working on the performance and scalability of deep learning systems across multi-GPU clusters. It suits experienced systems engineers who combine deep learning knowledge with low-level optimization expertise and want to shape the infrastructure that powers modern AI.
Listed on NVIDIA’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
This role involves building and optimizing the infrastructure that trains Cohere's large-scale AI models, working at the intersection of research and production engineering. You'll design scalable training systems, improve compute efficiency, and collaborate with top researchers to ship frontier models to enterprise customers.
Listed on Cohere’s careers site · Apply there ↗
NVIDIA
2 locations · $184,000 – $287,500
NVIDIA's GPU Communications Libraries and Networking team is hiring a senior architect to design next-generation communication systems that accelerate deep learning and HPC workloads across massive GPU clusters. This role combines systems design, performance optimization, and hardware-software co-design for some of the world's largest computing platforms.
Listed on NVIDIA’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
This role optimizes training performance for Cohere's large language models, focusing on throughput and accelerator utilization through software engineering and low-level kernel optimization. It suits engineers with strong systems programming skills who want to work on cutting-edge AI infrastructure and training infrastructure at scale.
Listed on Cohere’s careers site · Apply there ↗
NVIDIA
$152,000 – $287,500
NVIDIA seeks a deep learning engineer to integrate advanced communication technologies into AI frameworks like PyTorch and JAX, optimizing multi-GPU performance for training and inference workloads. The role involves hands-on development with communication libraries, compiler improvements, and performance analysis across large-scale distributed systems.
Listed on NVIDIA’s careers site · Apply there ↗
Cohere
Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote
Cohere is seeking a Senior Member of Technical Staff to design and develop multimodal AI systems that integrate text, speech, and vision capabilities. This role suits experienced machine learning engineers who want to work on frontier models at scale with a focused team and exceptional compute resources.
Listed on Cohere’s careers site · Apply there ↗
NVIDIA
$168,000 – $310,500
NVIDIA is seeking a senior physical design engineer to develop and optimize chip design methodologies across their product line, with focus on power, performance, and area efficiency on advanced semiconductor nodes. This role combines deep expertise in EDA tools and physical design flows with emerging opportunities to apply machine learning to accelerate design cycles.
Listed on NVIDIA’s careers site · Apply there ↗
NVIDIA
Santa Clara, CA · $152,000 – $287,500
NVIDIA seeks a senior software engineer to design and maintain optimized communication runtimes and system software for GPU clusters in AI and HPC applications. This role suits someone with deep systems knowledge who wants to work on foundational infrastructure powering machine learning and scientific computing.
Listed on NVIDIA’s careers site · Apply there ↗
NVIDIA
$152,000 – $287,500
NVIDIA seeks a performance engineer to optimize communication libraries (NCCL, NVSHMEM, UCX) that connect thousands of GPUs in deep learning and HPC systems. This role suits someone with systems software expertise who enjoys diagnosing performance bottlenecks across GPU clusters and networking stacks.
Listed on NVIDIA’s careers site · Apply there ↗
NVIDIA
Full-time · Santa Clara, CA · $184,000 – $356,500
Lead a team developing high-performance GPU communication libraries at NVIDIA, working on technologies like NCCL and NVSHMEM that enable distributed deep learning and HPC applications. This role combines technical depth with management responsibility, requiring you to design features, mentor engineers, and collaborate across the organization to deliver libraries that push the boundaries of what's possible at scale.
Listed on NVIDIA’s careers site · Apply there ↗