Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.
106 open listings · Page 1 of 5
Roles posted by organizations here, alongside roles we found on employers’ own careers sites. Crawled roles say so on the card and send you to the employer to apply. Only show organizations on Career App
AeroVironment
Full-time · Huntsville, AL
This role develops and validates physics-based 3D synthetic environments for military simulation and testing, requiring deep expertise in radiometric rendering, graphics programming, and real-time systems integration. It suits experienced software engineers comfortable leading technical work on complex modeling systems within defense contracting.
Listed on AeroVironment’s careers site · Apply there ↗
Visa Inc.
Full-time · $123,000 – $190,900
This role involves building secure, production-grade AI agent systems and generative AI capabilities for enterprise use across Visa's business units. It suits engineers with hands-on experience developing AI applications who want to work on infrastructure and tools that scale across a large financial organization.
Listed on Visa Inc.’s careers site · Apply there ↗
Accenture
This role leads the design and delivery of AI-ready infrastructure platforms for enterprise clients, handling GPU-based systems, cloud migrations, and high-performance computing environments. It suits experienced infrastructure leaders comfortable with client engagement, team leadership, and emerging AI compute technologies.
Listed on Accenture’s careers site · Apply there ↗
NVIDIA
$152,000 – $287,500
NVIDIA seeks an Infrastructure Solutions Architect to guide deployment of next-generation data center GPUs and networking platforms at hyperscale, serving as a technical bridge between product teams and large enterprise customers. This role combines hands-on infrastructure expertise with cross-functional leadership to accelerate adoption of NVIDIA technologies globally.
Listed on NVIDIA’s careers site · Apply there ↗
Intel
Full-time · $170,500 – $315,490
Intel is seeking a software engineer to design and optimize neural network performance libraries in the oneDNN project, supporting AI frameworks across CPUs and GPUs. The role suits developers with deep expertise in performance engineering, low-level optimization, and linear algebra who want to influence AI infrastructure used across the industry.
Listed on Intel’s careers site · Apply there ↗
Unity
Mountain View, CA · $278,100 – $417,100
This role leads the engineering effort to run advanced AI models efficiently on consumer devices within a web-native runtime, optimizing for speed, memory, and power. It suits a senior systems engineer with deep expertise in model deployment, GPU programming, and real-time performance optimization who thrives on closing the gap between research models and shipped products.
Listed on Unity’s careers site · Apply there ↗
NIO
Full-time · San Jose, CA · $192,100 – $249,600
NIO is hiring a senior engineer to build production inference systems for large language and vision models across cloud and edge devices in their autonomous vehicle platform. This role suits someone with deep experience optimizing AI workloads on accelerators who wants to ship real-world impact at scale.
Listed on NIO’s careers site · Apply there ↗
Motional
Full-time · Boston, MA · $165,000 – $197,698 · Remote
This role involves developing and optimizing AI models for autonomous vehicle perception systems, requiring hands-on experience with machine learning deployment, robotics integration, and GPU acceleration. It suits engineers with automotive software backgrounds who are comfortable moving models from research to production and iterating based on real-world performance.
Listed on Motional’s careers site · Apply there ↗
Together AI
Full-time · San Francisco, CA · $200,000 – $290,000
A Research Engineer role focused on building and optimizing large-scale training infrastructure for foundation models at Together AI. This suits engineers who combine systems expertise with ML knowledge and enjoy translating research into production systems that serve real customers.
Listed on Together AI’s careers site · Apply there ↗
Crusoe
2 locations · $250,000 – $300,000
Crusoe is seeking an experienced engineer to optimize how large language models run in production, focusing on inference performance, cost, and reliability. This role combines deep systems work with customer collaboration, taking optimizations from concept through to deployed, monitored services.
Listed on Crusoe’s careers site · Apply there ↗
Deepgram
USA | Remote · $152,000 – $208,000
This role leads the design and delivery of ML infrastructure, model serving systems, and AI tooling that enable Deepgram's research and engineering teams to build and deploy voice AI models at scale. You'll coordinate across research, engineering, and product to translate ML requirements into production systems while optimizing for cost, latency, and developer velocity.
Listed on Deepgram’s careers site · Apply there ↗
Modal
New York, NY · $150,000 – $350,000
Modal is hiring a researcher to lead inference optimizations for their LLM serving platform, focusing on techniques like speculative decoding and quantization that reduce cost and latency. This role suits someone with a background shipping inference systems or research who can independently drive projects from conception through deployment.
Listed on Modal’s careers site · Apply there ↗
Pika Labs
Palo Alto, CA · $250,000 – $350,000
A senior or staff-level role optimizing AI model inference performance at a video generation startup. This suits engineers with deep expertise in GPU acceleration, distributed systems, and deploying large language models and video models efficiently at scale.
Listed on Pika Labs’s careers site · Apply there ↗
Menlo Ventures Portfolio
New York, NY · $405,000 – $485,000 · Remote
Anthropic seeks a Staff Engineer to lead the technical direction of their inference runtime—the foundational layer serving Claude to millions of users across GPU, TPU, and Trainium platforms. This role combines hands-on systems work in Rust and Python with cross-organizational leadership, suited to someone with deep expertise in high-performance infrastructure who has anchored a platform through scale.
Listed on Menlo Ventures Portfolio’s careers site · Apply there ↗
XPeng
Internship · Santa Clara, CA
This role involves optimizing AI inference performance on embedded automotive platforms, focusing on latency reduction and power efficiency for production vehicles. It suits someone with strong systems-level expertise in C++ and performance analysis who wants to work on autonomous driving at scale.
Listed on XPeng’s careers site · Apply there ↗
Anyscale
San Francisco, CA · $170,000 – $245,000 · Remote
Anyscale is hiring an engineer to optimize large-scale LLM inference systems, building infrastructure that developers can use to run machine learning models efficiently from laptop to cluster. This role suits someone with distributed systems expertise who wants to work on high-performance AI infrastructure and contribute to open-source projects like Ray and vLLM.
Listed on Anyscale’s careers site · Apply there ↗
Bend Studio
Full-time · San Mateo, CA · $193,300 – $289,900
This role focuses on architecting and optimizing video codecs for PlayStation platforms and cloud gaming, requiring deep expertise in compression standards and systems-level performance tuning. It suits experienced engineers with strong C/C++ fundamentals who want to work on large-scale multimedia infrastructure affecting millions of users.
Listed on Bend Studio’s careers site · Apply there ↗
Cohere
Internship · Canada, Europe, United States, United Kingdom · Remote
This is a machine learning internship at an AI company where you'll work on training and deploying large-scale foundation models, developing new training techniques, and collaborating with product teams. It suits students passionate about NLP and applied machine learning who want hands-on experience with cutting-edge models and infrastructure.
Listed on Cohere’s careers site · Apply there ↗
AST SpaceMobile
Full-time · Midland, TX
This role oversees day-to-day operations of an on-site AI lab for a spacecraft company, managing physical security, IT infrastructure, and a part-time team that labels images to train computer-vision models. It suits someone with lab or secure-facility management experience who is comfortable enforcing strict information-security rules and supervising hourly workers in a manufacturing environment.
Listed on AST SpaceMobile’s careers site · Apply there ↗
Agility Robotics
Full-time · Fremont, CA · $155,000 – $241,000
This role leads the development of real-time motion planning and navigation systems for humanoid robots operating in logistics and manufacturing environments. It suits experienced robotics engineers who want to own critical systems that enable autonomous locomotion, collision avoidance, and multi-robot coordination at scale.
Listed on Agility Robotics’s careers site · Apply there ↗
Voyager
Full-time · Washington, DC · $165,000 – $250,000
This role develops fast physics simulation engines and AI-accelerated surrogates that let autonomous agents optimize hardware designs across millions of iterations. It suits computational scientists or engineers who thrive at the intersection of numerical methods, GPU computing, and machine learning, and who want to compress years-long engineering cycles into days.
Listed on Voyager’s careers site · Apply there ↗
Intuit
3 locations · $202,500 – $274,000
A Staff Machine Learning Engineer will design, build, and deploy data science models at scale within Intuit's data science team, working across the full pipeline from data preparation through A/B testing and impact analysis. This role suits experienced ML engineers comfortable with the mathematical foundations of machine learning and production software engineering practices who want to work on problems serving millions of users.
Listed on Intuit’s careers site · Apply there ↗
Runway
Remote · $260,000 – $370,000
This role involves developing computer vision and generative models for creative AI tools, bridging research and production to enable artists and scientists to work with world simulation technology. It suits researchers with strong machine learning fundamentals who want their work to reach real users quickly.
Listed on Runway’s careers site · Apply there ↗
AeroVironment
Full-time · Huntsville, AL
This role develops physics-based 3D rendering software for realistic synthetic environments used in defense applications, combining radiometric modeling, graphics programming, and real-time system integration. It suits software engineers with graphics and Linux experience who want to work on complex visualization systems in a defense contracting environment.
Listed on AeroVironment’s careers site · Apply there ↗
NVIDIA
Full-time · Santa Clara, CA · $200,000 – $322,000
This role creates technical content and developer resources to help enterprise customers adopt NVIDIA's AI software stack, from documentation and demos to deployment guides and reference architectures. It suits someone with deep AI/ML experience who enjoys building clear explanations of complex systems and collaborating across product, engineering, and field teams.
Listed on NVIDIA’s careers site · Apply there ↗