Skip to main content
CareerApp

Skill

Model Optimization

Data Science, Analytics and AI/ML

A set of techniques for improving a machine learning model's efficiency, speed, or accuracy, including methods like quantization, pruning, distillation, and hyperparameter tuning. ML engineers apply these techniques to reduce a model's memory footprint and inference latency, particularly for deployment on constrained hardware like mobile devices or edge servers. It balances trade-offs between model size, computational cost, and predictive performance.

See who is hiring

Model Optimization in the job market

Last checked September 12, 2026

Open roles
7

on Career App right now

Employers
7

hiring for it

Median pay

not enough disclosed

Disclose pay

of these roles

Open roles requiring Model Optimization (7)

Research Staff, Voice AI Foundations

Deepgram

San Francisco, CA · $150,000 – $250,000 · Remote

You'll lead research into latent space models for voice AI, tackling fundamental challenges in audio compression, generative speech synthesis, and data efficiency. This role suits researchers who thrive on unsolved problems, move rapidly between theory and implementation, and want to pioneer entirely new approaches to making voice AI accessible at scale.

Listed on Deepgram’s careers site · Apply there ↗

Staff Machine Learning Engineer

Unity

Full-time · $218,400 – $283,900

This role involves optimizing state-of-the-art AI models to run efficiently on mobile and desktop devices within a browser-native runtime, handling everything from model export through kernel-level tuning to shipped features. It's ideal for a performance-focused engineer who thrives on closing the gap between research models and production on-device products, working with transformers, diffusion networks, and vision-language models across constrained hardware.

Listed on Unity’s careers site · Apply there ↗

AI Infrastructure Engineer

NIO

Full-time · San Jose, CA · $192,100 – $249,600

NIO is hiring a senior engineer to build production inference systems for large language and vision models across cloud and edge devices in their autonomous vehicle platform. This role suits someone with deep experience optimizing AI workloads on accelerators who wants to ship real-world impact at scale.

Listed on NIO’s careers site · Apply there ↗

Member of Technical Staff - Research, Post-Training

Modal

New York, NY · $150,000 – $350,000

Modal is seeking a research scientist to develop post-training methods and infrastructure for large language models, combining theoretical advances with practical deployment at scale. The role suits researchers with a track record in reinforcement learning and foundation models who want to bridge academic research and production systems.

Listed on Modal’s careers site · Apply there ↗

Software Engineer, AI Inference / HPC

Topaz Labs

Full-time · Dallas, TX · $110,000 – $150,000

You'll optimize AI model inference performance across Topaz Labs' image and video enhancement platform, serving as the bridge between research and production while working with hardware partners. This role suits engineers who want to apply systems and optimization expertise to a rapidly growing AI company with millions of users.

Listed on Topaz Labs’s careers site · Apply there ↗

ML Research Scientist (Health & Sensing)

Eight Sleep

Full-time · San Francisco, CA

This role develops machine learning and AI systems that analyze sleep and health sensor data to create personalized insights and adaptive thermoregulation for Eight Sleep's Pod device. It suits researchers with expertise in machine learning applications to health data who want to ship real products at scale.

Listed on Eight Sleep’s careers site · Apply there ↗

Apply for Staff Research Engineer, Model Efficiency at Cohere on their site

Staff Research Engineer, Model Efficiency

Cohere

Full-time · New York, NY · CA$250,000 – CA$535,000 · Remote

Cohere is hiring a Staff Research Engineer to optimize how their large language models run in production, focusing on inference efficiency across the full stack from architecture to hardware. This role suits researchers with deep expertise in LLM optimization and strong software engineering skills who want to ship real performance improvements at a fast-growing AI company.

Listed on Cohere’s careers site · Apply there ↗

Employers hiring for Model Optimization

7 in total, most open roles first.

Related skills

Curated neighbors in the taxonomy, whether or not employers ask for them together.

Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.