# LLM Algorithmic Optimization Engineer

Hiring organization: [NIO](https://career.thegoodapps.co/organizations/nio)

Canonical page: https://career.thegoodapps.co/jobs/677dcb76-2013-4056-837f-e42d8c1ec94b

Listed on NIO's own careers site. Applications go to them directly.

- Employment type: full time
- Location: San Jose, CA
- Salary: 143200 – 186000 USD per year

## Summary

NIO seeks an engineer to research and optimize large language models and multimodal models for efficient inference and deployment in electric vehicles, particularly in digital cockpit and autonomous driving systems. This role combines algorithmic research with systems-level optimization across heterogeneous hardware, suited to someone with strong foundations in deep learning, GPU architecture, and a track record in model optimization.

_Our summary, not NIO's wording._

## Skills named

C++, Python, PyTorch

## Required

- PhD or Master's in Computer Science, Computer Engineering, Applied Mathematics, Communications, Electronics, or related field
- Strong understanding of GPU/NPU architecture and optimization
- Proficiency in LLM and VLM architectures and algorithms
- Proficiency in Python
- Experience with PyTorch or similar AI training/inference tools
- Proficiency in C/C++
- Familiarity with at least one LLM inference engine
- Hands-on experience with ONNX or similar model-serving frameworks
- Ability to debug code in distributed computing environments

## Nice to have

- PhD in computer science, artificial intelligence, or related field (versus Master's)
- 3+ years of industry experience with inference optimization
- Experience optimizing deep learning models on resource-constrained edge devices
- Familiarity with microkernel architecture, Linux kernel, hypervisor, middleware, and application frameworks
- Published high-impact, innovative research papers

Apply on NIO's site: https://nio.wd3.myworkdayjobs.com/en-US/NIO_Careers/job/San-Jose-US/LLM-Algorithmic-Optimization-Engineer_R-000124
