$143,200 – $186,000
Listed on NIO’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
NIO seeks an engineer to research and optimize large language models and multimodal models for efficient inference and deployment in electric vehicles, particularly in digital cockpit and autonomous driving systems. This role combines algorithmic research with systems-level optimization across heterogeneous hardware, suited to someone with strong foundations in deep learning, GPU architecture, and a track record in model optimization.
Our summary, not NIO’s wording. The full posting is on their site.
What they ask for
Required
- PhD or Master's in Computer Science, Computer Engineering, Applied Mathematics, Communications, Electronics, or related field
- Strong understanding of GPU/NPU architecture and optimization
- Proficiency in LLM and VLM architectures and algorithms
- Proficiency in Python
- Experience with PyTorch or similar AI training/inference tools
- Proficiency in C/C++
- Familiarity with at least one LLM inference engine
- Hands-on experience with ONNX or similar model-serving frameworks
- Ability to debug code in distributed computing environments
Nice to have
- PhD in computer science, artificial intelligence, or related field (versus Master's)
- 3+ years of industry experience with inference optimization
- Experience optimizing deep learning models on resource-constrained edge devices
- Familiarity with microkernel architecture, Linux kernel, hypervisor, middleware, and application frameworks
- Published high-impact, innovative research papers