$38 – $46 / hour
Listed on NIO’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This internship focuses on optimizing large language models and multimodal models for efficient inference and deployment in NIO's autonomous vehicle systems. It suits recent computer science graduates or master's students with strong foundations in deep learning, GPU optimization, and systems-level thinking who want to bridge AI research with real automotive applications.
Our summary, not NIO’s wording. The full posting is on their site.
What they ask for
Required
- Currently pursuing or completed PhD or Master's degree in Computer Science, Computer Engineering, Applied Mathematics, Communications, Electronics, or related field
- Strong understanding of GPU/NPU architecture and optimization techniques
- Knowledge of LLM and VLM architectures and transformer-based algorithms
- Proficiency in Python
- Experience with PyTorch or similar AI training/inference tools
- Proficiency in C/C++
- Hands-on experience with ONNX or similar model-serving frameworks
- Familiarity with debugging in distributed computing environments
Nice to have
- PhD in computer science, artificial intelligence, or related fields
- 3+ years of relevant industry experience alongside a master's degree
- Experience optimizing deep learning model inference on hardware architectures
- Familiarity with microkernel architecture, Linux kernel, hypervisor, and middleware
- Published research record with high-impact papers
- Experience with LLM inference optimization on resource-constrained edge devices