$192,100 – $249,600
Listed on NIO’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
NIO is hiring a senior engineer to build production inference systems for large language and vision models across cloud and edge devices in their autonomous vehicle platform. This role suits someone with deep experience optimizing AI workloads on accelerators who wants to ship real-world impact at scale.
Our summary, not NIO’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 5+ years building and optimizing AI inference systems at scale
- Deep knowledge of LLM/VLM model internals and Transformer architectures
- Performance engineering expertise in kernel development and memory optimization
- GPU/NPU programming proficiency
- Strong C/C++ skills with track record of production software
- Computer architecture and systems programming foundation
- BS/MS in Computer Science or related field
Nice to have
- Master's or PhD in Computer Science, Electrical/Computer Engineering, or related field plus 5 years industry experience
- Experience building inference serving systems with batching, scheduling, and caching
- Hardware-aware model optimization expertise
- Edge and embedded AI experience with real-time constraints
- Open source or proprietary framework contributions