# LLM Algorithmic Optimization Engineer - Intern

Hiring organization: [NIO](https://career.thegoodapps.co/organizations/nio)

Canonical page: https://career.thegoodapps.co/jobs/56cfbfe2-7ea6-47e8-a8aa-d67fccc6edb9

Listed on NIO's own careers site. Applications go to them directly.

- Employment type: internship
- Seniority: Intern
- Location: San Jose, CA
- Salary: 38 – 46 USD per hour

## Summary

This internship focuses on optimizing large language models and multimodal models for efficient inference and deployment in NIO's autonomous vehicle systems. It suits recent computer science graduates or master's students with strong foundations in deep learning, GPU optimization, and systems-level thinking who want to bridge AI research with real automotive applications.

_Our summary, not NIO's wording._

## Skills named

C++, Python, PyTorch

## Required

- Currently pursuing or completed PhD or Master's degree in Computer Science, Computer Engineering, Applied Mathematics, Communications, Electronics, or related field
- Strong understanding of GPU/NPU architecture and optimization techniques
- Knowledge of LLM and VLM architectures and transformer-based algorithms
- Proficiency in Python
- Experience with PyTorch or similar AI training/inference tools
- Proficiency in C/C++
- Hands-on experience with ONNX or similar model-serving frameworks
- Familiarity with debugging in distributed computing environments

## Nice to have

- PhD in computer science, artificial intelligence, or related fields
- 3+ years of relevant industry experience alongside a master's degree
- Experience optimizing deep learning model inference on hardware architectures
- Familiarity with microkernel architecture, Linux kernel, hypervisor, and middleware
- Published research record with high-impact papers
- Experience with LLM inference optimization on resource-constrained edge devices

Apply on NIO's site: https://nio.wd3.myworkdayjobs.com/en-US/NIO_Careers/job/San-Jose-US/LLM-Algorithmic-Optimization-Engineer---Intern_R-000119
