$278,100 – $417,100
Listed on Unity’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role leads the engineering effort to run advanced AI models efficiently on consumer devices within a web-native runtime, optimizing for speed, memory, and power. It suits a senior systems engineer with deep expertise in model deployment, GPU programming, and real-time performance optimization who thrives on closing the gap between research models and shipped products.
Our summary, not Unity’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- 8+ years in software or ML engineering
- 4+ years focused on on-device or edge inference
- Shipped production deployments of transformer or diffusion models on mobile, desktop, or embedded hardware
- Hands-on WebGPU deployment experience with WGSL compute shader writing
- Deep expertise with at least one major inference runtime
- Low-level performance engineering with at least one GPU/compute API
- Understanding of quantization, pruning, weight sharing, and distillation techniques
- Knowledge of mobile SoCs and desktop/laptop GPUs
- Proficiency in TypeScript, JavaScript, and WGSL
- Technical leadership track record
Nice to have
- Experience shipping world-model, neural-rendering, or real-time generative pipelines on device
- Game engine or real-time graphics background
- Open-source contributions to ML inference frameworks or GPU libraries
- WebGPU specification and evolving compute features familiarity
- Compiler stack experience for kernel generation
- On-device benchmarking infrastructure and CI experience