Skip to main content
CareerApp
Intel

AI Infrastructure Engineer

Intel

full time · Mid

$170,500 – $315,490

Listed on Intel’s own careers site. You apply with them directly — we never stand between you and the employer.

What this role is

Intel is seeking a performance engineer to optimize Large Language Model inference on their next-generation GPUs, working across the full stack from kernel development to open-source framework contributions. This role suits engineers passionate about squeezing maximum throughput from hardware and collaborating with the broader AI infrastructure community.

Our summary, not Intel’s wording. The full posting is on their site.

Skills this role names

Log in to see which of these are already on your profile.

What they ask for

Required

  • Bachelor's degree in Computer Science, Software Engineering, AI/ML or related field with 4+ years experience (or Master's with 3+ years, or PhD)
  • 3+ years of software engineering experience in GPU computing, AI systems, or HPC
  • Proficiency in modern C++ and Python
  • Ability to read and modify complex systems-level code

Nice to have

  • CPU/GPU architecture knowledge
  • Understanding of LLM architectures and inference paradigms (attention, KV caching, continuous batching, speculative decoding, prefill-decode disaggregation)
  • Prior open-source contributions to vLLM, SGLang, PyTorch, or llama.cpp
  • Experience writing and optimizing custom GPU kernels with Triton, SYCL, CUDA/CUTLASS, or similar DSLs
  • Experience with multi-node inference orchestration
  • Daily use of AI coding agents to accelerate workflow

Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.