# Senior Backend Engineer, Inference Platform

Hiring organization: [Together AI](https://career.thegoodapps.co/organizations/together-ai)

Canonical page: https://career.thegoodapps.co/jobs/a308d18f-70f9-4123-8c05-7a97f1099f18

Listed on Together AI's own careers site. Applications go to them directly.

- Employment type: full time
- Seniority: Senior
- Location: San Francisco, CA
- Salary: 160000 – 250000 USD per year

## Summary

This role involves building and optimizing the core infrastructure that routes and balances inference requests across thousands of GPUs, working with cutting-edge AI hardware to make language models faster and more efficient at global scale. You'll partner with ML researchers to productionize new models while contributing to the open source tools that power the industry.

_Our summary, not Together AI's wording._

## Skills named

CUDA, Go, Kubernetes, Python, Rust, TypeScript

## Required

- 5+ years building large-scale distributed systems and API microservices
- Strong understanding of OS concepts: multi-threading, memory management, networking, storage performance
- Expert-level programming in Rust, Go, Python, or TypeScript
- Knowledge of modern LLMs and how they are served in production

## Nice to have

- Experience with the open source inference ecosystem (SGLang, vLLM, NVIDIA Dynamo)
- Kubernetes or container orchestration experience
- Familiarity with GPU software stacks (CUDA, Triton, NCCL) and HPC technologies (InfiniBand, NVLink, MPI)
- Bachelor's or Master's degree in Computer Science, Computer Engineering, or related field

Apply on Together AI's site: https://job-boards.greenhouse.io/togetherai/jobs/4835763007
