$200,000 – $290,000
Listed on Together AI’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
Platform Engineer building backend services and infrastructure for Together AI's model customization and evaluation platform. This role suits engineers experienced in production infrastructure, Kubernetes, and systems that serve both user-facing and research-oriented workloads at scale.
Our summary, not Together AI’s wording. The full posting is on their site.
Skills this role names
- Amazon Web Services (AWS)
- Ansible
- ArgoCD
- GitHub Actions
- Go
- Grafana
- Kubernetes
- Linux
- Microsoft Azure
- Prometheus
- Python
- Terraform
Log in to see which of these are already on your profile.
What they ask for
Required
- 3+ years building infrastructure or backend services in production
- Design, operate, and troubleshoot production Linux and Kubernetes platforms
- Strong software engineering in Python or Go
- Infrastructure automation with Terraform or Ansible
- Monitoring and observability with Prometheus or Grafana
- CI/CD pipeline experience with GitHub Actions or ArgoCD
- Cloud environment administration (AWS, GCP, or Azure)
- Experience with hybrid bare-metal and cloud environments
- Strong communication and documentation skills
- Comfortable working across the full stack from cluster operations to backend services
Nice to have
- Large-scale production systems with high reliability requirements
- Pipeline orchestration frameworks like Kubeflow, Argo Workflows, or Flyte
- GPU workload management on HPC clusters
- Experience with NVIDIA networking stack (NCCL, Mellanox, GPUDirect RDMA)
- Deployment of AI training or inference services
- Networking fundamentals including TCP/IP, DNS, routing, load balancing, TLS
- Open-source contributions or maintenance