# Staff Platform Engineer, Service Infrastructure

Hiring organization: [Together AI](https://career.thegoodapps.co/organizations/together-ai)

Canonical page: https://career.thegoodapps.co/jobs/87e5d631-a63b-4bbf-a6a9-2e1122e02bcc

Listed on Together AI's own careers site. Applications go to them directly.

- Employment type: full time
- Seniority: Staff
- Location: San Francisco, CA
- Salary: 240000 – 280000 USD per year

## Summary

Together AI seeks a hands-on Staff Platform Engineer to define and evolve service infrastructure strategy across their product foundations, including Kubernetes, AWS, and Terraform patterns. This role suits experienced infrastructure leaders who thrive on turning repeated operational problems into scalable, reusable solutions and coordinating across teams to improve reliability and standards.

_Our summary, not Together AI's wording._

## Skills named

Amazon CloudFront, Amazon EKS, Amazon Virtual Private Cloud (VPC), Amazon Web Services (AWS), ArgoCD, Datadog, DNS, GitHub Actions, GitOps, Go, Helm, Istio, Kubernetes, Microsoft Sentinel, Prometheus, Python, SSL/TLS, Terraform, TypeScript

## Required

- 7+ years in platform engineering, service infrastructure, SRE, distributed systems or cloud infrastructure
- Deep production experience with Kubernetes including EKS, Helm, ArgoCD, Argo Rollouts, ingress, autoscaling, secrets, service identity, networking and progressive delivery
- Strong Terraform experience including module design, infrastructure CI/CD, policy enforcement and safe self-service workflows
- Experience operating CDNs, load balancers, DNS, TLS, and traffic management
- Proficiency in at least one infrastructure automation language such as Go, Python or TypeScript
- AWS experience including EKS, IAM, VPC, load balancing, Route 53, CloudFront and ECR
- Direct experience with observability systems including metrics, logs, traces, dashboards, alerting, SLOs and incident response
- Proven ability to lead cross-functional technical initiatives across product, infrastructure, networking and security teams
- Strong written communication with experience producing design docs, migration plans and technical standards
- Staff-level judgment to define ambiguous problems, make pragmatic tradeoffs and influence without authority

## Nice to have

- Experience building internal developer platforms or paved-path service frameworks used across teams
- Experience embedding infrastructure best practices into product engineering at scale
- Experience with service mesh or zero-trust infrastructure such as mTLS, SPIFFE, SPIRE, Cilium, Istio, Linkerd or Envoy
- Experience with policy-as-code systems such as OPA, Gatekeeper, Kyverno or Sentinel
- Experience with multi-region, multi-cluster, hybrid-cloud or cross-provider service networking
- Experience with supply-chain security, image signing, SBOMs, vulnerability management or compliance automation

Apply on Together AI's site: https://job-boards.greenhouse.io/togetherai/jobs/5180690007
