Skip to main content
CareerApp

Staff Platform Engineer, Service Infrastructure

Together AI

San Francisco, CA · full time · Staff

$240,000 – $280,000

Listed on Together AI’s own careers site. You apply with them directly — we never stand between you and the employer.

What this role is

Together AI seeks a hands-on Staff Platform Engineer to define and evolve service infrastructure strategy across their product foundations, including Kubernetes, AWS, and Terraform patterns. This role suits experienced infrastructure leaders who thrive on turning repeated operational problems into scalable, reusable solutions and coordinating across teams to improve reliability and standards.

Our summary, not Together AI’s wording. The full posting is on their site.

What they ask for

Required

  • 7+ years in platform engineering, service infrastructure, SRE, distributed systems or cloud infrastructure
  • Deep production experience with Kubernetes including EKS, Helm, ArgoCD, Argo Rollouts, ingress, autoscaling, secrets, service identity, networking and progressive delivery
  • Strong Terraform experience including module design, infrastructure CI/CD, policy enforcement and safe self-service workflows
  • Experience operating CDNs, load balancers, DNS, TLS, and traffic management
  • Proficiency in at least one infrastructure automation language such as Go, Python or TypeScript
  • AWS experience including EKS, IAM, VPC, load balancing, Route 53, CloudFront and ECR
  • Direct experience with observability systems including metrics, logs, traces, dashboards, alerting, SLOs and incident response
  • Proven ability to lead cross-functional technical initiatives across product, infrastructure, networking and security teams
  • Strong written communication with experience producing design docs, migration plans and technical standards
  • Staff-level judgment to define ambiguous problems, make pragmatic tradeoffs and influence without authority

Nice to have

  • Experience building internal developer platforms or paved-path service frameworks used across teams
  • Experience embedding infrastructure best practices into product engineering at scale
  • Experience with service mesh or zero-trust infrastructure such as mTLS, SPIFFE, SPIRE, Cilium, Istio, Linkerd or Envoy
  • Experience with policy-as-code systems such as OPA, Gatekeeper, Kyverno or Sentinel
  • Experience with multi-region, multi-cluster, hybrid-cloud or cross-provider service networking
  • Experience with supply-chain security, image signing, SBOMs, vulnerability management or compliance automation

Turn on analytics and we load Google Analytics: Google gets the pages you open and what you do here — searches, jobs you view, jobs you apply to — and sets its own cookies. Leave it off and the only cookies we set are your login, your theme, and this answer. Accept All also records a yes to advertising, which nothing uses yet. Privacy Policy.