Staff+ Software Engineer, Capacity Engineering
Menlo Ventures PortfolioNew York, NY · full time · Staff
$320,000 – $485,000
Listed on Menlo Ventures Portfolio’s own careers site. You apply with them directly — we never stand between you and the employer.
What this role is
This role owns the data pipelines, observability systems, and operational tooling that track and optimize Anthropic's massive multi-cloud GPU and accelerator fleet. It suits engineers who enjoy moving between data engineering, systems operations, and observability work, and who want to directly influence infrastructure decisions affecting billions of dollars in compute spending.
Our summary, not Menlo Ventures Portfolio’s wording. The full posting is on their site.
Skills this role names
Log in to see which of these are already on your profile.
What they ask for
Required
- Production system design and operation experience
- Production-quality Python and SQL code
- Deep experience with at least one major cloud provider (AWS, GCP, or Azure)
- Observability tooling stack experience including Prometheus, PromQL, and Grafana
- Ability to gather requirements and work across organizational boundaries in ambiguous environments
Nice to have
- Capacity planning or resource management experience at hyperscale or large ML environment
- Scheduling and packing efficiency optimization experience
- Multi-cloud data ingestion and billing export normalization experience
- Total cost of ownership and forecasting experience
- Accelerator infrastructure familiarity (GPU, TPU, Trainium metrics)
- Experience building internal data products with self-service access and schema contracts
- Storage efficiency and lifecycle program experience at exabyte scale