Senior Engineering Manager, Capacity Engineering
Anthropic · San Francisco, CA | New York City, NY | Seattle, WA
About this role
**Senior Engineering Manager, Capacity Engineering** **About Anthropic** Anthropicís mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. **About the Role** Lead the Capacity Engineering team responsible for managing one of the largest and fastest-growing infrastructure fleets in the industry. You'll own the data, tooling, and operational systems that let Anthropic plan, measure, and maximize utilization across compute resources - one of the company's largest areas of spend. This is a hands-on leadership role: stay close enough to systems to review designs, make architectural calls, and step into incidents when needed - while spending most of your time on people, priorities, and cross-organizational alignment. **Key Areas of Responsibility** • **Data Platform** - Pipelines that ingest occupancy and utilization telemetry, normalize billing across cloud providers, and serve tables for research engineers, finance, and leadership • **Planning and Assurance** - Cluster health tooling, capacity planning platforms, alerting, and systemic fixes to scheduling and fragmentation • **Efficiency** - Measuring and improving hardware utilization across training, inference, and evals **What You Bring** ✓ Experience managing software/infrastructure engineering teams with hiring and people development ✓ Strong technical background in production systems (data engineering, infrastructure, distributed systems, observability) ✓ Familiarity with major cloud providers (AWS, GCP, Azure), Kubernetes, and modern observability stacks ✓ Track record setting and executing engineering roadmaps in ambiguous, high-autonomy environments ✓ Excellent communication skills across technical and business stakeholders ✓ Comfort owning operational responsibility and on-call management **Preferred Qualifications** • Experience with capacity planning, resource management, or FinOps at hyperscalers or large-scale ML environments • Familiarity with accelerator infrastructure and GPU/TPU metrics • Experience with multi-cloud billing and telemetry normalization • Background building internal data products with self-service access • Experience with scheduling, packing efficiency, or workload optimization **Compensation & Logistics** 💰 Annual Salary: $405,000 – $485,000 USD 📍 Hybrid (25% office time minimum) 🎓 Bachelor's degree or equivalent required 🛂 Visa sponsorship available *We encourage applications even if you don't meet every qualification. We value diverse perspectives and actively work to support underrepresented groups in tech.*
Listing freshness
CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.