CronJobs

devops-sre jobs

Engineering Manager, Scheduler and Fleet Efficiency

Anthropic · San Francisco, CA | New York City, NY

hybridsenior$405,000–$405,000Posted Sep 1, 2026KubernetesCluster SchedulingDistributed SystemsJob OrchestrationResource ManagementML Infrastructure

Apply on the employer site

About this role

## About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems—safe and beneficial for users and society. ## About the role Anthropic’s compute fleet is one of the largest and most varied in the world. The Scheduler team builds the scheduling layer for Anthropic’s Kubernetes fleet, the tools researchers and engineers use to launch and manage jobs, and the systems that help the fleet run as efficiently as possible. As an **Engineering Manager** for Scheduler, you’ll lead a team building infrastructure that the research and product organizations depend on—partnering closely with capacity planning, research, inference, and product to make efficient use of the fleet. ## Key responsibilities - Lead and grow a team of engineers building the scheduling platform, job-launch tooling, and fleet-efficiency systems; own planning, execution, and delivery against milestones - Set technical direction for scheduling, placement, queueing, and quota across the compute fleet - Partner with capacity planning, research, inference, and product teams to bring workloads onto the “paved path” and improve scheduling decisions - Drive the roadmap for scheduler capabilities, fleet utilization, and developer experience for launching/managing jobs - Define and track metrics for fleet efficiency and scheduling quality (e.g., utilization, queue wait, job-start latency) and hold the team accountable - Create clarity in an ambiguous, fast-moving environment where compute demand often exceeds supply - Take an inclusive, equitable approach to hiring, coaching, and career development; sustain a high-performing, healthy team - Represent the team across engineering and contribute to engineering-wide initiatives as part of Anthropic’s engineering management group ## Minimum qualifications - Experience managing and growing a team of software engineers - Hands-on software engineering background as an individual contributor prior to moving into management - Experience building or operating large-scale distributed or infrastructure systems in production - Working knowledge of Kubernetes and cluster scheduling concepts (resource requests/limits, affinity, priority & preemption, custom schedulers/controllers) - Excellent written and verbal communication skills, including the ability to create clarity across teams ## Preferred qualifications - 5+ years of engineering management experience leading infrastructure/platform/compute teams - Experience owning a cluster scheduler, job orchestration system, or resource manager at scale - Familiarity with scheduling ML workloads on accelerators and tradeoffs between utilization, fairness, and latency - Experience building developer tooling relied on daily by other engineers - Background in observability or incident response for control-plane systems; improving production reliability - Track record of building a culture of belonging and engineering excellence - Low ego, high empathy, and leading by example ## Logistics - **Minimum education:** Bachelor’s degree or equivalent combination of education/training/experience - **Required field of study:** Field relevant to the role (via coursework, training, or professional experience) - **Minimum years of experience:** Correlates with internal job level requirements - **Location-based hybrid policy:** Expected to be in one of Anthropic’s offices at least **25%** of the time (some roles may require more) - **Visa sponsorship:** Yes—however, not every rol

Listing freshness

CronJobs last confirmed this listing 19h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.

Browse all software engineering jobs →

Follow fresh jobs in Discord