Engineering Manager, Scheduler And Fleet Efficiency

USD 405,000-485,000 per year
MIDDLE
✅ Hybrid
✅ Visa Sponsorship

Tech Stack

Communication @ 3 Engineering Management @ 5 Kubernetes @ 3 Observability @ 3

Details

Anthropic is seeking an engineering manager to lead the Scheduler team, which builds the scheduling layer for Anthropic's Kubernetes fleet, job-launch tooling, and fleet-efficiency systems. The team supports compute-intensive research and product workloads by determining where jobs run, managing demand when compute supply is constrained, and improving fleet utilization, scheduling predictability, and developer experience.

Responsibilities

  • Lead and grow a team of engineers building the scheduling platform, job-launch tooling, and fleet-efficiency systems, owning planning, execution, and delivery against key milestones.
  • Set technical direction for scheduling, placement, queueing, and quota across Anthropic's compute fleet.
  • Partner with capacity planning, research, inference, and product teams to bring workloads onto the paved path and make efficient scheduling decisions.
  • Drive the roadmap for scheduler capabilities, fleet utilization, and the developer experience of launching and managing jobs.
  • Define and track metrics for fleet efficiency and scheduling quality, including utilization, queue wait, and job-start latency.
  • Create clarity for the team and stakeholders in an ambiguous, fast-moving environment where demand for compute routinely exceeds supply.
  • Lead hiring, coaching, and career development with an inclusive and equitable approach.
  • Represent the team across the engineering organization and contribute to engineering-wide initiatives as a member of Anthropic's engineering management group.

Requirements

  • Experience managing and growing a team of software engineers.
  • A hands-on software engineering background as an individual contributor before moving into management.
  • Experience building or operating large-scale distributed or infrastructure systems in production.
  • Working knowledge of Kubernetes and cluster scheduling concepts, including resource requests and limits, affinity, priority and preemption, and custom schedulers or controllers.
  • Excellent written and verbal communication skills, including the ability to create clarity across teams.
  • Preferred: 5+ years of engineering management experience, including leading infrastructure, platform, or compute teams.
  • Preferred: Experience owning a cluster scheduler, job orchestration system, or resource manager at scale.
  • Preferred: Familiarity with scheduling machine-learning workloads on accelerators and the tradeoffs between utilization, fairness, and latency.
  • Preferred: Experience building developer tooling used by engineers daily.
  • Preferred: A background in observability or incident response for control-plane systems, with a track record of improving production reliability.
  • Preferred: A track record of building a culture of belonging and engineering excellence.
  • Preferred: Low ego, high empathy, and a habit of leading by example.
  • Minimum education is a bachelor's degree or an equivalent combination of education, training, and experience. The field of study must be relevant to the role as demonstrated through coursework, training, or professional experience.

Compensation And Benefits

  • Annual salary: $405,000–$485,000 USD.
  • Competitive compensation and benefits.
  • Optional equity donation matching.
  • Generous vacation and parental leave.
  • Flexible working hours.
  • Office space for collaboration.

Work Arrangement And Immigration

  • Location-based hybrid policy: Staff are expected to work from one of Anthropic's offices at least 25% of the time, although some roles may require more office time.
  • Anthropic sponsors visas and makes reasonable efforts to obtain visas for successful candidates, with support from an immigration lawyer.

More jobs at Anthropic

Similar jobs