Staff + Senior Software Engineer, Scaling

USD 320,000-485,000 per year
SENIOR
✅ Hybrid
✅ Visa Sponsorship

Tech Stack

AI AWS @ 4 Algorithms Azure @ 4 Distributed Systems @ 4 GCP @ 4 Kubernetes @ 4 LLM @ 3 Machine Learning @ 4 Networking Observability Python @ 6 Rust @ 6

Details

Anthropic's Inference team builds and scales the systems that serve Claude to millions of users worldwide. The team operates compute-agnostic inference deployments across diverse AI accelerators and cloud platforms, covering intelligent request routing, fleet-wide orchestration, scaling, and networking.

Responsibilities

  • Design, build, and maintain distributed systems serving Claude at global scale.
  • Develop resilient systems that adapt to real-world events in real time.
  • Build intelligent request routing, load balancing, and traffic management systems across thousands of accelerators and multiple cloud providers.
  • Maximize compute efficiency and optimize fleet costs through autoscaling and orchestration of production, research, and experimental workloads.
  • Build and operate production-grade deployment pipelines for releasing new models.
  • Provide high-performance inference infrastructure for next-generation model development.
  • Integrate new AI accelerator platforms and support inference for new model architectures.
  • Design routing algorithms, manage multi-region deployments and geographic routing, and analyze observability data to tune production performance.

Requirements

  • Significant software engineering experience, particularly with distributed systems.
  • A results-oriented approach, flexibility, and willingness to take on work outside the formal job description.
  • Interest in learning about machine learning systems and infrastructure.
  • Ability to work in environments where technical excellence drives business results and research breakthroughs.
  • Experience with high-performance, large-scale distributed systems is preferred.
  • Experience implementing and deploying machine learning systems at scale is preferred.
  • Experience with load balancing, request routing, or traffic management systems is preferred.
  • Familiarity with LLM inference optimization, batching, and caching strategies is preferred.
  • Experience with Kubernetes and cloud infrastructure such as AWS, GCP, or Azure is preferred.
  • Proficiency in Python or Rust is preferred.
  • Bachelor's degree or equivalent combination of education, training, and experience in a relevant field, or equivalent demonstrated through coursework, training, or professional experience.

Benefits

Anthropic offers competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and office collaboration space. Staff are expected to work from an Anthropic office at least 25% of the time. Anthropic sponsors visas where possible and retains an immigration lawyer to assist with visa applications.

More jobs at Anthropic

Similar jobs