Skip to main content

Staff + Sr. Software Engineer, Cloud Inference

Anthropic
Hybrid - San Francisco, CA, at least 25% in officeUpdated 1d ago
Compensation
$320k–$485k
Published range · Top quartile for Engineering (520 listings)
Location
Hybrid - San Francisco, CA, at least 25% in office
Remote eligibility
Employment
Full-time
Staff / Principal
Role family
Engineering
AI / ML
Apply on job-boards.greenhouse.io
Job actionsApply now
Job actionsApply now

About the job

About the role

Anthropic's mission is to create reliable, interpretable, and steerable AI systems. The Cloud Inference team scales and optimizes Claude to serve developers and enterprises across AWS, GCP, Azure, and future cloud service providers (CSPs). The team owns the end-to-end product of Claude on each cloud platform, from API integration and intelligent request routing to inference execution, capacity management, and day-to-day operations.

Key responsibilities

  • Design, build, and own backend services and infrastructure that serve Claude across multiple CSPs, accounting for differences in compute hardware, networking, APIs, and operational models
  • Work cross-functionally with internal inference, product API, systems, and security teams, and with CSP partners to stand up the full serving stack on new cloud platforms, resolve operational issues, and influence provider roadmaps
  • Build and evolve CI/CD automation systems, including validation and deployment pipelines, that reliably ship new model versions to millions of users across cloud platforms
  • Design interfaces and tooling abstractions across CSPs that enable cost-effective inference management and reduce per-platform complexity
  • Contribute to capacity planning, autoscaling, and workload routing strategies that match supply with demand and direct requests to the most cost-effective accelerator and region
  • Analyze observability data across providers to identify performance bottlenecks, cost anomalies, and regressions

Minimum qualifications

  • Significant software engineering experience with high-performance, large-scale distributed systems serving millions of users
  • Experience building or operating services on at least one major cloud platform (AWS, GCP, or Azure), with exposure to Kubernetes, Infrastructure as Code, or container orchestration
  • Curiosity about LLM serving; prior inference or ML experience is not required
  • Experience working with external partners to align goals and deliver impact
  • Highly autonomous, able to take ownership of problems end-to-end

Preferred qualifications

  • Direct experience working with CSPs to scale infrastructure or products across multiple platforms
  • Hands-on experience with capacity management, cost optimization, or resource planning at scale
  • Solid understanding of multi-region deployments, geographic routing, and global traffic management
  • Proficiency in Python or Rust

Compensation

Annual salary: $320,000—$485,000 USD.

Logistics

Minimum education: Bachelor's degree or equivalent combination of education, training, and/or experience. Location-based hybrid policy: all staff are expected to be in one of Anthropic's offices at least 25% of the time. Visa sponsorship is offered.

Skills & tags

What you can verify before applying

Compare the essentials before you leave: pay, remote scope, employment type, source, and the employer apply destination.