Skip to content
← Back to job listings

Staff + Senior Software Engineer (Cloud Inference)

Anthropic · San Francisco, United States

External listingfull-timeabout 2 months ago

About The Role

Join Anthropic, a leading AI safety and research company, as a Staff + Senior Software Engineer on the Cloud Inference team. In this role, you will design, build, and own backend services and infrastructure that serve our AI model, Claude, across multiple cloud service providers. You will work cross-functionally with internal teams and external partners, contribute to capacity planning and workload routing strategies, and analyze observability data to identify performance bottlenecks and cost anomalies. This position offers a comprehensive benefits package, including health insurance, paid parental leave, flexible paid time off, and competitive salary and equity packages.

  • Design, build, and own backend services and infrastructure that serve Claude across multiple cloud service providers (CSPs), accounting for differences in compute hardware, networking, APIs, and operational models.
  • Work cross-functionally with internal inference, product API, systems, and security teams, among others, and with CSP partners to stand up the full serving stack on new cloud platforms, resolve operational issues, and influence provider roadmaps.
  • Build and evolve CI/CD automation systems, including validation and deployment pipelines, that reliably ship new model versions to millions of users across cloud platforms without regressions.
  • Have experience working with external partners to align goals and deliver impact
  • Are curious about LLM serving; prior inference or ML experience is not required
  • Are a fast learner who can quickly ramp up on new technologies, hardware platforms, and provider ecosystems
  • Have significant software engineering experience, with a strong background in high-performance, large-scale distributed systems serving millions of users
  • Thrive in cross-functional collaboration with both internal teams and external partners
  • Have experience building or operating services on at least one major cloud platform (AWS, GCP, or Azure), with exposure to Kubernetes, Infrastructure as Code, or container orchestration
  • Are highly autonomous and take ownership of problems end-to-end, including work that falls outside your job description
  • Direct experience working with CSPs to scale infrastructure or products across multiple platforms, navigating differences in networking, security, privacy, billing, and managed service offerings
  • Hands-on experience with capacity management, cost optimization, or resource planning at scale across heterogeneous environments
  • Proficiency in Python or Rust
  • Solid understanding of multi-region deployments, geographic routing, and global traffic management
  • Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
  • Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this
  • We encourage you to apply even if you do not believe you meet every single qualification
  • Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
  • Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience
  • Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices

This is an external listing. JobSpring does not represent or verify the employer. Report this listing