320K–485K USD / year

Staff + Sr. Software Engineer, Cloud Inference Launch Engineering

AIHybrid — San Francisco, CA | Seattle, WA
Published on 2026-09-28
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

The role involves working on the Cloud Inference team to optimize and scale Claude across various cloud platforms. You'll be responsible for the validation pipeline for inference servers and load balancers, ensuring that model launches and performance improvements are executed correctly and efficiently.

About the company

Anthropic builds reliable and interpretable AI systems, including the Claude family of models, for businesses and developers. The company focuses on AI safety research and practical AI assistants for enterprise and consumer use.

The team

Cloud Inference team

Stack

PythonRustAWSGCPAzureKubernetes

What you'll do

  • Oversee frontier model launches and inference for new architectures
  • Integrate new inference features to cloud platforms
  • Identify and fix inference behavior gaps across platforms
  • Design CI/CD infrastructure for inference server
  • Improve validation speed and reliability
  • Analyze observability data to address performance bottlenecks

What we're looking for

  • Strong interest in LLM serving
  • Significant software engineering experience
  • Track record of building automation or test infrastructure
  • Experience with major cloud platforms
  • Collaboration skills with internal and external teams
  • Ability to quickly learn new technologies

Nice to have

  • LLM inference optimization skills
  • Experience with multi-region deployments
  • Familiarity with global traffic management
  • Experience working with CSP partner teams

Benefits

  • Equity donation matching
  • Generous vacation and parental leave
  • Flexible working hours
View original job post