Staff and Senior Software Engineer Inference
AI safety and research company building reliable, interpretable, and steerable AI systems, including the Claude product family and developer platform.
Maintainer signals as of 9/24/2026
Funding history
About Anthropic
Anthropic PBC develops frontier AI systems and deploys them through Claude products and the Claude Platform, with a stated focus on safety, interpretability, and steerability.
Skills
Candidate Availability
Required and preferred rules are kept separate and reflect the wording in the original posting.
About the Role
You will design, build, and operate distributed systems that serve models at scale. You will develop request routing, load balancing, traffic management, autoscaling, and fleet orchestration systems. You will build deployment pipelines, integrate accelerator platforms, improve inference performance, and provide infrastructure that supports model research and production workloads.
Requirements
- Significant software engineering experience with distributed systems
- Interest in machine learning systems and infrastructure
- Experience with high-performance large-scale distributed systems
- Experience deploying machine learning systems at scale
- Knowledge of load balancing, request routing, or traffic management
- Familiarity with LLM inference optimization, batching, and caching
- Experience with Kubernetes and cloud infrastructure
- Proficiency in Python or Rust
Responsibilities
- Design, build, and maintain distributed inference systems
- Develop resilient systems that adapt to real-world events
- Build request routing, load balancing, and traffic management systems
- Autoscale and orchestrate production, research, and experimental workloads
- Build and operate model deployment pipelines
- Provide high-performance inference infrastructure
- Integrate AI accelerator platforms and support new model architectures
Benefits
- Optional equity donation matching
- Generous vacation
- Parental leave
- Flexible working hours
