A complete application in a minute — tailored resume and cover letter, ready to send.
Anthropic in Seattle seeks a senior software engineer to design, build, and maintain distributed inference systems that serve Claude at scale. You will tackle routing, load balancing, and fleet orchestration across cloud providers, accelerators, and Kubernetes, with a focus on performance and reliability across millions of users.
You should have extensive experience with large-scale distributed systems, Python or Rust, and ML systems at scale; familiarity with AWS/GCP/Azure, model serving
Care about the societal impacts of your workThrive in environments where technical excellence directly drives both business results and research breakthroughsSignificant software engineering experience, particularly with distributed systemsDesire to learn more about machine learning systems and infrastructureWillingness to pick up slack, even if it goes outside your job descriptionResults-oriented, with a bias towards flexibility and impactWe encourage you to apply even if you do not believe you meet every single qualificationExperience with high-performance, large-scale distributed systemsExperience with Kubernetes and cloud infrastructure (AWS, GCP, Azure)Familiarity with LLM inference optimization, batching, and caching strategiesExperience with load balancing, request routing, or traffic management systemsProficiency in Python or RustExperience implementing and deploying machine learning systems at scale