Staff+ Software Engineer, ML Inference Path
San Francisco, CA
Posted 4d ago
About the job
The Safeguards ML Inference Path team designs, builds, and operates the production infrastructure that powers Claude's ML based safety systems. We collaborate closely with safety researchers and inference engineers to bring new classifiers and novel classes of ML defenses to production. We own the research → production transfer of new safety technologies that is on the critical path for every Claude model launch. We build for scale, serving thousands of ML classifiers for all requests on the token generation path and for every platform Claude runs on. We are looking for engineers who have deep expertise in productionizing ML systems, working at the intersection of machine learning, large-scale distributed systems, and AI safety, developing the platforms and tools that enable our safeguards to operate reliably at scale.
Responsibilities
- Design and build scalable ML infrastructure for real-time safety deployments.
- Build monitoring and observability tools for safety-critical applications.
- Collaborate with research teams to productionize safety research.
- Optimize inference latency and throughput for real-time safety evaluations.
- Implement automated testing, deployment, and rollback systems for ML models.
- Partner with Safeguards, Security, and Alignment teams to deliver infrastructure.
- Contribute to the development of internal tools and frameworks for safety research and deployment.
Requirements
- Proficiency in Python and experience with ML frameworks like PyTorch, TensorFlow, or JAX.
- Understanding of distributed systems principles and experience building high-throughput, low-latency systems.
- Experience building automated deployment pipelines and evaluation infrastructure.
- Experience implementing A/B testing frameworks and experimentation infrastructure for ML systems.
- Results-oriented with a bias towards reliability and impact in safety-critical systems.
- Enjoy collaborating with researchers and translating research into production systems.
- Care deeply about AI safety and the societal impacts of your work.
Benefits
- Annual compensation range: $320,000 - $485,000 USD
- Visa sponsorship available