Machine Learning Infrastructure Engineer, Safeguards Research

San Francisco, CA | New York City, NY

Posted 9d ago

Job Location

San Francisco, CA | New York City, NY

Tech Stack

Remote Work Policy

On-site

Categories

AI Infrastructure Engineer

About the job

Anthropic's Safeguards team is responsible for developing systems that detect and mitigate misuse of AI models. This role focuses on building and owning the infrastructure that supports the research efforts of this team. You will create the tooling, pipelines, and abstractions that enable researchers to run experiments, train detection methods, and select detections for launch efficiently and reliably. This position bridges the gap between research and production, ensuring fast iteration for researchers and dependable results for detection systems as models evolve. The ideal candidate will have a proven ability to solve large-scale systems and data problems and a strong desire to deepen their machine learning expertise.

Responsibilities

  • Build and scale infrastructure and data pipelines for Safeguards machine learning research.
  • Own training, evaluation, and scoring workflows for researchers, prioritizing speed from idea to result.
  • Design tooling and interfaces, including libraries and command-line tools, for direct researcher use.
  • Incorporate correctness and sanity checking into the stack to maintain trustworthy results.
  • Transition high-value research workflows from experiments to production-grade jobs.
  • Enhance throughput, cost-efficiency, and reliability of large-scale inference and scoring workloads.
  • Collaborate with researchers and engineers to understand and anticipate evolving workflow needs.

Requirements

  • Strong software engineering fundamentals and Python proficiency.
  • Experience building and operating data-intensive or distributed systems in production.
  • Experience building tooling or infrastructure used by other engineers or researchers.
  • Comfort working across the research-to-deployment pipeline.
  • Ability to debug performance and correctness issues in unfamiliar stacks.
  • Strong written and verbal communication skills.
  • Collaborative approach to technical decisions.

Benefits

  • Annual compensation range: $350,000 - $500,000 USD
  • Visa sponsorship available
  • Hybrid work policy (at least 25% in office)

About Anthropic

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.