Staff+ Software Engineer, Safeguards Review Tooling

San Francisco, CA

Posted 14d ago

Job Location

San Francisco, CA

Tech Stack

Remote Work Policy

On-site

Categories

Applied AI Engineer

About the job

The Safeguards team is seeking engineers to build the Review Tooling team, which develops systems used by humans and AI to investigate potential harms and take enforcement actions across Anthropic's products and cloud platforms. This foundational role involves owning the tools safety investigators rely on, including analytics, privacy-preserving primitives, and a sandbox environment for rapid iteration. You will also drive the scaling of review through automation, enabling AI to augment human capabilities while keeping human judgment central where it matters most. The speed, clarity, and reliability of this tooling are critical for identifying harmful behavior, making sound enforcement decisions, and feeding signal back into model training. You will collaborate closely with policy, operations, data science, legal, and privacy teams to ensure effective, accurate, and trustworthy enforcement systems.

Responsibilities

  • Build investigation, review, and enforcement tooling for first-party and third-party platforms, including case queues, investigation views, decision logging, and account-actioning workflows.
  • Develop reusable APIs, data storage, and backend services for quickly and safely launching new review workflows.
  • Scale review through automation, enabling effective use of AI by reviewers and building towards AI-assisted and AI-driven review processes.
  • Translate enforcement and investigation needs into reliable systems by partnering with policy, operations, legal, privacy, and data science stakeholders.
  • Implement necessary guardrails for sensitive internal tools, such as granular permissions, audit trails, data-access controls, and reviewer wellbeing features.
  • Instrument shipped tools to surface metrics on queue health, reviewer throughput, and decision quality, ensuring evolution alongside privacy primitives and data retention commitments.

Requirements

  • Technical background in full-stack or platform engineering with the ability to engage in architecture and design discussions.
  • Experience shipping internal tools or platforms with demanding operational users and measurably improving their workflows.
  • Experience collaborating cross-functionally with non-engineering partners like operations, policy, or legal teams.
  • Excellent communication skills, including explaining technical tradeoffs to non-technical stakeholders.
  • A passion for the societal impacts of AI and a desire to make powerful systems safer.

About Anthropic

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.