Engineering Manager, Safeguards Review Tooling

San Francisco, CA

Posted 16d ago

Job Location

San Francisco, CA

Tech Stack

Remote Work Policy

On-site

Categories

Applied AI Engineer

About the job

Anthropic is seeking an Engineering Manager to lead their Review Tooling team, which is responsible for building systems that help humans and AI investigate potential harms and take enforcement actions across Anthropic's products. This foundational role involves owning the tools safety investigators rely on, including analytics, privacy-preserving primitives, and a sandbox environment for rapid iteration. The manager will also drive the strategy for scaling review through automation, enabling Claude to extend human reviewer capabilities while keeping people in the loop for critical judgment. This position requires close collaboration with policy, operations, data science, and legal teams to ensure effective, accurate, and trustworthy enforcement systems.

Responsibilities

  • Lead, grow, and develop a team of engineers building investigation, review, and enforcement tooling.
  • Define the vision and roadmap for the review tooling platform, including analytics, privacy-compatible data access, and a sandbox for new interfaces.
  • Drive the team's strategy for scaling review through automation, enabling effective use of Claude and building toward Claude-assisted and Claude-driven review workflows.
  • Partner with policy, operations, legal, privacy, and data science stakeholders to translate needs into reliable systems.
  • Ensure review tooling evolves with privacy primitives and data retention commitments to maintain user trust.
  • Create clarity for the team and stakeholders in an ambiguous environment.
  • Hire and coach top technical talent with an inclusive and equitable approach.
  • Contribute to engineering-wide initiatives as part of the engineering management community.

Requirements

  • Experience managing software engineering teams, including hiring, coaching, and developing engineers.
  • Technical background in full-stack or platform engineering, with ability to engage in architecture and design discussions.
  • Experience shipping internal tools or platforms with demanding operational users and measurably improving their workflows.
  • Experience working cross-functionally with non-engineering partners (operations, policy, legal).
  • Excellent communication skills, including explaining technical tradeoffs to non-technical stakeholders.
  • Care about the societal impacts of AI and a desire to make powerful systems safer.

Benefits

  • Annual compensation range: $405,000 - $485,000 USD
  • Visa sponsorship available

About Anthropic

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.