Engineering Manager, Safeguards

London, UK

Posted 12d ago

Job Location

London, UK

Tech Stack

Remote Work Policy

On-site

Categories

Applied AI Engineer

About the job

Anthropic is building reliable, interpretable, and steerable AI systems to be safe and beneficial for users and society. The Safeguards team is crucial in ensuring our models and products are developed and deployed safely. We are seeking an Engineering Manager to lead the Review Tooling team. This team develops the systems that both humans and AI, like Claude, use to investigate potential harms and take enforcement actions across Anthropic's products and third-party platforms. This is a foundational role where you will own the tools that safety investigators rely on, including analytics, privacy-preserving primitives, and a sandbox for rapid iteration. You will also drive the scaling of review through automation, integrating Claude to extend human reviewer capabilities while ensuring human judgment remains central where needed. Close collaboration with policy, operations, data science, and legal teams is essential to ensure effective, accurate, and trustworthy enforcement systems.

Responsibilities

  • Lead, grow, and develop a team of engineers building investigation, review, and enforcement tooling.
  • Define the vision and roadmap for the review tooling platform, including analytics, privacy-compatible data access, and a sandbox for new interface development.
  • Drive the team's strategy for scaling review through automation, enabling effective use of Claude and building Claude-assisted/driven workflows.
  • Partner with policy, operations, legal, privacy, and data science stakeholders to translate needs into reliable systems.
  • Ensure review tooling evolves with privacy primitives and data retention commitments to maintain user trust.
  • Create clarity for the team and stakeholders in an ambiguous and evolving environment.
  • Hire and coach top technical talent with an inclusive and equitable approach.
  • Contribute to engineering-wide initiatives as part of Anthropic's engineering management community.

Requirements

  • Experience managing software engineering teams, including hiring, coaching, and developing engineers.
  • Technical background in full-stack or platform engineering, with ability to engage in architecture and design discussions.
  • Experience shipping internal tools or platforms with demanding operational users and measurably improving their workflows.
  • Experience working cross-functionally with non-engineering partners (operations, policy, legal).
  • Excellent communication skills, including explaining technical tradeoffs to non-technical stakeholders.
  • Care about the societal impacts of AI and a desire to make powerful systems safer.

Benefits

  • Annual compensation range: £325,000 — £390,000 GBP
  • Visa sponsorship available

About Anthropic

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.