Staff+ Fullstack Software Engineer, Safeguards Engineering

$405k - $485k • San Francisco, CA | New York City, NY

Posted 8h ago

Remote Work Policy

On-site

Categories

Applied AI Engineer

About the job

Anthropic is seeking fullstack software engineers to join its Safeguards team. This role focuses on building and enhancing safety and oversight mechanisms for AI systems, including monitoring models, preventing misuse, and ensuring user well-being. You will develop systems to detect unwanted model behaviors and enforce terms of service and acceptable use policies, applying your technical skills to uphold principles of safety, transparency, and oversight. Team placement is flexible and determined after the interview process based on your interests, experience, and organizational needs.

Responsibilities

  • Design and build internal review tools for analysts to investigate abuse, including dashboards, case queues, and evidence viewers.
  • Develop interfaces for human supervision of AI agents, such as streaming output, interruption/approval flows, and action provenance tracking.
  • Create visualizations and exploration tools to surface abuse patterns to research teams.
  • Own the frontend architecture for tools handling sensitive data at scale, focusing on performance, correctness, and end-to-end access controls.

Requirements

  • Bachelor’s degree in Computer Science, Software Engineering, or comparable experience.
  • Proficiency in TypeScript and modern React.
  • Comfort working with Python for backend development.
  • Ability to work across the full stack.
  • Strong communication skills, with the ability to explain complex technical concepts to non-technical stakeholders.

Benefits

  • Annual compensation range: $405,000 - $485,000 USD
  • Visa sponsorship available

About Anthropic

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.