Applied Risk Standards Specialist

Remote San Francisco FullTime

Posted 11h ago

Job Location

San Francisco

Tech Stack

Remote Work Policy

Fully remote

Employment Type

FullTime

Categories

Applied AI Engineer

About the job

OpenAI is seeking an Applied Risk Standards Specialist to lead the development of standards for applied AI risk in various applications, including mental health, youth safety, age assurance, functional efficacy, reliability, and privacy. This role involves translating internal research, safety policies, and evaluation methods into technically sound, measurable, and adaptable standards to foster public trust and support responsible AI deployment. You will collaborate closely with internal teams such as Safety Systems, Research, Product, Model Policy, and Legal, as well as external stakeholders and standards bodies, to author proposals, negotiate requirements, and represent OpenAI. The position also includes contributing to frontier AI assurance, risk-management standards, and evaluation approaches, while ensuring emerging requirements are integrated into internal planning.

Responsibilities

  • Lead the development of applied AI risk standards for areas like mental health, youth safety, age assurance, efficacy, reliability, and privacy, focusing on evaluations and assessments.
  • Translate research, internal policies, and evaluation methods into credible test methods, assessment criteria, and standards grounded in real system behavior.
  • Integrate emerging requirements into internal planning and clarify their implications for evaluations, controls, and assurance.
  • Collaborate with technical and GRC partners on assessor competence, independence, conflicts of interest, evidence access, reporting, and the distinction between audits and technical evaluations.
  • Contribute to drafting, representation, and coordination on frontier risk-management and independent-assessment standards.
  • Lead standards-body workstreams, negotiate proposals, and coordinate internal input to reduce the burden on technical experts.
  • Maintain clear approvals, negotiating positions, contribution records, and implementation handoffs to ensure external commitments are connected to technical owners.

Requirements

  • Technical depth in evaluation science, AI evaluations, red teaming, AI safety, or trust and safety.
  • Demonstrated ability to translate technical expertise into credible standards or assessment methods.
  • Experience authoring or materially shaping standards, evaluation criteria, control frameworks, or normative technical contributions.
  • Ability to translate broad safety objectives into clear requirements and identify evidence to demonstrate their fulfillment.
  • Understanding of the differences among organizational risk-management processes, model and system evaluations, mitigation testing, and broader assurance claims.
  • Ability to evaluate sensitive, context-dependent outcomes, including uncertainty and variation, without overstating assessment capabilities.
  • Understanding of how assessor competence, independence, incentives, and access affect assessment credibility.
  • Skill in building consensus without sacrificing technical rigor, and knowing when to compromise, challenge, or escalate decisions.
  • Ability to earn trust with technical teams while communicating clearly with legal, executive, policy, and external participants.
  • Precise writing skills, discretion, and the ability to independently manage workstreams from proposal through handoff.

About OpenAI

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.