1 open role mentioning empirical frameworks
OpenAI
This role focuses on understanding and mitigating real-world risks arising from model misalignment as AI models become more autonomous and operate over longer timeframes. You will investigate how misaligned behavior emerges across extended interactions, such as models pursuing incorrect objectives, taking unsafe shortcuts, or circumventing constraints. The insights gained will be translated into behavioral policies, evaluations, monitoring systems, and safeguards to improve and validate model behavior. This position is ideal for individuals passionate about transforming alignment and safety concerns into concrete, empirically validated improvements for advanced AI systems.