1 open role mentioning probing
Scale AI
Scale Labs is seeking talented researchers to join a new team focused on policy research, bridging the gap between AI research and global policymakers to make informed, scientific decisions about AI risks and capabilities. This role will develop and apply post-training methods and interpretability techniques to make frontier AI systems safer and better understood. You will design and run post-training pipelines, develop interpretability-informed evaluations, and collaborate with policymakers, engineers, and other researchers to translate findings into actionable safety standards and best practices.
$216k - $270k