1 open role mentioning evaluation harnesses
Scale AI
Scale Labs is seeking a Research Scientist focused on Frontier Risk Evaluations to join a new team dedicated to policy research. This role bridges the gap between AI research and global policymakers, aiming to inform scientific decisions about AI risks and capabilities. The team tackles complex problems in agent robustness, AI control protocols, and AI risk evaluations to help governments, industry, and the public understand and mitigate AI risks while promoting AI adoption. This position involves designing and creating evaluation measures, harnesses, and datasets for assessing the risks posed by frontier AI systems, with opportunities to collaborate broadly and publish findings.
$216k - $270k