The Safety Research team at OpenAI is focused on mitigating AI threats to global security that could scale to an extreme level of severity. This involves measurement (monitoring and predicting AI capabilities), mitigation (ensuring safeguards are adequate), and coordination (setting mitigation targets). This is urgent, fast-paced work with far-reaching implications. This role involves pushing the boundaries of frontier models to shape our empirical grasp of AI safety concerns. You will own the scientific validity of frontier preparedness capability evaluations, designing new evaluations grounded in real threat models (including CBRN, cyber, and other frontier-risk areas) and maintaining existing ones. You will define datasets, graders, rubrics, and threshold guidance, producing auditable artifacts for leadership. Key responsibilities include identifying emerging AI safety risks and new methodologies, building and refining evaluations of frontier AI models, designing and building scalable systems for evaluations, and contributing to risk management and best practice guidelines for AI safety evaluations.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed