The Model Policy team within Safety Systems works to ensure that frontier models behave safely and reliably in real-world environments by designing policies that define safe model behavior. The Model Policy Manager will focus on the safety of multimodal models, shaping how OpenAI identifies, evaluates, and addresses risks in multimodal AI models such as GPT-Live and ChatGPT Images, as well as multimodal capabilities in frontier AI models. This role involves designing and maintaining model policies for audio, image, video, and omni-modal behavior, translating theories of harm and threat models into behavioral safety policies, evaluation criteria, grading guidance, and safeguards. The role also includes identifying and analyzing safety regressions and failure patterns to inform policy iteration, and developing policy artifacts to support model training, evaluation, and deployment. The position requires partnering with AI researchers, domain experts, and product teams to operationalize policy into measurable model behavior. The role is based in San Francisco, CA, with a hybrid work model of 3 days in the office per week.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
No Education Listed