Model Policy Manager, Multimodal SafetyNew$266K–$335K

The opportunity

Safety Systems manages the complete lifecycle of safety efforts for OpenAI’s frontier models, ensuring our models are deployed responsibly and have a positive impact on society. Our work spans diverse research and engineering initiatives—from system-level safeguards and model…

What you'll do

  • Safety at every step

  • OpenAI GPT6 System Card

  • OpenAI Model Spec

  • ChatGPT Images 2.5

  • Design and maintain model policies for audio, image, video, and omni-modal behavior.

  • Translate theories of harm and threat models into behavioral safety policies,: evaluation criteria, grading guidance, and safeguards.

What they're looking for

  • Identify and analyze safety regressions and failure patterns to identify gaps: in existing policies and inform policy iteration.
  • Develop policy artifacts that support model training, evaluation, and: deployment, including behavior instructions, human-data campaigns, golden sets, and evaluations.
  • Partner with AI researchers, domain experts, and product teams to: operationalize policy into measurable model behavior.
  • Have strong judgment about the real-world risks of advanced multimodal AI systems.