Model PolicyActive$207K–$295K

Hybrid · San FranciscoOperations

The opportunity

About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency. Within Safety Systems, the Model Policy team aligns model behavior with desired human values and norms.

What they're looking for

  • Study real-world deployments to identify where model behavior succeeds,: fails, or drifts from the intended safety posture.
  • Combine longer-horizon safety research with hands-on launch and deployment work.
  • Contribute to system cards, safety reports, policy documentation, launch: reviews, and external communications on OpenAI's approach to model safety and risk mitigation.
  • Design and run human data campaigns, including gold set construction,: labeling guidance, calibration, adjudication, and eval coverage analysis, to ensure policies can be reliably measured and improved.
Model Policy at OpenAI | Role Match