The opportunity
About the Team Our Safety Systems team is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency. Within Safety Systems, the Model Policy team aligns model behavior with desired human values and norms.
What they're looking for
- Study real-world deployments to identify where model behavior succeeds,: fails, or drifts from the intended safety posture.
- Combine longer-horizon safety research with hands-on launch and deployment work.
- Contribute to system cards, safety reports, policy documentation, launch: reviews, and external communications on OpenAI's approach to model safety and risk mitigation.
- Design and run human data campaigns, including gold set construction,: labeling guidance, calibration, adjudication, and eval coverage analysis, to ensure policies can be reliably measured and improved.