Researcher, Safety OversightActive$295K–$445K

The opportunity

The Safety Systems team is responsible for various safety work to ensure our best models can be safely deployed to the real world to benefit the society, and is at the forefront of OpenAI's mission to build and deploy safe AGI, driving our commitment to AI safety and fostering a culture of trust and transparency.

What you'll do

  • Develop and refine AI monitor models to detect and mitigate known and: emerging patterns of misuse and misalignment.

  • Set research directions and strategies to make our AI systems safer, more aligned, and more robust.

  • Evaluate and design effective red-teaming pipelines to examine the end-to-end: robustness of our safety systems, and identify areas for future improvement.

  • Conduct research to improve models’ ability to reason about questions of: human values, and apply these improved models to practical safety challenges.

  • Coordinate and collaborate with cross-functional teams, including T&S, legal,: policy and other research teams, to ensure that our products meet the highest safety standards.

  • Are excited about OpenAI’s mission of building safe, universally beneficial: AGI and are aligned with OpenAI’s charter

What they're looking for

  • Show enthusiasm for AI safety and dedication to enhancing the safety of: cutting-edge AI models for real-world use.
  • Bring 4+ years of experience in the field of AI safety, especially in areas: like RLHF, human-AI collaboration, fairness & biases.
  • Hold a Ph.D. or other degree in computer science, machine learning, or a related field.
  • Thrive in environments involving large-scale AI systems.