Researcher, Misalignment ResearchActive$295K–$445K

San FranciscoTechnology

The opportunity

About the Team Safety Systems sits at the forefront of OpenAI’s mission to build and deploy safe AGI, ensuring our most capable models can be released responsibly and for the benefit of society. Within Safety Systems, we are building a misalignment research team to focus on the…

What they're looking for

  • Create automated tools and infrastructure to scale automated red‑teaming and stress testing.
  • Conduct research on failure modes of alignment techniques and propose improvements.
  • Publish influential internal or external papers that shift safety strategy or: industry practice. We aim to concretely reduce existential AI risk.
  • Partner with engineering, research, policy, and legal teams to integrate: findings into product safeguards and governance processes.