The opportunity
Preparedness is a critical Safety Research team at OpenAI, which is focused on mitigating AI threats to global security that could scale to an extreme level of severity.
What you'll do
Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems.
Mitigation. Keeping misuse safeguards, alignment tools, and security measures: on track to adequately address extreme threats that might arise in the future.
Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness: framework , and partnering with other staff to achieve these targets.
Design and implement mitigation components for model-enabled cybersecurity: misuse—spanning prevention, monitoring, detection, and enforcement—under the guidance of senior technical and risk leadership.
Integrate safeguards across product surfaces in partnership with product and: engineering teams, helping ensure protections are consistent, low-latency, and scale with usage and new model capabilities.
Evaluate technical trade-offs within the cybersecurity risk domain (coverage,: latency, model utility, and user privacy) and propose pragmatic, testable solutions.
What they're looking for
- Collaborate closely with risk and threat modeling partners to align: mitigation design with anticipated attacker behaviors and high-impact misuse scenarios.
- Execute rigorous testing and red-teaming workflows, helping stress-test the: mitigation stack against evolving threats (e.g., novel exploits, tool-use chains, automated attack workflows) and across different product surfaces—then iterate based on findings.
- Have a passion for AI safety and are motivated to make cutting-edge AI models safer for real-world use.
- Bring demonstrated experience in deep learning and transformer models.