The opportunity
Models are becoming increasingly capable—moving from tools that assist humans to agents that can plan, execute, and adapt in the real world. Mitigating the frontier risks resulting from these capabilities is paramount to OpenAI’s ability to continue deploying models safely.
What you'll do
Measurement. Monitoring and predicting the evolving capabilities of frontier AI systems.
Mitigation. Keeping misuse and misalignment safeguards, alignment tools, and: security measures on track to adequately address extreme threats that might arise in the future.
Coordination. Setting mitigation targets by maintaining OpenAI’s preparedness: framework , and partnering with other staff to achieve these targets.
Design and implement mitigation components for model-enabled cybersecurity: misuse—spanning prevention, monitoring, detection, and enforcement—under the guidance of senior technical and risk leadership.
Integrate safeguards across product surfaces in partnership with product and: engineering teams, helping ensure protections are consistent, low-latency, and scale with usage and new model capabilities.
Evaluate technical trade-offs within the cybersecurity risk domain (coverage,: latency, model utility, and user privacy) and propose pragmatic, testable solutions.
What they're looking for
- Collaborate closely with risk and threat modeling partners to align: mitigation design with anticipated attacker behaviors and high-impact misuse scenarios.
- Execute rigorous testing and red-teaming workflows, helping stress-test the: mitigation stack against evolving threats (e.g., novel exploits, tool-use chains, automated attack workflows) and across different product surfaces—then iterate based on findings.
- Have a passion for AI safety and are motivated to make cutting-edge AI models safer for real-world use.
- Bring demonstrated experience in deep learning and transformer models.