The opportunity
The Recursive Self-Improvement (RSI) team works across research, engineering, product, and infrastructure to build AI systems that accelerate and ultimately conduct high-quality research at OpenAI. We work to automate real research workflows and improve research productivity by…
What you'll do
Design evaluations for research judgment, hypothesis generation and testing,: and long-horizon experiment execution.
Turn real research workflows and model failures into data and evaluation flywheels.
Improve model research capabilities through agent harnesses, synthetic data,: RL environments, and model training.
Build and maintain safe, reliable integrations between our models and OpenAI’s research infrastructure.
Develop research agents, experiment-orchestration systems, and sandboxed: runtimes that support real research workflows.
Create metrics and economic models to understand RSI’s current and future: effects on research productivity, model capabilities, and the safety of internal deployments.
What they're looking for
- Have research or engineering experience across LLM training, model: evaluations, agent systems, synthetic data, research infrastructure, or large-scale distributed systems.
- Are a strong generalist who can move between open-ended research and: practical implementation, turning ambiguous problems into clear results.
- Collaborate effectively across the full stack, including systems, data, model: training, evaluations, and other research teams.
- Are comfortable building and maintaining the data pipelines, tooling, and: infrastructure needed to support emerging AI capabilities.