The opportunity
The Synthetic RL team develops reinforcement learning methods that leverage synthetic data, environments, and feedback to train and evaluate frontier AI models. The team explores approaches such as self-play, simulators, and other synthetic evaluations to push model capability,…
What you'll do
Research and develop reinforcement learning algorithms
Design and run experiments to study training dynamics and model behavior at scale
Collaborate with engineers and researchers to integrate successful approaches into model training pipelines
Have a strong background in reinforcement learning, machine learning research, or related fields
Have strong engineering and statistical analysis skills
Enjoy exploring new problem spaces where data, objectives, and evaluation are imperfect or evolving
What they're looking for
- Are motivated by seeing research ideas influence real-world AI systems