Researcher, Synthetic RLActive$295K–$445K

The opportunity

The Synthetic RL team develops reinforcement learning methods that leverage synthetic data, environments, and feedback to train and evaluate frontier AI models. The team explores approaches such as self-play, simulators, and other synthetic evaluations to push model capability,…

What you'll do

  • Research and develop reinforcement learning algorithms

  • Design and run experiments to study training dynamics and model behavior at scale

  • Collaborate with engineers and researchers to integrate successful approaches into model training pipelines

  • Have a strong background in reinforcement learning, machine learning research, or related fields

  • Have strong engineering and statistical analysis skills

  • Enjoy exploring new problem spaces where data, objectives, and evaluation are imperfect or evolving

What they're looking for

  • Are motivated by seeing research ideas influence real-world AI systems