The opportunity
Our Reinforcement Learning teams are central to advancing our AI systems, contributing to every Claude model and driving the autonomy and coding gains in our latest releases. The work spans computer use, code generation through RL, fundamental RL research for large language…
What you'll do
Deliver a regular read on the ground truth in RL research, covering: performance against baselines, experiment results, day-to-day health, and incidents
Work with RL org leads on prioritizing, ranking, and tracking the state of experiments
Drive research reviews end to end in partnership with set the agenda, make: sure the right context is in the room ahead of time, and close the loop on what gets decided
Establish processes and frameworks that bring structure to an unstructured: research setting without slowing researchers down
Collaborate with research leads, infrastructure engineers, and data: operations to identify blockers, prioritize competing needs, and make technical trade-off decisions