The opportunity
We're hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to safely write correct, fast code for accelerators.
What you'll do
Developing systems that enable models to use computers effectively
Advancing code generation through reinforcement learning
Pioneering fundamental RL research for large language models
Building scalable RL infrastructure and training methodologies
Enhancing model reasoning capabilities
Invent, design and implement RL environments and evaluations.
What they're looking for
- Conduct experiments and shape our research roadmap.
- Deliver your work into training runs.
- Collaborate with other researchers, engineers, and performance engineering: specialists across and outside Anthropic.
- Have expertise with accelerators (CUDA, ROCm, Triton, Pallas), ML framework programming (JAX or PyTorch).