Research Engineer, Performance RL (Reinforcement Learning)Active$350K

The opportunity

We're hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to safely write correct, fast code for accelerators.

What you'll do

  • Developing systems that enable models to use computers effectively

  • Advancing code generation through reinforcement learning

  • Pioneering fundamental RL research for large language models

  • Building scalable RL infrastructure and training methodologies

  • Enhancing model reasoning capabilities

  • Invent, design and implement RL environments and evaluations.

What they're looking for

  • Conduct experiments and shape our research roadmap.
  • Deliver your work into training runs.
  • Collaborate with other researchers, engineers, and performance engineering: specialists across and outside Anthropic.
  • Have expertise with accelerators (CUDA, ROCm, Triton, Pallas), ML framework programming (JAX or PyTorch).
Research Engineer, Performance RL (Reinforcement Learning) at Anthropic | Role Match