The opportunity
We're hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to design silicon.
What you'll do
Developing systems that enable models to use computers effectively
Advancing code generation through reinforcement learning
Pioneering fundamental RL research for large language models
Building scalable RL infrastructure and training methodologies
Enhancing model reasoning capabilities
Invent, design, and implement RL environments and evaluations for agentic RTL: generation, design (including formal) verification, physical design optimization.
What they're looking for
- Work on cross-cutting RL considerations such as EDA-tool latency optimization and proxy rewards.
- Conduct experiments and shape our roadmap.
- Deliver your work into research and production training runs.
- Collaborate with other researchers and engineers across and outside Anthropic.