The opportunity
Multi-Agent systems are becoming an increasingly important part of how AI is deployed, whether via fast small-model subagents inside a product, or large groups of agents solving very large problems . Training Claude to be maximally effective and safe within large groups is a…
What you'll do
Help create and optimize environments and data for model training that: maximize Claude’s performance or ease of use on agentic tasks
Ideate, develop, and compare the performance of different agent harness: configurations (eg memory, context management, communication architectures for agents)
Design and implement rigorous quantitative benchmarks for large scale agentic tasks
Work with our product org to find solutions to our most vexing challenges in applying agents to our products
Have experience with large-scale RL on language models
Have experience training multi-agent systems
What they're looking for
- Enjoy going deeply into the roots of a problem and understanding its foundations, rather than its surface.
- Have good communication skills and an interest in working with other researchers on difficult tasks
- Have a passion for making powerful technology safe and societally beneficial
- Are excited for a mission-driven org with fast-paced, impactful work