The opportunity
AI is becoming vitally important in every function of our society. At Scale, our mission is to accelerate the development of AI applications.
What you'll do
Train state of the art models, developed both internally and from the: community, to deploy to our enterprise customers.
Research cutting edge algorithms to integrate directly into our training stack.
Design solutions that enable complex multi-agent systems to directly learn: from both process + outcome based rewards.
+ years of LLM training in a production environment
Experience with post-training methods like RLHF/RLVR and related algorithms like PPO/GRPO etc.
Publications in top conferences such as NEURIPS, ICLR, or ICML within the last two years
What they're looking for
- PhD or Masters in Computer Science or a related field