The opportunity
We're looking for a research engineer who believes that visual and spatial reasoning are core to fully unlocking the capabilities of LLMs. On the Vision team, you'll own the end-to-end process of creating training data and RL environments targeting visual knowledge work:…
What you'll do
Own the data strategy for vision capabilities end-to-end, from building evals and scaling RL environments
Manage technical relationships with external data vendors, including writing: task specifications, evaluating visual data and annotation quality, and iterating on reward design
Develop and improve QA frameworks that catch reward hacking and ensure environment quality at scale
Run generalization experiments to measure how data strategy changes improve: multimodal capabilities on held-out evaluations
Partner with pretraining, RL, and product teams, and do the science that: shows we’re all rowing in the same direction
Have 7+ years of ML, computer vision, and software engineering experience: through industry, academia, or other projects
What they're looking for
- Have experience with reinforcement learning, reward design, or training data: curation for large language or vision-language models
- Are familiar with the architecture, training, and operation of large vision language models
- Are comfortable managing technical vendor relationships and iterating quickly on feedback
- Are results-oriented, with a bias towards flexibility and impact