The opportunity
We are looking for research engineers to build the safety and oversight mechanisms that govern how our models handle biological knowledge. As a bio safety researcher, you will spend your time: designing and running capability evaluations against frontier models, generating and…
What you'll do
Design, build, and run capability evaluations to assess what new models can: do in the biological domain, and turn results into concrete deployment recommendations
Develop training and evaluation datasets for our safety classifiers, working: with internal and external threat modeling experts to ground them in realistic risk
Train, tune, and iterate on safety classifiers with ML engineers, optimizing: jointly for adversarial robustness and low false-positive rates
Build the tooling and pipelines that make evaluation and classifier development fast and repeatable
Analyze classifier and eval performance against production traffic, identify gaps, and prioritize improvements
Design and run red-teaming and stress-testing of safeguards as threats, models, and product surfaces evolve
What they're looking for
- Experience working with large language models, including prompting, fine-tuning, or evaluation
- Experience training or deploying classifiers or other ML systems in production
- Experience developing ML methods for biological systems or biological data
- Familiarity with adversarial robustness, red-teaming, or safety evaluation of ML systems
- Have at least 8 years of hands-on experience in life sciences, with deep: expertise in areas such as molecular biology, drug discovery, or computational biology
- Experience leading complex technical projects across multiple stakeholder groups