Researcher, Alignment CoT MonitorabilityNew$250K–$445K

Hybrid · San FranciscoOperations

The opportunity

About the Team The CoT Monitorability team at OpenAI studies whether and when the chain-of-thought of frontier reasoning models is monitorable enough to support scalable oversight. We study how to measure monitorability , which training mechanisms affect monitorability, and…

What they're looking for

  • Produce externally publishable research when results advance the broader science of alignment.
  • Have strong hands-on experience training, evaluating, or debugging large ML models, especially LLMs.
  • Have deep curiosity, interest in alignment, and high agency.
  • Bring depth in alignment, interpretability, model behavior, empirical ML, or adjacent research.