Research Engineer, Model EvaluationsActive$500K

Remote · WorldwideTechnology

The opportunity

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole.

What they're looking for

  • Run experiments to characterize how prompting, sampling, and scaffolding: choices affect results on internal and industry benchmarks
  • Communicate evaluations and their results to internal stakeholders and, where appropriate, external audiences
  • Strong Python programming skills, including production or research infrastructure
  • Experience building or operating distributed systems, data pipelines, or: other infrastructure that needs to be reliable at scale