The opportunity
The Statsig team at OpenAI builds and operates the experimentation platform that powers product development, measurement, and decision-making across the company. We partner closely with product, engineering, and infrastructure teams to ensure experiments are trustworthy,…
What you'll do
Drive the statistical direction and technical strategy for OpenAI’s experimentation platform
Design and improve experimentation methodologies used across product and research teams
Build pragmatic solutions to real-world experimentation challenges, balancing: rigor with operational simplicity
Improve the reliability and trustworthiness of experiment results, including: detection and prevention of bias, logging issues, and data quality failures
Developscalable analytical systems and pipelines in Python and distributed compute environments
Partner with engineers and product teams to improve experiment design, metric: quality, and decision-making practices
What they're looking for
- Lead investigations into complex experimentation anomalies and measurement failures
- Establish best practices for experimentation governance, interpretation, and statistical correctness
- Mentor other data scientists and raising the overall technical bar for experimentation and causal inference
- Experience building, scaling, or operating experimentation platforms at a large technology company