As asked
An experiment shows the variant beat control on clicks but the result is not statistically significant. The team wants to ship it. What do you say?
Sample answer outline
Separate the observed lift from the evidence for it: a positive point estimate that is not significant means the data is consistent with no real effect. Check sample size and whether the test reached its planned power, look at the guardrail metrics, and ask whether the team peeked early. Recommend either running longer to reach the planned sample or shipping with eyes open if the downside is low, but be clear about the risk.
Expect these follow-ups
- How would you explain not significant without saying it failed?
- What guardrail metrics would you insist on checking first?
- How does peeking inflate the false positive rate?