Statistical power
On this page
Definition
Power is the probability of rejecting the null hypothesis when the true effect equals a specified size. It is written , where is the Type II error rate. Power of 0.8 means an effect of the required size will be found in 80% of repetitions; in the rest the test misses it and returns a non-significant result.
How to compute
Power rises with sample size, effect size, and significance level, and falls with variance. For a two-sample mean, approximately: per arm. Take the variance from historical data and the effect from the MDE the product deems meaningful.
Pitfalls
Compute power before the experiment: post-hoc power on an observed effect is meaningless and almost always low. Power does not describe the effect you found; it describes the sensitivity of the design. Several metrics in one test lower the power of each, and CUPED restores it by shrinking .