Underpowered Design

Suppose a valid power grid contains the following rows for a 28-period test:

effect_size power Monte Carlo interval valid
0.04 0.31 0.27–0.35 true
0.06 0.55 0.50–0.60 true
0.08 0.73 0.68–0.77 true
0.10 0.84 0.80–0.88 true

At an 80% target, the grid-based MDE is 10%. This does not mean that the campaign is expected to deliver 10%, or that effects below 10% are zero. It means the configured procedure detected smaller injected effects less often than the planning threshold under this DGP.

If commercially plausible lift is 4%, the design is underpowered for that decision. Do not solve the problem by reporting an optimistic grid point or by choosing a longer period after seeing outcomes. Consider more comparable donors, lower-noise outcomes, a defensible larger treated population, a longer pre-specified measurement window, or a different design. If none is available, record the design as infeasible before launch.

Also inspect failures. A row with high estimated power and valid: false is not a usable design result. Power is divided by successful simulations, so ignored failures could otherwise make it look better than it is.