What the 5 percent actually promises
The standard experiment closes with a significance test: if the observed difference would arise by chance less than 5 percent of the time under no true effect, declare significance.
The fine print matters more than the ritual. That 5 percent false positive guarantee is a property of a procedure, not of a number, and the procedure it belongs to is: choose a sample size in advance, collect exactly that much data, test once. Every piece of the guarantee is conditional on the once.
Why once? Because the p-value's meaning is how often would chance produce this, under a defined way of looking. Look differently, look repeatedly, stop when you like what you see, and you are running a different procedure whose false positive rate is something else entirely, usually something much worse, while still stamping 5 percent on the report.
Key idea: statistical guarantees attach to the whole decision procedure, including when you look and why you stopped. Change the looking and you change the guarantee, even though every individual calculation along the way was performed correctly. That is what makes the trap in this lesson so quietly destructive: no step looks wrong.

