The uncomfortable base rate
Start with the number that built the experimentation industry. Ron Kohavi, who ran experimentation platforms at Microsoft and Airbnb and co-wrote the standard text, Trustworthy Online Controlled Experiments, has reported repeatedly that at Microsoft only about a third of tested ideas improved the metrics they were designed to improve; roughly a third made no measurable difference, and a third actively hurt. Teams at other heavily-experimenting companies have published similar or harsher ratios, with well-designed, expert-reviewed ideas failing at rates that embarrass everyone's intuition.
Sit with what that means: the median confident product idea, backed by expertise, design review and conviction, does nothing or does damage. Nobody can tell which third an idea belongs to in advance, including the experts proposing it.
Key idea: experimentation is not a statistics hobby; it is the institutional admission that intuition about what helps users is wrong most of the time, made survivable by a machine that checks every idea cheaply. Companies like Booking.com and Airbnb run their product process on that admission, testing essentially everything user-facing.
The rest of this course is that machine: what it is, how it breaks, and the engineering that keeps its answers honest at scale.

