Skip to content
Kudos AI

Experimentation

A 1-part series, ordered from first to latest.

  1. The Experiment That Was Going to Win Anyway

    A test with 2,000 users per arm reports effects 2.4 times too large. An A/A test checked ten times comes out significant 19% of the time. Twenty independent null metrics produce a winner 64% of the time and twelve null segments 46%. Four numbers, one cause, and the decisions that have to be made before the data arrives.

    11 September 2026 · 8 min read