How It Works
Version A is the current state, the baseline. Version B contains a single, deliberate change. Both versions run under conditions as similar as possible, and a pre-defined metric decides which wins. The test only produces trustworthy results when the sample is large enough, the metric is agreed in advance, and only one variable changes between A and B.