A/B Testing

A/B testing is a controlled experiment that compares two versions of a web page, email, ad, or feature by randomly splitting traffic between them and measuring which version performs better on a defined success metric.

Also known as: split testing, bucket testing, randomized controlled trial

Why A/B testing matters

A/B testing replaces opinions with evidence. Instead of debating whether a green button or blue button will convert better, you show each to a random half of your audience and let the data decide. This rigorous approach to optimization eliminates the HiPPO effect (Highest Paid Person's Opinion) and builds a culture of data-driven decision making.

The cumulative effect of consistent A/B testing is transformative. A 5% improvement from one test may seem modest, but running 20 tests per year with a 30% win rate and 5% average lift compounds to a 34% total improvement. Companies that build strong experimentation programs consistently outperform competitors who rely on gut feel.

A/B testing also reduces risk. Instead of redesigning your entire checkout flow and hoping for the best, you can test each change individually, measure its impact, and only keep changes that improve performance. This iterative approach prevents costly mistakes and builds confidence in every change you ship.

The newsletter

Join our KISS newsletter

One short read a week on what actually moves revenue, in a free email. Read by 10,000+ operators and founders.

No spam. Unsubscribe in one click.

A/B Testing examples

E-commerce

A furniture retailer A/B tests their product page layout, moving customer reviews from a tab to an inline section visible without clicking. The inline variant increases add-to-cart rate by 12% and shows no negative impact on page load speed.

SaaS

A SaaS company tests two pricing page structures: a three-tier plan comparison vs a single recommended plan with an option to see alternatives. The single-plan variant increases signups by 18% and reduces time-to-decision by 40%.

How to Track in KISSmetrics

Use KISSmetrics alongside your A/B testing tool to get deeper insights into test results. While testing tools measure aggregate conversion rates, KISSmetrics tracks how each variant affects individual user behavior over time. This lets you see whether a variant that wins on immediate conversion also wins on retention, lifetime value, and downstream engagement.

Common Mistakes

  • -Ending tests too early based on initial results that have not reached statistical significance.
  • -Testing trivial changes (button color, font size) while ignoring high-impact elements like value proposition, pricing, and page structure.
  • -Not defining a primary success metric before the test starts, leading to cherry-picking the metric that shows the desired result.
  • -Running multiple tests on the same page simultaneously without controlling for interaction effects.
  • -Ignoring segment-level results - a test that shows no overall effect may have significant positive effects for one segment and negative for another.

Pro Tips

  • +Calculate required sample size before starting a test to know how long it needs to run for reliable results.
  • +Test big, bold changes first (different value propositions, layouts, offers) before optimizing details.
  • +Use KISSmetrics to track the long-term impact of winning variants on retention and revenue, not just the immediate conversion metric.
  • +Build a test backlog prioritized by potential impact (traffic volume times expected improvement) to focus on the highest-value experiments.
  • +Document every test result - wins, losses, and inconclusive - to build institutional knowledge about what works for your audience.

Related Terms

Further Reading

KISSmetrics

Build your business intelligence layer for free.

KISSmetrics records the variant as a property on the person, so every other report splits by variant without extra setup.