The CTA button best practices that hold up under testing come down to six things: write the copy around what the visitor receives rather than what they have to do, use first-person framing, keep it to two to five specific words, make the button the highest-contrast element in its immediate context, repeat it inline after each persuasive section, and test copy before you test anything visual.
Notice what is not on that list. Button color, the thing people ask about first, matters only as a proxy for contrast. The famous red-versus-green tests were measuring how much the button stood out from its page, not anything intrinsic to the color.
The newsletter
Join our KISS newsletter
One short read a week on what actually moves revenue, in a free email. Read by 10,000+ operators and founders.
No spam. Unsubscribe in one click.
I.The label is the lever, the visual is a floor
Copy and appearance are not two variables of the same kind. One changes what the visitor believes they are agreeing to; the other only decides whether they saw the button at all.
A.Why the label carries the decision
The button label is the last thing a visitor reads before committing, and it is the only element on the page that describes the transaction they are about to enter. Everything above it argued for the product. The label answers a narrower question: what happens in the next three seconds, and what do I get for it. That is why the wording moves behaviour in a way no visual property can. Colour cannot reduce uncertainty about what a click does.
Three properties of the wording do most of the work. The first is what the verb describes. “Submit,” “register” and “sign up” name the effort the visitor expends. “Get,” “start,” “claim,” “see” and “access” name what arrives. The same action framed as receipt rather than expenditure reads as a smaller ask.
The second is specificity, which is really a way of pricing the risk. “Download Now” leaves the visitor to imagine what follows the click, and people imagine badly: a form, a sales call, a mailing list. “Download the 50-Page Guide” closes that gap. Under two words is too vague to mean anything, over five stops reading as a button, and inside that range every word should either be the action or be the value.
The third is grammatical person, which is the best-documented finding in the area. Michael Aagaard’s Unbounce test pitted “Start your free 30 day trial” against “Start my free 30 day trial,” and the first-person version won by a wide margin. The mechanism is ownership: “my” has the visitor completing the action in their own voice, while “your” keeps it as something being asked of them. Later replications land smaller. The direction rarely changes.
Reported lift from two published CTA copy tests
A/B Test viewB.What the visual actually buys
None of that means appearance is free to ignore, because the visual properties are not competing with copy. They are a threshold that has to be met before copy can be read at all, and once met they stop paying. A red button on a mostly blue page beats a blue one; the same red button on a red page loses to it. Contrast, not colour, is the variable, and the requirement is simply that the CTA be the most prominent element in its immediate context.
Whitespace is the one people underrate, and it is doing two separate jobs. Visually, isolation reads as importance: a button surrounded by space is a decision point, a button wedged between paragraphs is page furniture. Mechanically, space around a tap target reduces mis-taps on mobile, which is a conversion loss that never appears in any report because a mis-tap looks like a visitor who chose not to click. That is the shape of everything in this sub-part. The floors do not create wins. They remove invisible losses, which is why they are worth meeting once and then never testing again.
II.The button is the wrong unit
Everything in the first part treats the CTA as an element on a page. It is better understood as a moment in a decision, and that reframing changes both where it goes and what you measure.
A.Placement is a claim about how much persuasion is needed
The above-the-fold argument is usually conducted as if there were a right answer. There is not, because placement encodes an assumption about the visitor’s state. Above-the-fold works when intent is already formed: someone who clicked an ad for free project management software knows what they want, and any delay is friction. When the decision needs education, at high price points, for complex services, for a brand the visitor has never heard of, that same button arrives before there is any reason to say yes, and its real function becomes navigational. It tells the visitor what the page is for.
Inline CTAs solve the other case. Placed after a persuasive section, a proof point or a testimonial, they catch the visitor at the moment the argument landed. On a long page, repeat them every two or three persuasive sections and keep every instance identical, because two differently worded buttons make the visitor choose between offers rather than accept one. Sticky bars are a hedge for mobile, where the ready moment and the scroll position are uncorrelated. Keep them thin, and treat them as a supplement to inline placement rather than a replacement.
The same logic extends past position into wording, once you know anything about who is reading. A first-time visitor and a returning visitor who has read three product pages are at different points in the same decision, and the button that fits one wastes the other.
Matching the ask to the state the visitor is in
Populations view| Visitor state | The ask that fits | The ask that wastes the slot |
|---|---|---|
| First visit, arrived from search | See how it works | Choose your plan |
| Read two or more product pages | Start free trial | See how it works |
| Returned to pricing twice | Choose your plan | Download the guide |
| Trial started, no key action yet | Finish setup | Start free trial |
Page context is the cheap version of the same idea and requires no behavioural data at all. A post about conversion optimisation should offer something about conversion optimisation. A comparison page should acknowledge that the reader is running an evaluation. A generic sign-up button on a specific page is the mismatch above, committed on purpose.
B.Which forces the measurement downstream
If a CTA is a moment inside a decision rather than an element on a page, then the click is not the outcome. It is one step, and optimising it in isolation has a predictable failure. Click rate rewards whatever makes clicking feel cheapest, and the cheapest promise is usually the one furthest from the truth. A button that overstates what follows lifts the click and thins every step after it, and the report shows a win.
From CTA impression to paid, one variant
Funnels report viewA Funnels report that begins at the button click and ends at the paid conversion is the only view in which the overpromise shows up as what it is. It also changes which tests you can read at all: click rate reaches significance quickly and revenue does not, so a variant that looks decided at the click step is frequently undecided where it counts. That is an argument for fewer, longer tests rather than for measuring the easier thing.
III.Running it so the results accumulate
One good test changes one page. What compounds is a sequence run in an order that makes each result readable, and written down so the next person inherits it.
A.The order, and why it is that order
Copy first, placement second, visual design third. The ordering is not about which matters most in the abstract; it is about which produces a readable result soonest. Copy changes are free to ship, produce the largest documented effects, and are unambiguous when they win. Placement changes are cheap but interact with page length and traffic mix, so they need more care in reading. Visual changes are the smallest effects on the least traffic-efficient variable, which means they are the tests most likely to run for six weeks and conclude nothing.
Sample size is the constraint nobody plans for. The number of visitors needed rises sharply as the effect you are looking for gets smaller, which is exactly the situation you create by testing a button radius before you have tested a verb. Run the tests in descending order of expected effect and you spend your traffic on the questions that can actually be answered with it. Our guide to statistical significance covers how to size one honestly, and the testing primer covers the setup.
Two reading errors are worth naming because they both produce confident wrong answers. The first is calling a test at the moment it crosses significance, which is peeking: a metric wandering across a threshold will cross it eventually whether or not anything changed, so the stopping rule has to be set before the test starts. The second is mixing traffic. A variant that wins on paid social and loses on branded search can report as flat, and the flat result is the one that gets filed. Split the read by source wherever you have the volume for it, and if you do not have the volume, at least check that the mix did not shift mid-test.
One test at a time on one button, too. Simultaneous changes to wording and placement produce a result you cannot attribute, and an unattributable win teaches you nothing transferable, which defeats the point of running the programme at all.
B.What makes a result outlive the test
Most CTA testing evaporates. The variant ships, the ticket closes, and eighteen months later someone runs the same test because the result lived in a spreadsheet nobody inherited. What accumulates instead is a written record of patterns that turned out to be true for your audience specifically: that first-person framing wins here, that urgency underperforms with this segment, that the understated button beats the loud one on the pricing page. Those patterns are worth more than any individual lift, because they are the priors you bring to every new page.
Urgency is the clearest example of a pattern you have to establish locally. “Only 3 spots left” works when it is true and is discounted automatically by visitors who have seen countdown timers reset on refresh. Value framing has no such decay, applies whether or not a deadline exists, and does not spend trust you will need later. It is the better default, and manufactured scarcity is the one tactic worth declining even when it tests positive, because the cost lands outside the window the test measures.
The mechanical blocker on all of this is usually instrumentation. If measuring a new variant requires an engineer to add a tracking call, the programme runs at the speed of the sprint board and most variants never get measured past the click. Autocapture removes the dependency by recording interactions with elements that already exist, so the funnel from button to payment is available for a variant the day it ships. For the page-level context around the button itself, see our landing page conversion guide and form optimisation strategies.
Verdict
Colour is the wrong question and always was. Treat the visual properties as a specification to satisfy once, not a variable to explore: contrast against the immediate surroundings, at least 44 pixels of tap height, real space on every side. Meet that, then stop. Every hour spent on button aesthetics after the floor is met is an hour not spent on the label, and the label is where the documented effects are, because it is the only part of the button that tells the visitor what they are agreeing to.
Then stop optimising the button and start optimising the moment. Write the label as what the visitor receives, in their own voice, in two to five specific words. Put it where the argument for clicking has just been made, repeat it identically, and match the ask to what you already know about the reader. And measure every variant to the paid conversion rather than the click, because the single most reliable way to raise a click rate is to overstate what follows it, and a testing programme that stops at the click will find that trick and recommend it. If you can only enforce one rule from this article, enforce that one.
One analytics idea a week
Short, specific, written by the team building the product. No digest, no roundup.
Continue Reading
What Is a Landing Page Conversion? Definition, Formula, and 2026 Benchmarks
A landing page conversion is the percentage of visitors who complete the page target action. The formula is simple, but most teams measure it wrong because session-based analytics fragments the data. Here is what the metric actually means, how to calculate it, and what good looks like in 2026.
Read articleIntroduction to A/B Testing: How to Run Experiments That Actually Work
A/B testing is the most reliable way to improve conversion rates. But most tests fail because of poor methodology, not poor ideas. This guide shows you how to run tests that produce trustworthy results.
Read articleExit Intent Strategies: Recovering Visitors Before They Leave
Exit intent overlays can recover 10-15% of abandoning visitors when done right. But most are done wrong: too aggressive, poorly targeted, and never measured. Here is the analytics-driven approach.
Read article