Mida's agent reads your own results and tells you what is worth testing next.

Try it for FREE now

Do social proof A/B tests work?

Least often of anything we track — but the wins are among the biggest. Social proof changes — reviews, ratings, logos, counts — win 9% of the time, near the bottom of the table. 79% change nothing measurable. When one does land the median moves +25.8%, which is a large effect earned rarely.

The widest gap in this benchmark between how often something is recommended and how often it works. Add social proof because it is true and useful, not because you expect a lift.

The numbers

Tests analysed100+
Beat control9%
No measurable difference79%
Lost to control12%
Median lift when it won +25.8%

Typical winning range: +14.5% to +37.7% for the middle half of winners. Median traffic per variant was 2122 visitors. Most ran on homepages (61), product pages (61), other pages (18).

What counts as a social proof test

Testimonials, star ratings, customer logos, user counts or review snippets are added, moved or removed.

Control and variant wireframe for a social proof A/B test
Control on the left, variant on the right. Only the changed element is highlighted.

Adjacent categories, and how often they win: layout (19%), hero image (17%), styling (15%).

Social proof tests by page type

page typewin ratetests
homepages 15%50+
product pages 8%50+

Only page types with at least 30 social proof tests appear here.

Social proof tests that won

  • statement citing 40,000+ self-employed people in Germany as social proof, framed as a numbered step
  • Replaces the taking capsules info block with a subscription benefits list when one-time purchase is selected, restoring original content when subscription is chosen
  • Adds a new USP list block with icons and text (e.g. shipping/returns/guarantee) below existing content
  • statement inviting people to join a claimed 40,000+ Belgian freelancers
  • statement citing 15,000+ positive reviews establishing trust as a debt relief partner, plus an Excellent rating badge based on 15,000+ reviews

Social proof tests that did not

  • Inserts a new green in-stock/ships-today status message below the purchase options label
  • scarcity message stating the item is almost sold out with only a few remaining
  • Hides existing star rating/trust badge and inserts a new trust badge image with review score above the product gallery
  • Replaces features list text with new guarantee/cancellation/trust-count bullet points and restyles them (teal bold text), plus adjusts padding on a heading element

These are individual tests, not rules. Each ran on one site, with one audience, against one page we are not showing you. They are picked to be illustrative rather than sampled at random, and a change that won here can lose on your page for reasons none of this captures. Read them as prompts for what to test, not as findings to copy.

Why the rate looks like this

Social proof competes for attention with everything else on the page rather than replacing it. Adding a block of reassurance does not remove a reason to hesitate; it hopes to outweigh one. That is a weaker mechanism than removing friction or changing what gets seen first.

How this compares

change typewin ratetests
Layout 19%351
Hero image 17%160
Form 16%81
CTA colour 16%32
Styling 15%1414
Price framing 13%117
CTA copy 12%315
Body copy 11%337
Split URL 11%1164
Headline 10%918
Social proof (this page) 9%196

FAQ

So is social proof useless?

No — 9% is not zero, and the winners move more than most. It is that the expectation set by how universally it is recommended is far above what it delivers.

Do star ratings beat testimonials?

Not answerable from this data. Both appear on both sides.

Where does it work best?

Homepages carried the most tests in this set, which reflects where people put it more than where it works.

Who ran these tests

This is every Mida account that ran a readable test — in-house marketers, founders, product teams, and agencies working on client sites. Nothing here is filtered by who ran the experiment or how experienced they are.

Low win rates are normal in experimentation, including at the top end. Microsoft's experimentation team, reporting on its own platform, found that only about one third of ideas improve the metric they were designed to improve — and that roughly another third actively hurt it. That is a dedicated experimentation organisation with research, prioritisation and review behind every test.

A mixed population like this one runs below that. The gap is roughly what disciplined practice buys you: ideas grounded in research rather than opinion, one variable at a time, and tests built so the result can actually be read.

A win is a variant that beat its control on that test's primary goal with a statistically significant result. Tests that never got enough traffic to say anything either way are excluded. The methodology has the full detail, including what these numbers cannot tell you.