Think Build Implement Repeat
London, UK +44 7367 067226
WhatsApp FOLLOW f in X
Shopify & eCommerce

Changing Things Without Losing Sales

Last updated:

Split testing needs volume you may not have

Reliable split testing requires a substantial number of conversions per variant. A store with forty orders a month cannot reach statistical confidence on a small change within any useful timeframe.

Running an underpowered test and acting on the result is worse than not testing, because it produces confident conclusions from noise.

Sequential testing instead

  1. Record four weeks of baseline before changing anything
  2. Make one change
  3. Measure four weeks after
  4. Compare, allowing for seasonality
  5. Then make the next change

It is slower and it is honest about what small-store data can support.

Prefer changes with obvious effects

  • Showing delivery cost earlier
  • Adding prices where there were none
  • Halving the page load time
  • Adding reviews to product pages
  • Removing a required field from checkout

Those produce effects large enough to see without sophisticated measurement. Fine-grained optimisation needs volume you probably do not have.

Watch for confounds

ConfoundGuard
SeasonalityCompare against the same period last year
A campaign runningNote what else changed
Traffic source shiftingSegment by source
Stock availabilityNote anything that went out of stock

Write down what you changed

A simple log — date, what changed, why, what happened — is worth more than any testing tool for a small store. Six months later it is the only record of what actually worked.

Most stores make dozens of changes a year and can recall almost none of them.

Frequently asked questions

How much traffic before split testing is viable?

Enough for a few hundred conversions per variant within a reasonable period. Below that, sequential testing is more honest.

Can we test more than one thing?

Not if you want to know which worked. One change per period, in a small store.

What if the result is ambiguous?

Usually it means the change was too small to detect. Prefer bigger changes with bigger expected effects.

Should we use a testing tool?

At low volume, a spreadsheet log and four-week comparisons is more useful and considerably cheaper.

Keep reading

Changing things and hoping?

A four-week baseline and a change log costs nothing and tells you what actually worked.

Book a free 30-minute call Get a project estimate WhatsApp us

Related services

What we build for problems like this one

Shopify DevelopmenteCommerce Development