Solutions · Testing and measurement
Honest split testing and measurement for Shopify stores
Most CRO tests read inconclusive, and most of the rest are called early. This page says how Liftable measures, which parts are built, and exactly where they have and have not run, so a data team can hold it to the method before anything is switched on.
- Method
- Reviewed 6 October 2026
Peeking, and why most results are not real
A fixed-horizon test read every morning and stopped on the first good day inflates its false-positive rate several times over. A sequential test (Liftable uses a mixture sequential probability ratio test) is built to be read continuously: its p-value is always valid, so a cron job can read it every hour without changing what a result means. The method page has the thresholds.
Starting a test the store cannot read
Before a test starts, a power check takes the store’s last 28 days of sessions on that surface and the metric’s variance, and computes how long it would take to detect the pre-registered effect. If that is longer than the limit, the test is refused and the page says how many sessions it would need. As a rough guide, about 50,000 sessions a month is where a storefront change can be measured in a few weeks; below that, only larger effects read in that time.
What makes a dollar figure
Winning the test is not a revenue claim. A shipped change keeps a 10% holdback on the old version for 30 days, and the receipt compares real Shopify orders between the two. Only that comparison puts a number next to a dollar sign, and refunds that arrive later shrink it. No store has a verified line on its receipt yet, and the site says so wherever it mentions one.
What Liftable does about it today
What is live today. Liftable finds the opportunity and our founding-team programme implements and measures the first fix with you. Automated testing is not live on any merchant’s store yet.
- Available now the scan, the evidence, the drafts, the approval
- With the founding team implementing and measuring the first fix
- Coming automated tests on a merchant’s store
What you get
- A pre-registration stored with every test: the metric, the minimum detectable effect, the alpha and the guardrails, before any traffic is split.
- A power check that refuses an underpowered test and says what traffic it would need, instead of letting it run for a month and read nothing.
- Guardrails on revenue per session, contribution margin, checkout errors and page speed, armed from the first read; a sample-ratio check at every read.
- A receipt that counts only tests that won, measured against a holdback on real orders, restated when refunds arrive.
- The founding team running the first measured fix with you, with a published result even when the test loses.
Who it is for
A good fit
- Brands with the traffic to read a test (about 50,000 sessions a month as a rough guide) and a team that would rather hear ‘not proven’ than a flattering number.
- Data teams who want the method in writing before an install.
Not the right tool
- Stores below the traffic line that want a weekly stream of ‘winners’. The power check will refuse most of those tests, and we tell you which side of the line you are on first.
- Teams with their own hypothesis backlog and a visual editor habit: Shopify Rollouts or Shoplift run what you design.
Questions
Has Liftable run a test on a real store?
A shipping test has run on Liftable’s own development store. No automated test has run on a merchant’s store, and no store has a verified-revenue receipt. The founding team implements and measures the first fix with each founding brand by hand.
Can I see the p-value?
For a sequential test, yes: it is always valid, so reading it changes nothing. Liftable shows no p-value for a fixed-horizon test and calls no winner outside the boundary a test was registered with.
What is a proxy metric and when is it used?
Add-to-cart rate or checkout-start rate, read instead of conversion rate when a store cannot read revenue in time. A proxy test carries a revenue guardrail, never ships automatically and is never shown as verified revenue. The experiment page labels it as a proxy.
What stops a bad test from hurting sales?
Guardrails: a drop in revenue per session or contribution margin past the stop rule, a rise in checkout errors or in page load time, or a sample-ratio mismatch each stop the test. A stopped test keeps its record and says why it stopped.
See which of these your store has. The scan is free, takes about a minute, and needs no install.
Scan your storeGuides
- Testing · 5 minHow to split-test when you don’t have the trafficMost stores can’t read a purchase-level A/B test in a sensible time. Why, and the adjustments that make testing work at the volumes stores actually have.Read
- Testing · 7 minHow to test a fix without breaking your live storeTest through a route that never edits your theme’s code, read the change first, run one test per surface, keep the rollback to one switch. The safe sequence.Read
- Testing · 7 minCan you automate conversion optimisation end to end?Five of the six steps in conversion optimisation can run without a person. The sixth should not, and the difference matters before you buy anything.Read
Terms
Integrations
Other solutions
- Cart and checkoutShipping surprise, discount-code hunting and the cart drawer: where orders are lost after the add to cart.Read
- Product pagesThe ten things a product page gets wrong before the shopper reaches add to cart.Read
- Mobile conversionThe part of the mobile gap that is yours to close: fold, taps, speed and popups.Read
- Site speed and appsThe apps that slow the store, the ones doing nothing, and the ones doing the same job twice.Read
Last reviewed against the product on 6 October 2026.
Find what's holding your Shopify store back
Liftable reads your storefront as a shopper would and shows what looks wrong on your own pages, desktop and mobile: friction, slow pages, missing trust and apps doing nothing, with the evidence for each. Free, in about a minute.
Scan your store.
See where your store is losing revenue and what to fix first. It is free, takes about a minute, and needs no install.