ASOhack
Türkçe · Español
Back to Blog
ASO Fundamentals

Why Your ASO Tests Aren't Showing Results (How to Fix It)

The five real reasons your ASO A/B tests show no significant difference between variants — and what indie developers should test instead.

ASOhack TeamMay 19, 20266 min read

You set up an A/B test in App Store Connect or Play Console. Ran it for 21 days. Result: "no significant difference between variants."

This is the most common experience for indie ASO testers. The temptation is to give up on testing entirely. Don't. The issue is usually solvable.

This is the diagnostic.

Reason 1: insufficient sample size

The most common issue.

The math

For a test to show a 10% lift with 95% confidence at 5% baseline conversion:

Required sample size per variant ≈ 30,000 product page views

If you're getting 200 views/week, you'd need 150 weeks of testing. Impossible.

What to do

  • Test bigger effects. Don't try to detect a 5% lift; try 20-30%.
  • Run longer. 28-day tests > 14-day tests.
  • Combine variants. Test one variant + control, not 3 variants competing.
  • Accept lower confidence. 80% confidence is OK for indie if you can't reach 95%.

Some tests are simply not feasible at indie volume. Recognize them + skip.

Reason 2: too many variables

You changed the icon AND the first screenshot AND the subtitle simultaneously. Won.

Now you don't know what worked.

What to do

  • One variable per test. Change the icon. Test for 28 days. Conclude. Then change the screenshot.
  • Sequencing matters. Test highest-impact variables first (icon, first screenshot).
  • Document each test. Keep a log of what you changed + what you saw.

Reason 3: the test variant isn't different enough

You changed your icon's background color slightly. Tested. No significant difference.

It's not that the test failed; the variant just isn't meaningfully different.

What to do

  • Test bigger creative shifts. Different character / focal point, not different color.
  • Push variants harder. If you're testing two "similar" variants, expect no result.
  • Use AI / designer help. A designer's variant pushes harder than yours might.

Reason 4: noisy traffic

Your traffic comes from:

  • Paid acquisition (which already has its own creative variants).
  • Branded search (people searching your app name — they'll convert regardless).
  • Referrals (warm intent — they'll convert regardless).
  • Organic search (the actual ASO traffic).

Only the organic search traffic is sensitive to listing changes. The rest is noise.

What to do

  • Look at conversion by source. App Store Connect / Play Console show this.
  • Filter to organic search. That's where ASO matters.
  • Apply your variant findings to organic-search-sourced traffic.

If 80% of your traffic is branded search, your ASO tests will show small absolute lifts.

Reason 5: external factors moved during the test

You launched the test. Two days in, a competitor's app got featured by Apple. Your traffic mix shifted.

Or: Apple updated their algorithm during your test. Or: A holiday created seasonal demand. Or: You ran a press campaign that drove warm traffic.

Any of these confound the test.

What to do

  • Avoid testing during launches / major events.
  • Check for external factors before declaring a winner.
  • Re-run the test if you're suspicious.

Reason 6: you're measuring the wrong thing

You tested for "install conversion." Variant A and B were equal.

But Variant A's installers had 30% higher D7 retention. You missed that.

What to do

  • Measure downstream effects. Conversion isn't the only metric.
  • Track cohort retention by variant if you can.
  • Look at LTV per variant for subscription apps.

A variant that converts equally but retains better is the better variant.

Reason 7: the test was too short

You ran for 7 days. Saw a 30% lift. Stopped the test. Variant A won, right?

Probably wrong. 7-day tests show patterns that disappear over 14-21 days. Weekend traffic differs from weekday. Algorithm learning needs time.

What to do

  • Run for 14-28 days minimum.
  • Don't peek and declare winners early.
  • Set sample size threshold before declaring significance.

Reason 8: confounding tests running simultaneously

You're testing the icon AND running a Product Hunt launch AND iterating on the in-app onboarding.

You can't separate the effects.

What to do

  • Test one thing at a time.
  • Schedule tests with no concurrent major changes.
  • If multiple tests are unavoidable, accept lower confidence.

Reason 9: the variant is just bad

Sometimes the variant you tested is genuinely worse. The test correctly showed no improvement.

What to do

  • Don't get attached. A failed test isn't failure — it's learning.
  • Try a different variant. Don't double down on the same direction.
  • Get external input. Designer, founder, or community feedback can suggest better variants.

Reason 10: you chose a low-leverage test

Testing trivial variables ("change button color") rarely shows lift. The leverage isn't there.

What to do

Test highest-leverage things first:

  1. Icon: ~30% conversion lift possible.
  2. First screenshot: ~20% lift possible.
  3. App preview video (presence vs absence): ~30% lift possible.
  4. Subtitle phrasing: 5-15% lift.
  5. Screenshot order: 5-10% lift.

Anything beyond these usually shows <5% lift.

What to actually test

Based on impact:

High-impact tests (worth running)

  • Icon variants (3-4 distinct directions).
  • First screenshot (outcome-led vs UI-led).
  • App preview video (with vs without).
  • Subtitle copy (outcome vs feature).

Medium-impact tests

  • Screenshot order.
  • Second / third screenshot variants.
  • Localization of specific elements.

Low-impact tests (skip)

  • Button color in screenshots.
  • Decimal in pricing display.
  • Minor visual polish.

The compounding insight

Even if individual tests show no significant difference, the act of testing builds intuition. After 10-20 tests, you start to know what works in your category.

This compounded intuition is more valuable than any single test result.

Common testing mistakes

  • Stopping tests early. 7-day tests rarely conclusive.
  • Testing trivial variables. No leverage.
  • Multiple simultaneous tests. Can't separate.
  • Ignoring sample size. Indie apps often don't reach significance.
  • Not documenting. Compound intuition lost.
  • Quitting on ASO testing after a few failures. Testing pays back over years.

What to do when no test ever wins

If you've run 5-10 tests with no clear winners:

  1. Test bigger variants. Smaller variants don't show lift.
  2. Re-examine your baseline. Is conversion actually problematic?
  3. Talk to users. Quantitative testing has limits; qualitative reveals issues.
  4. Get expert input. A specialist can suggest variants you'd miss.

When to skip A/B testing

For very early-stage indie apps:

  • Volume too low for significance.
  • Time better spent on product / acquisition.
  • Just ship the best variant your judgment can produce.

A/B testing becomes valuable when you have enough traffic + time.

Try the tools

Ready to Optimize Your App Store Listing?

Try our free ASO tools — no signup required.