Kova Commerce · Hyde & Hare · Meta testing framework

How we'll test your Meta ads

We test ideas, not single ads. Each idea gets its own space, a fair budget and a clear pass or fail. The winners move into one main campaign, where a small set of proven ads gets most of the money.

1 · The structure

Two campaigns: one finds winners, one spends on them

Testing takes a small, fixed share of the budget. Everything that proves itself moves across to Scale, and Scale stays lean so every ad in it gets enough spend to keep learning.

Campaign 1 · Testing
10-15%

One ad set per idea. Each idea has 3 to 5 versions and its own set budget, so every idea gets a fair read.

Ad set · winner

Idea A

Moves to Scale →
Ad set

Idea B

Ad set

Idea C

Campaign 2 · Scale
85-90%

Only proven ads. Meta moves budget to whatever's working best that week.

Proven ads · 10 to 15 max
Winner inWeakest out
i
We never split the audience. Persona and stage of awareness live in the message, not in the targeting. Every ad set goes to the same broad audience and Meta finds the right people for each message. That keeps everything learning in one place instead of starving small audiences.
2 · What goes into the test ad sets

We test whole ideas first, then refine the winners

Creative matters from the first second, so we never test words on their own. Each idea is one message with creative made to say it: the video, the words on screen and the copy all say the same thing. Once an ad has proven itself, we refine it one change at a time.

What we test
Idea

One message plus the creative made to say it.

Inside an idea
Version

One ad. Each idea has 3 to 5.

Start of a version
Hook

The first 3 seconds: shot, words on screen, first line spoken.

What it looks like
Format

Talking to camera, demo, review card, static.

1

Test ideas

Question: which idea sells best? One ad set per idea. This is how most of our testing runs.

ChangesEverything, idea to idea
Stays the sameProduct and audience
Ad set 1 · Persona A
Their main reason to buy told by a real person to camera
Ad set 2 · Persona B
A different buyer's need shown by the product doing the job
Ad set 3 · Persona C
What worries a third buyer answered by a customer review
3-5 versions Each ad set holds 3 to 5 versions of its idea, with different creative or a different opening.
Proven
winner
2

Test one change (A/B test)

Question: what gets even more out of an ad that already works? Only for specific cases.

ChangesOnly the hook, or the headline on an image
Stays the sameThe proven video or image, and the copy
Version A · Same video
Opens on the product hook 1
Version B · Same video
Opens on a customer's words hook 2
Version C · Same video
Opens on the offer hook 3

Also for: the same image with a different headline on it.

Even spend Meta's A/B test tool gives every version an equal share, so each one gets a fair read.
Why not three hooks side by side?

Inside a normal ad set Meta picks a favourite early and gives it nearly all the budget, so the other versions never get a fair read. That's fine when we're judging the whole idea, but it can't tell us which hook is best. For that we use the A/B test tool.

Why A/B tests are only for winners

Every version needs about 3× the target cost per sale on its own, so a three-way test costs about three times as much as one idea test. It's worth it on an ad that's already earning, not on an untested one.

Persona gets its own ad set

A different buyer has a different reason to buy: a beanbag as a gift is a different message from a beanbag as a statement piece for yourself. So each persona gets its own ad set, with the hook and the copy made for them.

Stage lives inside every ad

People see the same ads many times, in no set order, so one ad does the early work and the closing work depending on who's watching. Each ad carries copy for different stages as its text options, and Meta shows each person the one that fits.

The "why" comes from the test log: every idea is labelled by its message, its format and its proof, so patterns show up across rounds.

3 · How we build each test

Start with what we want to learn, and the build follows

Every test answers one question. We go down these in order and stop at the first yes.

What do we want to learn?
Question 1

Are we changing just one thing on an ad that's already proven?

A different hook on the same video, or a different headline on the same image.

YES
A/B test

Meta's A/B test tool

One version per hook or headline. Same copy on every version. Spend is split evenly so each one gets a fair read.

Example C below
NO
Question 2

Are we comparing different reasons to buy?

Different buyers, or different messages, for the same product.

YES
One ad set per reason

Idea test, split by buyer

Each ad set gets its own hook and copy made for that buyer. The product stays the same, so the buyer is what we're comparing.

Example B below
NO
Question 3

Do we have one message and want to find the best way to show it?

A new idea, or a new take on one that's working.

YES
One ad set for the idea

Idea test with different creative

3 to 5 versions: carousel, video, a re-cut, a static. The same copy on all of them, with 3 to 5 text options that vary the writing and the stage.

Example A below
NO
Question 4

Do we only want better words?

The creative is fine; we think the copy could work harder.

YES
No new test

Add text options to ads already running

New writing structures and stage copy go in as extra text options. We read how each one does as a steer, not a verdict.

Three worked examples

What each branch looks like when we build it.

A · The cosy corner

Question 3 · Idea test

We want to learn: does "build a cosy corner" sell, and which way of showing it works best?

Ad set · Idea: Build a cosy corner
  • Version 1 · How-to carousel rug, cushions, throw, step by step
  • Version 2 · How-to video the corner coming together
  • Version 3 · Re-cut of the video opens on the finished corner
  • Version 4 · Static the finished corner, one line on it
CopyThe same on every version, with 3 to 5 text options: last season's best copy, problem then solution, a real review, the how-to steps, and a short one.
How we read itJudged against the other ideas running that week. If it wins, the version Meta favoured moves to Scale.

B · The beanbag: gift or statement piece?

Question 2 · Split by buyer

We want to learn: which buyer should lead the beanbag ads in November? Run in October, so the winner is ready for the peak.

Ad set 1 · Beanbag as a gift
  • Opens on it being given the moment it's unwrapped
  • Opens on who it's for the person who has everything
  • A real review from a gift buyer as the ad
Ad set 2 · Beanbag as a statement piece
  • Opens on the room the beanbag as the centrepiece
  • Opens on the material close up on the sheepskin
  • A real review from someone who bought it for themselves as the ad
CopyWritten for each buyer. Gift copy talks about the person receiving it; statement copy talks about your home.
How we read itAd set against ad set. Each has its own budget, so both buyers get a fair read.

C · A proven hot water bottle video, three openings

Question 1 · A/B test

We want to learn: which opening gets the most out of a video that's already sold well?

A/B test · Same video, only the first 3 seconds change
  • Version A · Opens on the product the original
  • Version B · Opens on a customer's words a real review on screen
  • Version C · Opens on the offer the complimentary gift, first
CopyExactly the same on all three, so the opening is the only difference.
How we read itSpend is split evenly. The winning opening goes into Scale as the new version of the ad.
4 · The rules

A fair budget for every idea, and a clear pass or fail

Every rule is set against your target cost per sale, so it scales with the account.

Budget per idea
3×

About three times the target cost per sale before we call it.

Stop early
2×

Spent twice the target cost with no sale? We pull it.

Read after
7 days

Judged on 7-day click sales, not on who merely saw it.

Judge the idea
Idea vs idea

Meta favours one version early. Versions with little spend are unread, not losers.

Example only · if the target cost per sale were £30
No sale by here → stop
Verdict
£0launch £602× target £903× target
5 · What happens to each idea

Three outcomes, each with a next step

Every idea ends up in one of three places. Nothing sits in testing forever.

After 7 days
Read

the idea against the other ideas running that week

Win

Moves to Scale

Hit the target cost per sale. It's copied into Scale as the same post, so it keeps its likes and comments, and the weakest Scale ad comes out.

Unclear

One more week or a re-cut

Close to target, or it didn't get enough spend to judge. Top it up, or change one thing and run it again.

Lose

Stop and log it

Hit 2× target with no sale, or clearly behind. We write down what we learned so the next idea is smarter.

6 · The weekly rhythm

Same loop every week

A steady rhythm stops us reacting to one good or bad day.

MON

Launch and graduate

New ideas go live. Last week's winners move to Scale.

TUE-SUN

Hands off

Tests run untouched so the read is clean.

WEEKLY

Tiredness check

We check the top spenders in Scale for signs they're wearing out.

EVERY TEST

One line in the log

What we tested, what happened, what we learned.

7 · Choosing the next test

Four questions every Monday, in this order

Before anything launches, we work down the list. The first question with an answer sets that week's tests.

Protect
1

Is anything in Scale tiring?

If a top ad is wearing out, the first job is its replacement: ideas close to what made it work, or a fresh hook on the same ad.

Follow on
2

What did last week tell us?

Every log line ends with "so next we test…". A win gets refined, a loss moves to a different idea, an unclear one gets another week.

Plan ahead
3

What's coming up?

Work back from stock, launches and offers. Beanbags peak in November, so the gift vs statement piece test runs in October and the winner is ready in time.

Explore
4

What have we never tried?

Buyers with no idea built for them yet, and claims or formats we've never used. This is where new winners come from.

When there's more than one candidate, we score each on three things and run the highest first:

Score 1
Money at stake

How much revenue rides on the answer.

Score 2
Evidence it could win

Reviews, past ads, what's working for others.

Score 3
Cost to make

Can we build it from what we already have?

The weekly mix

Roughly one test that builds on a winner and one or two that try something new. The testing budget covers several ideas a week, so making the creative is usually the limit, not the money.

Say it before we launch

Every test starts with what we expect to happen. If we can't say which result would change our next move, it isn't worth testing yet.

8 · The test log

Every test leaves a lesson behind

One shared log, one line per test. Every idea is labelled by its message, format and proof, so after a few rounds we can see what's really winning, e.g. one message beating the rest whether it's a video or a static. Over time this becomes the playbook for what Hyde & Hare customers respond to.

TestMessage · format · proofWhat we expectResultWhat we learned
Persona A ideaMain reason to buy · talking to camera · a customer says itSpeaking to their main reason beats a general messageFilled in after 7 daysFilled in after 7 days
Hook A/B test on a winnerSame video · 3 different openingsOpening on the offer beats opening on the productFilled in after 7 daysFilled in after 7 days