ToolsTemplate Library · Meta Ads · Planning

Creative Testing Log template.

One hypothesis, one variable, one verdict per row — with a four-test sequence modeled so the discipline is obvious.

Download the CSV Free · 9 columns · 4 pre-filled rows · updated July 2026

What's inside — the exact file

This is the complete template, not a sample. The rows are worked examples — replace them with your data (and delete them before any platform import).

test_idhypothesisvariable_testedaudiencestart_dateend_dateresultdecisionnext_test
CT-001Customer-quote hooks beat feature hooksPrimary text hookProspecting LAL 1%CT-002
CT-002Short quote beats long quoteText lengthProspecting LAL 1%CT-003
CT-003Winning text + video beats staticFormatProspecting LAL 1%CT-004
CT-004Same winner holds for retargetingAudience transferRetargeting 30d

What this template is for

Most 'creative testing' is spray-and-pray with a dashboard: five ads launched, one 'wins', nobody can say why, and the learning evaporates. This log imposes the discipline that makes testing compound: one hypothesis, one variable, one verdict per row — with a four-test sequence modeled so the discipline is visible.

It's a lab notebook for your ad account. Six months of filled rows is a playbook of what your audience actually responds to — the asset that survives ad fatigue, account resets, and team changes.

How to use it, step by step

  1. Write the hypothesis before the ad. 'Problem-led hooks beat feature-led hooks for cold traffic' is testable. 'Try some new creatives' is not. If the hypothesis can't be wrong, it isn't one.
  2. Change exactly one variable. Hook, visual format, offer framing, CTA — one per test. Change two and the result attributes to neither; the variable_tested column enforces the honesty.
  3. Give tests a fair fight. Same audience, same budget structure, a pre-agreed run window long enough for significance at your volume. Ending tests early because one 'looks like it's winning' is how noise becomes strategy.
  4. Record the verdict and the decision separately. Result is what the data said; decision is what you'll do (scale it, iterate on it, kill the line of inquiry). Writing both is what turns a test into knowledge.
  5. Chain the next test from the last. The next_test column makes testing sequential: each verdict seeds the following hypothesis. Four tests that build beat ten that wander.
  6. Review the log quarterly for patterns. Individual tests answer small questions; the accumulated log answers the big one — what does our audience consistently respond to? That review is where creative strategy actually comes from.

What each column means

test_idSequential ID — makes the chain referenceable.
hypothesisThe falsifiable claim being tested.
variable_testedThe one thing that differs between variants.
audienceWho saw it — results only compare within the same audience.
start_dateTest start.
end_datePre-agreed end — set before launch.
resultWhat the data said, with the numbers.
decisionWhat you're doing about it.
next_testThe follow-up hypothesis this verdict suggests.

Common mistakes to avoid

  • Testing five variables at once and learning nothing attributable — the log's one-variable rule exists because everyone breaks it under deadline pressure.
  • Calling winners at day two on a handful of conversions — small samples produce confident noise.
  • Recording results nowhere, so the same failed hook gets retested every six months by whoever's new.
  • Testing trivia (button colors) while the big levers (hook, offer, format) go unexamined — test in order of expected impact.

Questions, answered

How long should a creative test run?

Until it reaches enough conversions per variant to trust the difference — at typical lead-gen volumes, one to two weeks is common, and pre-committing the window in end_date is what keeps the answer honest. High-volume ecommerce can call tests faster; low-volume B2B needs patience or bigger differences.

Does Meta's own A/B test tool replace this log?

The experiments tool handles the mechanics (clean splits, significance) and is worth using — but it doesn't remember your hypotheses, decisions, or the chain of learning across tests. The log is the institutional memory layer on top of whatever mechanism runs the test.

What should we test first?

The hook — the first line and first visual second. It's the highest-leverage variable in nearly every account because it gates whether anything else gets seen. The pre-filled four-test sequence in the template starts exactly there.

This is a Market Disrupt worksheet, not a vendor file. Import-format templates follow each platform's documented layout as of July 2026 — platforms evolve, so validate against current documentation before a large import.

Want the account run, not just audited?

Paid social measured against pipeline, not platform dashboards — with the CRM wiring that makes the numbers honest.

Talk through your ad account