Ad Format Testing

An operator's read on Ad Format Testing: the parts that move, the way to apply them, and where to ground your numbers. Built for experimentation leads, analysts, and growth teams.

By David Schaefer · LinkedIn · Updated · 9 min read · 3 sources cited

Key takeaways

  • Ad Format Testing is a topic within Experimentation — a concrete choice, not a vague best practice.
  • Break the goal into named inputs, each with a single accountable owner.
  • Use public benchmarks for orientation; measure your own baseline for targets.
  • Skipping the current-state audit is the fastest way to fix the wrong thing.
  • Pair every primary number with a counter-metric so the goal cannot be gamed.

What Ad Format Testing covers

Ad Format Testing sits inside Experimentation -- the discipline of running controlled tests to find causal impact, from A/B and multivariate tests to geo experiments and lift studies -- and this page makes it concrete enough to act on. Keep that distinction.

Strip the jargon and a simple operating idea is left. Ad Format Testing belongs to Experimentation — the discipline of running controlled tests to find causal impact, from A/B and multivariate tests to geo experiments and lift studies. The aim on this page is practical: a working handle, not a dictionary entry. The frequent error is keeping it abstract when it should be specific. Hold it as a definite call you can argue for and change later.

Experimentation is the discipline of running controlled tests to determine causal impact — including A/B tests, multivariate tests, geo experiments, and platform-native lift tests.

Apply this whenever you need to know if a change causally improves outcomes versus selection effects, seasonality, or coincidence.

Useful sources to read next to this include Optimizely, GeoLift from Meta, Evan Miller's calculators, and the CXL Institute. These reference points keep a debate from restarting from zero each quarter. The rest is mechanics built on that foundation.

How Ad Format Testing works in practice

Ad Format Testing becomes tractable once you separate what you control from what you only watch, then improve them one at a time. Use that as the anchor.

What looks like a black box is a short list of moving parts. You break the goal into parts, give each part an owner, and watch how the parts move. When it works, every contributor knows the number they are accountable for.

Ad Format Testing — what to track, and why
ElementWhat it is
SignalThe measurable change that tells you it worked.
OwnerThe single person accountable for the number.
DecisionThe action a given reading should trigger.
Counter-metricThe number you watch so you are not gaming the goal.

Daily checks catch breakage, monthly reviews catch drift, quarterly resets catch strategy gaps. The idea is plain; the discipline to keep using it is the rare part.

How to apply Ad Format Testing

Four steps carry most of the value: definition, instrumentation, a controlled test, a written review. That part is non-negotiable.

  1. Define the term out loud. Write one sentence everyone agrees with. If two people would describe it differently, you have found your first problem.
  2. Instrument before you optimize. Confirm the metric is captured accurately first. Untrustworthy data turns every later test into a guess.
  3. Change one thing and test it. Compare against a proper baseline and move one thing. That isolation is what makes the finding trustworthy.
  4. Review on a cadence and write it down. Capture what happened and the next step in writing. The trail is what turns a test into institutional knowledge.

Hold the sequence. Instrumenting before defining measures the wrong thing precisely. Everything below is an elaboration of that one point.

Grounding Ad Format Testing in real numbers

Use external benchmarks to orient the numbers, then trust your own measured baseline. Everything else follows from it.

An industry average is a starting question, not a finishing answer. Numbers travel badly between industries, channels, and business models. Use it below to confirm rough direction before trusting your own data.

Claim: The IAB sets the standard viewable-impression threshold at 50 percent of pixels in view for one second for display. Source: [IAB]. Context: A served impression and a viewed one are not the same line in a report.

Numbers here that carry no citation are RGM analysis -- patterns seen across audits, not published facts. It earns trust only once your own numbers confirm it.

Common mistakes with Ad Format Testing

Failures cluster around three causes: no clear definition, isolated optimization, and an unguarded goal. Read that line again.

The mistakes that quietly cost the most
  • Confusing a correlation in the dashboard for a cause.
  • Reporting the number without naming the decision it should drive.
  • Optimizing ad format testing in isolation without checking the downstream business effect.

None of these are exotic. They are the default failure modes. A short pre-mortem on these saves a long post-mortem later.

Quick answers

How should a team treat Ad Format Testing day to day?
As a recurring decision, not a one-time setting. Name it, measure it, and revisit it on a cadence so the choice stays matched to the current goal.
Can small teams use Ad Format Testing?
Yes. Smaller teams often apply it better because fewer handoffs mean the person who owns the lever also owns the number.
Where do RGM observations fit here?
Any pattern labelled RGM analysis comes from reviewing real accounts. It is offered as a tested hypothesis, never as a substitute for measuring your own data.

Frequently asked

What is Ad Format Testing in simple terms?

Ad Format Testing is a topic within Experimentation, the discipline of running controlled tests to find causal impact, from A/B and multivariate tests to geo experiments and lift studies. In plain terms, this page treats it as a recurring decision your team can make with a shared definition instead of restarting the debate each time.

Why does Ad Format Testing matter?

It matters because it shapes how budget, effort, and attention get allocated. When ad format testing is defined and measured well, spend follows what works; when it is fuzzy, spend follows whoever argues hardest.

How do you measure Ad Format Testing?

Pick one primary number, instrument it cleanly, and pair it with a counter-metric so you are not gaming the goal. Then compare against a pre-change baseline rather than an industry average.

What references help with Ad Format Testing?

Useful reference points include Optimizely, GeoLift from Meta, Evan Miller's calculators, and the CXL Institute. Tools matter less than a clean definition and trustworthy measurement; a good tool on a bad definition still produces a misleading dashboard.

What is the most common mistake with Ad Format Testing?

Optimizing it in isolation. A local improvement that ignores the downstream business effect can look like a win on the dashboard while costing money elsewhere.

How often should you review Ad Format Testing?

Daily checks catch breakage, monthly reviews catch drift, quarterly resets catch strategy gaps. The point is a fixed rhythm, so slow drift gets caught before it becomes a quarter-sized problem.

Sources cited on this page

  1. CXL Experimentation — cxl.com/blog
  2. Evan Miller — www.evanmiller.org
  3. Meta GeoLift — facebookincubator.github.io/GeoLift