Learn how to run video ad testing step by step, from sample size to survey questions. Start your test with a free template today.

White outline of Goldie, the SurveyMonkey mascot

Summary:

  • Follow a structured sequence to test a video ad: define objectives, select a testing method (monadic vs. sequential monadic), determine sample size, and build a focused survey.
  • Avoid common pitfalls like under-sampling, testing too many variables, ignoring statistical significance, and poor audience recruitment.
  • Pre-testing creative reduces wasted spend, diagnoses performance issues, and builds benchmarks to ensure evidence-based media decisions.

Video ad testing is the practice of showing a video ad, or several versions of one, to a sample of your target audience before it goes live, then measuring how they react.

Instead of guessing which cut or message will land, you get direct feedback on recall, clarity, emotional response, and purchase intent while there is still time to make changes.

This guide walks through the full process: setting an objective, choosing between monadic and sequential monadic testing, calculating a sample size, writing survey questions that predict performance, fielding the study, and turning results into a creative decision.

You will also find ready-to-use question wording and a list of mistakes that quietly wreck test results.

Video ads are expensive to get wrong. Test yours with a real audience before you spend on production.

A video ad test only produces useful answers when the setup is deliberate. Treat the steps below as a sequence, not a menu.

Before you write a single question, decide what decision this test needs to inform. Common objectives include:

  • Picking a winner among two or more finished cuts
  • Diagnosing why an existing ad underperforms
  • Validating a storyboard or animatic before full production
  • Confirming that a message lands the same way across audience segments

Your objective determines which metrics matter most. A launch-decision test should weight purchase intent and overall appeal heavily. A diagnostic test should lean on open-ended feedback and message clarity questions instead.

Monadic testing shows each respondent only one video ad. Nobody compares options side by side, so you avoid comparison bias and get a cleaner read on how an ad performs on its own merits. The tradeoff: you need a separate, sufficiently large group of respondents for every ad you test, which raises the total sample size and cost.

Sequential monadic testing shows each respondent two or more ads, one after another, and asks the same questions after each. This lets you test more concepts with a smaller total sample and get a direct, within-person comparison. The tradeoff is order bias, since the first ad someone sees can color how they judge the next one, so rotate ad order across respondents to cancel it out.

Choose monadic for a small number of well-developed concepts where you want the most rigorous, unbiased read on each. Choose sequential monadic when you are screening several early-stage concepts on a tighter budget.

Sample size drives whether your results are noise or signal. For monadic video ad testing, 200 to 300 respondents per ad variant is a commonly cited industry range for detecting meaningful differences in attributes like appeal or purchase intent. Sequential monadic designs can often work with a smaller total sample, since every respondent evaluates multiple ads.

A few factors push the number up or down:

  • Testing across several demographic subgroups requires enough respondents in each one.
  • Ads with subtle differences need larger samples to detect a real gap.
  • A tight screener that filters out most of your panel means fielding a larger gross sample to hit your net target.

If you are unsure where to start, a sample size calculator can translate your target confidence level and margin of error into a concrete number.

Structure your survey in this order:

  1. Screener questions to confirm the respondent matches your target audience
  2. A viewing prompt that embeds or links to the video ad
  3. Core evaluation questions covering recall, clarity, emotional response, and purchase intent
  4. Open-ended questions for unprompted reactions
  5. Demographic questions at the end, so they do not bias earlier answers

Keep the survey focused: once a questionnaire creeps past 30 questions, completion rates drop and answer quality suffers, so trim anything that does not map back to your objective.

Your results are only as good as the people answering.

Recruit respondents who resemble your actual target market, whether that means current customers, a lookalike audience, or a panel filtered by the same age, income, or interest criteria you use in media buying.

A market research audience panel can help you reach a specific demographic quickly if your own contact list is too small.

Launch the survey and monitor completion rates as responses come in. Watch for:

  • Speeders who finish implausibly fast
  • Straight-lining, where someone picks the same answer on every scale question
  • Uneven completes across subgroups

Fielding windows for a targeted panel typically run from a few hours to a couple of days, depending on how niche your audience criteria are.

Once responses are in, compare ads, or compare an ad against your own prior benchmarks, across each metric you set out to measure.

A widely used method is the Top 2 Box score, which combines the two most positive answer choices for a question into a single percentage.

If 35 percent of respondents rated an ad "extremely likely to buy" and 25 percent rated it "very likely," the Top 2 Box purchase intent score is 60 percent.

Layer in a statistical significance check before declaring a winner. A gap of a few points between two ads might sit within the margin of error, especially at smaller sample sizes, so do not treat every numeric difference as a real one. Read the open-ended responses too: quantitative scores tell you what happened, and the verbatim comments tell you why.

Translate results into a decision:

  • If one ad clearly outperforms on your priority metrics, move it forward into media buying.
  • If results are close or inconclusive, revisit the creative rather than picking a winner by a hair.
  • If a specific scene or line keeps coming up in open-ended feedback, flag it for the edit even if the overall score is strong.

Keep a record of scores across campaigns so you build your own internal benchmarks over time, rather than treating each test as a one-off.

Below are ready-to-use questions covering recall, message clarity, emotional response, and purchase intent. Adapt the wording to fit your brand voice, but keep the scale structure consistent across every ad you test so comparisons stay fair.

  1. What do you remember most about the video you just watched? (open-ended)
  2. Which brand or product was featured in the video? (open-ended, to check unaided brand recall)
  3. How much of the video do you recall watching? (all of it / most of it / some of it / very little of it)
  1. How clear was the main message of this video? (extremely clear / very clear / somewhat clear / not so clear / not at all clear)
  2. In your own words, what was this video trying to tell you? (open-ended)
  1. How did this video make you feel? (select all that apply: excited, curious, amused, indifferent, annoyed, confused, inspired, other)
  2. How enjoyable was this video to watch? (extremely enjoyable / very enjoyable / somewhat enjoyable / not so enjoyable / not at all enjoyable)
  1. How believable is the message in this video? (extremely believable / very believable / somewhat believable / not so believable / not at all believable)
  2. How relevant is this video to your needs or interests? (extremely relevant / very relevant / somewhat relevant / not so relevant / not at all relevant)
  1. If this product were available today, how likely would you be to purchase it based on this video? (extremely likely / very likely / somewhat likely / not so likely / not at all likely)

Round out the survey with a screener question up front and demographic questions at the end, such as age, gender, or income.

Global ad spend runs into the hundreds of billions of dollars a year, and video commands a growing share of that budget across streaming, social, and connected TV. Every dollar spent on a weak concept is a dollar that could have gone toward a stronger one, which is why pre-testing creative before it reaches paid media has become standard practice at disciplined marketing teams.

Testing before launch gives you real advantages over launching untested and hoping for the best:

  • Less wasted media spend. Catching a weak concept in a survey costs a fraction of what it costs to run that concept through a paid campaign and find out it underperforms.
  • A diagnosis, not just a verdict. A test shows which specific attribute, such as believability, relevance, or clarity, is holding a weaker ad back, so revisions target the right problem.
  • Segmented findings. Filtering results by age, gender, or other demographics shows whether one ad resonates broadly or plays much better with a specific slice of your audience.
  • Institutional knowledge over time. Testing consistently across campaigns builds a running benchmark, so new creative gets compared against your own historical performance, not a blank slate.

None of this replaces good creative judgment; it gives that judgment something concrete to stand on before impressions are on the line.

Even a well-intentioned test can produce misleading results if it falls into one of these traps.

Running a test with too few respondents per ad variant is the most common mistake. A small sample can produce a score that looks decisive but is really just statistical noise. If budget is tight, consider a sequential monadic design instead of stretching a monadic design too thin.

If two versions differ in music, voiceover, pacing, and call-to-action all at once, a winning score tells you almost nothing about which change drove the result. Isolate one or two variables per test.

A small gap between two ads might sit well within the margin of error for your sample size. Always check significance before declaring a winner, and treat close results as a tie rather than a victory.

A blurry upload or a slow-loading link can tank scores for reasons that have nothing to do with the creative itself.

Testing a niche product with a general population panel, or skipping screener questions, produces feedback from people who were never going to buy it.

Numeric scores tell you what happened; skipping verbatim feedback means missing why, often the more useful part of the test.

  • What sample size do I need for video ad testing?
  • What is the difference between monadic and sequential monadic video ad testing?
  • How is video ad testing different from a live in-platform A/B test?
  • How many video ads should I test at once?
  • What metrics should a video ad testing survey measure?

You do not need a big research team to test a video ad properly. You need a clear objective, the right method for your number of concepts, a sample size that can detect a real difference, and questions that map back to the decision you are trying to make.

If you are ready to put this process into practice, start from a video ad testing template to build your questionnaire in minutes, or try the video ad testing feature if you want a guided setup with built-in scorecards and an integrated respondent panel.

Two marketing employees, one reviewing a paper with brand strategy, and the other holding a printout of charts

SurveyMonkey can help you do your job better. Discover how to make a bigger impact with winning strategies, products, experiences, and more.

A man and woman looking at an article on their laptop, and writing information on sticky notes

A diary study is a qualitative research method where people log experiences over time. Learn when to use one and see real examples.

Smiling man with glasses using a laptop

Learn how to run a win-loss analysis with a repeatable framework, real interview questions and a free template. No CI vendor required.

Woman reviewing information on her laptop

Learn how to run a market assessment: a TAM, SAM, SOM walkthrough, sample questions, and a go/no-go checklist. Try SurveyMonkey free.