Skip to content

Audience and Message Research

Message Testing Scorecard

Compare candidate messages for unaided comprehension, relevance, credibility, differentiation, limitation awareness, recall, and ability to support a suitable action.

Free editable Markdown · Message strategists, user researchers, and brand teams ·

Download Markdown

Accessible HTML preview

Blank template

The downloaded file contains the same fields in editable Markdown.

Test setup

Decision to inform
[Where the message will be used]
Audience and situation
[Who is deciding what?]
Candidate labels
[A, B, C, without evaluative names]
Presentation context
[Text, page section, ad, email, or another]
Order method
[Randomized, rotated, or fixed with reason]
Recruitment criteria
[Relevant experience]
Participant code
[Non-identifying label]
Consent and session date
[Method and YYYY-MM-DD]

Candidate observation

Candidate
[A / B / C]
Unaided meaning
[Participant’s explanation]
Expected offer or action
[What do they think happens?]
Intended audience
[Who do they think it serves?]
Important words noticed
[Participant language]
Misunderstanding
[Invented feature, missed condition, or ambiguity]
Evidence requested
[What would make it credible?]
Likely next step
[What would they do?]

Dimension scorecard

Comprehension
[1–5 and evidence]
Relevance
[1–5 and connection to situation]
Credibility
[1–5 and reason]
Differentiation
[1–5 and compared alternative]
Limitation awareness
[1–5 and condition recalled]
Actionability
[1–5 and expected next step]
Recall after delay
[What remained, if tested]
  • Scores include participant reasoning.
  • The researcher did not correct interpretation before recording it.
  • One severe misunderstanding is not hidden by an average.

Decision synthesis

Strongest candidate by dimension
[Do not force one overall winner]
Material tradeoff
[What improves while something else weakens?]
Unsupported expectation
[Claim or implication to remove]
Revision
[Specific wording, proof, order, or offer change]
Product or policy issue
[Problem copy cannot solve]
Decision owner
[Name or role]
Retest requirement
[What and with whom]
Research limitation
[What this round cannot establish]

How to use this template

  1. Define the message decision, audience situation, candidate set, success criteria, and claims that must remain accurate.
  2. Standardize presentation, choose an order or randomization plan, and prepare neutral comprehension prompts.
  3. Test each message for unaided meaning before asking about preference, appeal, or ratings.
  4. Score dimensions with written evidence, recording misunderstandings, limitation awareness, and requested proof.
  5. Compare tradeoffs, select revisions rather than an automatic winner, and document what requires retesting.

Define the decision and control the comparison

State what will change based on the test: a homepage promise, service explanation, campaign frame, or audience-specific introduction. Keep candidate messages comparable in format, length, context, and visual treatment unless those are deliberate variables. Randomize order when possible because the first option can establish vocabulary and the last may be easier to remember. Avoid testing a polished design against plain text if the objective is wording. Recruit people who face the relevant decision and include variation likely to affect understanding; colleagues who already know the strategy cannot stand in for readers.

Ask for interpretation before preference

Show one message, remove or hide it when appropriate, and ask participants to explain what it means in their own words, whom it is for, what they expect, and what they might do next. Capture invented features, missed conditions, and overbroad interpretations. Then explore relevance and credibility using evidence from their situation. A favorite message can still be misleading, while a less exciting one may communicate accurately. Ratings are useful prompts, but the reason behind a score supplies the actionable evidence. Never coach a confused participant toward the intended answer.

Evaluate claims, limits, and action together

A strong message makes a useful promise that the audience can understand, believe for stated reasons, distinguish from alternatives, and act on appropriately. Test whether essential limitations remain visible rather than hidden in later copy. Ask what evidence the participant would need and whether the proposed next step matches their readiness. Analyze dimensions separately before choosing; an average can conceal excellent relevance and unacceptable credibility. Report differences and uncertainty, revise the underlying offer or proof when wording cannot solve the problem, and retest material changes with fresh participants.

See the fields in context

Fictional example: repair booking message

Alder Repair Desk and all participant responses are invented.

  • Candidate A: “Repairs without the wait” was memorable but led participants to expect immediate service.
  • Candidate B: “Choose a repair window and see the estimated turnaround before booking” was understood accurately.
  • Credibility: Participants wanted evidence explaining which device types received an estimate.
  • Decision: Revise B to name eligible categories and link the current scheduling explanation.
  • Rejected shortcut: The team does not average the scores and declare A the winner merely because its recall rating was higher.

Frequently asked questions

How many candidates should be tested at once?

Use the smallest set that represents meaningful strategic alternatives. Too many similar versions exhaust participants and encourage shallow preference picking. Early research may compare broad frames; later work can refine one selected direction.

Should participants rate messages on a numeric scale?

Ratings can structure comparison, but always ask for the reason and capture unaided meaning first. Treat scores as ordinal research prompts, not precise measurements, unless the study design and sample support statistical analysis.

Can a message win if people misunderstand it?

A material misunderstanding should block or revise it, even if preference is high. Attractive ambiguity may temporarily improve appeal while creating unsuitable decisions, disappointment, or compliance risk.

What if no candidate performs well?

Do not pick the least weak by default. Revisit audience evidence, the offer, proof, differentiation, and constraints. The research may show that the underlying proposition needs work rather than another round of synonyms.

File details

File name
message-testing-scorecard.md
Format
Markdown (.md)
Size
3 KB
Designed for
Message strategists, user researchers, and brand teams

Usage note: Use the scorecard to compare messages under the same research conditions, not to manufacture a winning average. Record participant explanations and misunderstandings beside ratings. Test only claims the organization can support, protect participant data, and do not present small qualitative samples as population estimates.