Email and Lifecycle Content
Email Subject Line Test
Compare truthful subject lines for one defined audience using a documented hypothesis, controlled variants, suitable metric, stopping rule, and decision record.
Free editable Markdown · Lifecycle marketers, email editors, and analysts ·
Accessible HTML preview
Blank template
The downloaded file contains the same fields in editable Markdown.
Experiment foundation
- Experiment ID
- [Durable identifier]
- Email and lifecycle moment
- [Message and trigger]
- Recipient audience
- [Eligibility and exclusions]
- Recipient task
- [What should they understand or do?]
- Evidence or observation
- [Why test this?]
- Decision to inform
- [Specific future choice]
- Hypothesis
- [Audience + change + expected behavior + reason]
- Experiment owner
- [Name or role]
Variants and controls
- Control subject
- [Exact text]
- Variant subject
- [Exact text]
- Single difference
- [Specificity, action, framing, or another]
- Preheader
- [Stable text or documented variable]
- Sender identity
- [Stable sender]
- Message body version
- [Exact approved version]
- Truth and condition check
- [How each subject is fulfilled]
- No variant uses false urgency, personalization, or implied relationship.
- Accessible characters and meaningful words appear early.
- Audience, timing, sender, and body are controlled.
Measurement plan
- Assignment method
- [Randomization and split]
- Required sample or run
- [Method and owner]
- Primary metric
- [Metric tied to the decision]
- Guardrail metrics
- [Delivery, complaints, unsubscribe, task failure, or another]
- Exclusion rules
- [Bots, duplicates, internal traffic, or another]
- Stopping rule
- [Written before launch]
- Segment analysis
- [Only pre-planned relevant groups]
- Data owner
- [Analyst or team]
Result and decision
- Run dates
- [Start and end]
- Variant counts
- [Delivered population]
- Primary result
- [Value and uncertainty]
- Guardrail result
- [Any harm or anomaly]
- Operational issue
- [Tracking or delivery concern]
- Decision
- [Adopt / Retain control / Follow up / Inconclusive]
- Scope of learning
- [Where it may apply]
- Next action and owner
- [Specific change]
How to use this template
- Define the recipient situation, email task, research evidence, and one wording decision the test will inform.
- Write truthful variants that differ on the chosen factor while keeping material conditions visible.
- Pre-register audience assignment, sample approach, primary metric, guardrails, exclusions, duration, and stopping rule.
- Quality-check the variants and message, then launch only after data, deliverability, and approval owners confirm readiness.
- Analyze the complete result, document uncertainty and harms, and record the scoped decision for future sends.
Form a useful hypothesis
State what audience uncertainty or motivation the variants test and why the result will change future work. “Question marks beat statements” is too broad to transfer safely. A better hypothesis connects the situation and wording choice: recipients awaiting an application decision may recognize a subject that names the status action more readily than one that names the program. Use prior research, support language, or campaign evidence instead of guessing that all recipients respond to curiosity. Choose one material difference—specificity, benefit framing, sender context, or action language—so the result can be interpreted.
Protect the comparison
Keep sender, preheader, message body, audience eligibility, timing, and delivery treatment stable unless the design explicitly tests them. Remove duplicates and ensure random assignment is implemented as planned. Estimate the required sample and analysis approach with someone qualified for the decision; do not repeatedly inspect a small test and stop when a preferred variant moves ahead. Define the primary metric, guardrails, minimum run, exclusion rules, and stopping condition before launch. Opens can be affected by privacy features and technical behavior, so include downstream signals that match the email’s real task.
Decide beyond the winning number
Review delivery, complaints, unsubscribes, click or task completion, segment differences, confidence, and operational anomalies. A subject may raise opens while confusing recipients or attracting people for whom the content is irrelevant. Record the exact variants and result window so later teams do not turn a contextual finding into a universal style rule. If the result is inconclusive, retain the honest baseline or run a properly designed follow-up; do not call a tiny numerical difference a winner. Archive the experiment with the final wording and what the team will change.
See the fields in context
Fictional example: workshop waitlist update
Pinewater Workshops, its list, and all results are invented.
- Hypothesis: People awaiting a place will recognize a status-specific subject more easily than a general program update.
- Control: “Pinewater workshop news.”
- Variant: “Your workshop waitlist status.”
- Control condition: The body and preheader are identical and explain that no place is guaranteed.
- Decision rule: Use task completion and complaint guardrails beside opens; do not stop early.
- Outcome: The fictional result is recorded as inconclusive, so the team retains the clearer task-specific subject without claiming a statistical winner.
Frequently asked questions
Is open rate enough to choose a subject line?
Usually not. Opens can be affected by technical privacy behavior and do not show whether the recipient understood or completed the intended task. Include delivery, complaint, unsubscribe, click, and task signals appropriate to the email.
How many variants should one test include?
Use the smallest set that answers the decision with available sample and analysis capacity. Several slight variants dilute the comparison and invite post-hoc storytelling. Two well-controlled options are often more useful than a large contest.
Can the test use urgency?
Only when the urgency is real, material, and supported in the email. State the actual deadline and consequence rather than manufacturing scarcity. A result obtained through misleading pressure should not become a reusable practice.
Does a winning variant work for every audience?
No. Preserve audience, lifecycle moment, sender, body, timing, and test date with the finding. Apply it beyond that context only with reason and further evidence.