CacheMagpie logoCacheMagpie· Community Library

Email marketing tests and benchmarks

Email is the only lifecycle channel where almost everything is testable, which is why most programs test the least consequential parts of it. Audience, trigger and cadence decide outcomes far more than subject line wording, and the largest untested surface in nearly every program is transactional mail. Test who receives a message and when before you test what it says.

  • Every testable surface in an email program, from audience to footer
  • Test ideas for each lifecycle stage, linked to the matching stage page
  • Deliverability and consent constraints that shape what you can run
  • How fast email learnings fade, and graded evidence from real programs

What is actually testable in email

An email test is only clean when you know which of these you changed. Most inconclusive tests changed three at once.

Audience and trigger

Who is eligible, what event starts the message, and who is suppressed. This is the highest leverage surface in the channel and the least often treated as a variable.

Timing and cadence

Send time, delay after the trigger, spacing between messages in a sequence, and how many messages the sequence contains.

Sender and envelope

From name, reply-to, subject line and preheader. Cheap to test, quick to read, and the surface where effects decay fastest.

Body and structure

Length, one idea against several, image weight, plain text against template, and where the primary action sits.

Offer and value

Incentive size and type, non-price levers such as shipping, guidance or reassurance, and whether an offer appears at all.

Transactional content

Confirmations, receipts, shipping notices and alerts. The most opened mail you send and, in most programs, the only mail nobody has ever tested.

The anatomy of a testable email

01audience and trigger

Who is eligible and what event starts the send.

02from name

Brand, person, or brand plus person.

03subject line

Claim, question, specificity, length.

04preheader

Extension of the subject, or a second claim.

05send time

Absolute clock time or time since the trigger.

06opening line

Reason for the message, stated or implied.

07body

One idea or several, image weight, length.

08primary action

Placement, wording, and how many there are.

09offer

Present or absent, price or non-price.

10footer

Preference options, frequency choice, unsubscribe clarity.

Every slot above is a variable you can hold constant or change. A clean email test moves one of them at a time.

Email test ideas by lifecycle stage

Each stage below has its own page with the full set of levers, guard metrics and graded evidence.

  • Test the series length against a single message, read on first purchase rather than opens.
  • Test holding the incentive back to the last message so it only reaches people who did not act.
  • Test prompts triggered by real progress against a fixed schedule.
  • Test one action per message against a checklist.
  • Test the delay before the first reminder across the full recovery window.
  • Test naming a specific blocker, such as shipping or returns, against a generic reminder.
  • Test reducing cadence for disengaged contacts, read on revenue per contact.
  • Test event triggered messages replacing a recurring campaign.
  • Test asking what went wrong against offering a reason to return.
  • Test a hard stop after two touches, with complaint rate as the primary read.

Constraints that decide what you can test in email

Email has the fewest hard limits of any lifecycle channel, but the soft limits are real and they punish volume.

  • Reputation is shared across your whole program, so a bad campaign degrades every later send from the same domain.
  • Consent and suppression rules under GDPR and similar regimes decide who is eligible, which constrains audience tests before creative ones.
  • Bulk sender requirements from the major inbox providers set complaint rate ceilings and one click unsubscribe expectations.
  • Open rate is no longer a clean metric where mail privacy pre-fetches images, so it cannot be the deciding read.
  • Rendering varies widely between clients, so a design test needs a rendering check before a results check.

How fast email learnings fade

Email learnings are the most durable in the lifecycle stack, but not uniformly. Structural findings about audience, trigger and cadence hold for years, because they describe how people behave rather than how a format feels.

Wording and format findings fade faster. A subject line style that stands out does so because it is unusual, and once it is common it is not unusual. Treat a subject line win as valid for a season, and re-run it rather than enshrining it.

Anything tied to inbox provider behaviour, tabs, clipping, image handling or open measurement, can change without notice. Keep the source date of a learning in view before you copy it.

Email is the slowest channel to burn out, but tests that rely on a new format or a surprise subject line still fade as the list gets used to them. Reads here are worth revalidating once a year.

Format learnings from the corpus

  • Subject line length and clarity are the most repeated test in the corpus, and short and specific tends to beat clever.
  • Send time tests read smaller than most teams expect once the audience is already engaged.

How to run a email test

Randomise on the contact and hold the audience definition constant across arms. Read on the outcome the message exists to produce, not on the open, and let a sequence finish before you judge it.

Set the sample size and the read date before the test ships, not after you have seen the first hour of data. If the audience cannot reach the sample you need inside a sensible window, change the test rather than the standard.

A test card template for Email

Hypothesis
Because contacts in [segment] respond to [reason], changing [one variable] from [control] to [variant] will increase [outcome] within [window].
Success criteria
The downstream outcome, measured on everyone assigned rather than on openers or clickers.
Guard metric
Unsubscribe rate, complaint rate and inbox placement.

Free to copy and use in your own program. Fill the brackets from a test you have read in the library.

Guard metrics: what a win must not cost

Email wins that cost reputation are borrowed, not earned.

  • Complaint rate, which the major providers treat as a hard ceiling.
  • Unsubscribe rate per contact rather than per send.
  • Inbox placement and bounce rate on the sending domain.
  • Engaged list size at the end of the test period.

Evidence from real Email programs

The library holds 106 graded tests run by Email teams: 97 report a win for the variant, 4 a loss, and 5 no clear difference.

Winning tests in this cell read in line with the rest of the library, across 65 comparable tests.

Direction, grade and source are free. Exact figures open up once you sign in.

What a winning email test looks like

A winning email test moves the outcome the message exists to produce, holds unsubscribes and complaints flat, and comes with a documented audience and window so someone else can repeat it.

The wins worth keeping are structural. If the result depends on a specific wording, expect to re-run it within a year.

Frequently asked questions

Does Email messages still work in 2026?

The library holds 106 graded tests here, and 97 of them report a win for the variant against 4 losses and 5 with no clear difference. That is enough to say the direction still holds, and not enough to promise a number for your own program.

What is the most credible email test in the library?

First-name in welcome subject line lifts opens. Personalizing the welcome subject with the guest's first name lifted opens an exact amount with no downstream drag. The win compounded into a small first-booking uplift.

What tends to fail in Email messages?

Maximizing promo-email revenue this quarter directly trades against churn, so frequency is a lifetime-value decision, not a revenue one. Losing tests are kept in the library on purpose, because knowing what did not move is as useful as knowing what did.

How fresh are these email learnings?

The newest record in this cell is from 2026 and the oldest from 2012. 1 record has been reviewed and vouched for by the CacheMagpie editor.