Matched markets
A stable historical cousin.
Compare the treatment geography with one or more places that behaved similarly over time. The fit must be stable, not merely similar on average after large swings.
Incrementality testing for DTC & ecommerce
Causal evidence for media allocation
When reports add up to more sales than the business made, use a credible test to decide what deserves the next dollar.
Talk with EricThe question
Testing is for a genuine dispute: two or three channels claiming the same sale, a view-through channel whose credit feels too far from the conversion, or a budget decision that click reporting cannot settle.
The useful question is not whether a dashboard is “right.” It is whether a strategy is adding to the business—and what changes if you spend less, hold steady, or scale it.
Does this channel create incremental demand, or mostly record demand created somewhere else?
When a channel belongs in the mix, what can a levels test teach about marginal response and scale?
Did the intervention cause a meaningful business outcome under the conditions tested?
A strong fit
Multi-million-dollar scale and multiple meaningful channels are useful heuristics—not universal gates. The test should follow the decision and the available signal.
Better to start elsewhere
The next step may be cleaner data, a narrower attribution question, or waiting for a credible learning opportunity.
The comparison design
A treatment market is where media is introduced, paused, or changed. The question is what would have happened there otherwise.
Use prior sales to assess candidate markets before launching the test.
Change media exposure in selected geographies while protecting the test conditions.
Measure the outcome against a credible counterfactual—not a convenient dashboard read.
Translate the result, uncertainty, and next test into a bounded allocation move.
Matched markets
Compare the treatment geography with one or more places that behaved similarly over time. The fit must be stable, not merely similar on average after large swings.
Synthetic control
Combine multiple geographies into a comparison designed to resemble the treatment market’s historical behavior. It can be powerful, but may require more of the country and limit concurrent tests.
The practical answer
Back-test both approaches against history. When volume and capacity allow, more than one credible comparison can increase confidence in a result.
The engagement
Clarify the decision, KPI, available data, current reports, seasonality, promotions, channel definitions, and the feasible learning path.
Validate candidate comparisons, define treatment and control conditions, assess what the test can detect, and establish the decision rule before launch.
Check that media is running—or withheld—where intended. Identify exposure bleed, related-channel effects, and signals that need explanation while the test is in flight.
Deliver a decision-ready readout, understandable charts, underlying documentation, uncertainty, and the next action or learning plan.
Practical expectations
The fastest work starts with accessible, consistently labeled data. Foundation work is sometimes the project—not an administrative prelude.
Questions before committing
Maybe. A platform can accurately report the conversions it observed while still being unable to answer how many would have happened without the advertising. Testing is most valuable when that difference would change a material allocation decision.
Yes. When a channel clearly belongs in the mix, a levels test can be more useful: intentionally testing a lower, normal, or higher level of spend to learn about marginal response and possible diminishing returns.
Start with historical data. The right answer depends on which comparison is the more stable predictor for the proposed treatment markets and whether the business has enough volume and test capacity to support one or both approaches.
No. Last click, MTA, MMM, and lift tests observe different evidence and answer different questions. The benefit is a clearer explanation of the disagreement and a more defensible decision—not a forced single number.
A practical first conversation
Bring the channels, reports, and decision in front of you. We’ll determine whether a readiness diagnostic, a scoped incrementality test, or another measurement step is the practical next move.
Talk with Eric about incrementality testing