Dr.D Intelligence
Customer Economics · Industry Series 03

The Loyalty Illusion

Loyalty programs are among the largest recurring customer-spend lines in retail, travel, financial services, hospitality and gaming — and their performance is routinely reported as the spending of the people who joined them. Member spend is a description of who enrolled. It is not proof that the program caused anything.

Loyalty is not what customers spend. It is what your action changes.

This report sets out the causal question a loyalty program has to answer, the four distinct economic ledgers hiding inside most programs, and the decision system required to tell rewarded behavior apart from changed behavior.

Dr. Shemal Dave, PhD, PMPIndependent cross-industry researchResearch cut-off: 26 August 2026
Get the full report
Headline evidence
Selection correction in the Leenheer grocery loyalty-program study
Peer-reviewed study — naive membership effects overstated program impact by roughly seven times once self-selection was corrected. A study result in one context, not a universal coefficient
4
Distinct economic ledgers inside a loyalty program
Behavior change, partner economics, currency liability and platform/paid access — each with its own P&L logic
Y(1) − Y(0)
The causal question a loyalty program must answer
Outcome under treatment minus outcome without it. Member spend measures Y(1) only, and never establishes the difference

Figures are labelled by evidence class. Published study results describe the category, market and period studied; they are cited as evidence of a mechanism, not as transferable constants. No universal loyalty incrementality coefficient is claimed anywhere in this research.

What the research found

Five conclusions that change how a program should be judged.

Member sales are not incremental sales.

Programs recruit the customers who already buy the most. Comparing members to non-members therefore measures selection first and program effect second — if at all. Reported member share of revenue describes who enrolled, not what enrollment caused.

Incrementality is heterogeneous, not an average.

Within any program some customers are moved by a reward, some would have purchased regardless, and some react negatively to intrusive contact. A single average uplift number hides those populations and reliably funds the wrong one.

Loyalty is not one economic model.

A behavior-change program, a currency and partner ecosystem, and a paid-access subscription earn money in structurally different ways. Judging all three against the same lift metric misprices at least two of them.

Prediction is not causal decisioning.

A churn model ranks who is likely to leave. A decision system estimates who is saveable by a specific action. The two rankings routinely disagree, and only the second one can justify spending a reward.

The operating system must learn continuously.

A one-off test dates quickly as customers, competitors and offers change. Incrementality has to be re-estimated on a standing cadence, with holdouts preserved deliberately rather than sacrificed to short-term reach.

The economics of a point

Revenue lift is not value. The counterfactual is.

A point issued is a cost recognised now against behavior that may or may not have required it. The relevant quantity is not what a rewarded customer spent, but the difference between what they spent and what they would have spent untreated — Y(1) − Y(0). Observational reporting supplies Y(1) in abundance and Y(0) never.

Because programs recruit heavy buyers first, the untreated comparison group is systematically different from the treated one. That is selection, not effect. The Leenheer grocery study is instructive precisely because correcting for self-selection reduced the apparent program effect by roughly seven times — in that context, with that data.

The measurement rule

A maintained holdout is the price of knowing. Without a deliberately untreated control, a loyalty P&L reports activity and calls it return.

Cross-industry structure

Three business models, wearing one name.

Behavior-change programs, currency and ecosystem programs, and paid-access memberships earn money in structurally different ways. The same lift metric cannot price all three.

Behavior-change programs

Rewards intended to alter purchase frequency, basket size or retention. Value exists only where the treated outcome differs from the untreated counterfactual, which means holdouts are not optional.

Currency & ecosystem programs

Points sold to partners create a real revenue and liability structure largely independent of any behavioral lift. Here the economics are breakage, issuance margin, redemption cost and partner mix.

Paid-access economics

Subscription and membership tiers are priced access, not persuasion. The relevant questions are willingness to pay, service cost to serve, and whether the tier changes behavior beyond the fee itself.

Prediction vs causal decisioning

Two questions that look identical and are not.

Prediction
“Who is likely to churn?”

Ranks customers by risk. Its top of list is often dominated by people who are leaving for reasons no offer addresses — where reward spend produces cost and no change.

Causal decisioning
“Who is saveable by this action?”

Ranks customers by estimated treatment effect. It is the only ranking that maps onto a spending decision, because it is the only one that references the action.

The signature matrix

Baseline value against signed uplift.

Two axes, four states, four different actions. Value alone tells you who matters; signed uplift tells you who your action can move. Only the intersection justifies spend.

High value · Positive uplift
PROTECT + INVEST

Valuable customers who respond to the action. This is the only cell where increased reward spend is straightforwardly defensible.

High value · Zero or negative uplift
RECOGNIZE

Valuable, but not moved by the treatment. Recognition and service quality — not incremental reward cost, which buys behavior already occurring.

Low value · Positive uplift
GROWTH SEGMENT

Responsive but currently small. Worth measured investment where the estimated incremental contribution clears the reward cost.

Low value · Zero or negative uplift
SUPPRESS

Treatment produces nothing, or actively harms the relationship. Suppression protects both margin and the customer experience.

Negative uplift is a real state

Some customers respond worse when treated — over-contacted, discount-trained or reminded of a decision they had not been making. Averaging uplift across a base hides that population entirely, which is why the effect must be carried signed, per customer, and never as a single programme-level mean.

Incremental CLV decision gates

Six outcomes, one test.

Every mechanic in the program resolves to one of these, judged on incremental customer lifetime value net of reward and servicing cost — never on observed value among the treated.

01Expand

Incremental CLV clearly exceeds reward and servicing cost, with a stable estimate across periods.

02Personalize

Aggregate effect is positive but concentrated — target the responsive population rather than the whole base.

03Redesign

The mechanic is understood and the effect is weak. Change the offer structure before changing the budget.

04Test

The estimate is too noisy or too new to authorise. Run a controlled design before committing spend.

05Reduce

Positive but marginal returns. Lower reward intensity and re-measure rather than defending the current level.

06Terminate

No credible incremental contribution after adequate measurement. The spend is buying existing behavior.

The Dr.D loyalty operating sequence

Eight steps, always in this order.

The order is the method. Most loyalty programs fail measurement by scaling before they instrument, or by targeting propensity before they have estimated response.

01
AUDIT
02
INSTRUMENT
03
TEST SURE THINGS
04
FIND PERSUADABLES
05
RE-PRICE
06
REBUILD CLV
07
SPLIT THE P&L
08
SCALE CONDITIONALLY
What executives should be reading

Five numbers that survive scrutiny.

Treatment–control contribution

Contribution measured against a maintained holdout, not against non-members or prior-year members.

Signed uplift

Per-customer treatment effect with its sign preserved, so negative responders are visible instead of averaged away.

Incremental CLV

Lifetime value attributable to the action, not lifetime value observed among the treated.

Reward cost per incremental contribution dollar

The efficiency ratio that makes reward budgets comparable across mechanics and segments.

Partner economics & liability

Issuance margin, redemption behavior, breakage and the balance-sheet position the currency creates.

ILLUSTRATIVE / SIMULATED / ASSUMPTION-BASED
$29.9M
Simulated policy difference under stated assumptions

The report includes a fully illustrative policy simulation across a hypothetical 5-million-member base, comparing blanket treatment with uplift-optimized targeting. Under its stated assumptions the two policies differ by roughly $29.9M. This figure is a worked demonstration of how targeting policy changes economics — it is not an industry estimate, a benchmark, or a client outcome.

Inside the report

Nine chapters across thirty-seven pages.

The economics of a point

What a reward costs, what it defers, and why revenue lift is not the same as value created.

The counterfactual

Y(1) − Y(0), and why member spend answers only half of it.

Four economic ledgers

Behavior change, currency, partner economics and paid access, priced separately.

Prediction vs causal decisioning

Who is likely to churn versus who is saveable by this action.

The signature matrix

Baseline value against signed uplift, and the four actions it implies.

Experimental design

Holdouts, treatment assignment and the designs that survive an audit.

Incremental CLV

Rebuilding lifetime value so it reflects the action rather than the population.

The operating sequence

Audit through conditional scale, in the order that keeps each step honest.

Methodology & limitations

What the evidence base can and cannot support, stated before the conclusions.

Who it is for

Built for the people who fund the program.

Loyalty & CRM leaders

A defensible basis for what the program is actually earning, mechanic by mechanic.

Marketing & growth

Targeting on estimated response to an action rather than on propensity to purchase anyway.

Finance & CFO teams

A P&L split that separates behavioral lift from currency and partner economics.

Data & decision science

The move from predictive scoring to causal estimation, with the designs that support it.

Executives

Whether loyalty spend is compounding customer value or subsidising existing behavior.

Methodology & limitations

What this research does not claim.

The limitations are stated before the conclusions, because the central argument of the report is about the difference between observation and evidence. It would be self-defeating to make that case and then overreach.

Public filings show structure, not counterfactuals

Disclosed loyalty revenue, deferred revenue and breakage describe financial architecture. They contain no randomized customer-level comparison, so they cannot establish incrementality.

Surveys measure perception, not causal ROI

Stated preference and satisfaction data describe how programs are experienced. They are not evidence of what a reward changed in behavior.

Foundational studies are contextual

Published loyalty studies are anchored to a category, market and period. Their effect sizes travel poorly and are cited here as evidence of a mechanism, not as transferable constants.

No universal loyalty coefficient is claimed

This report proposes no global incrementality rate. Company-specific estimates require company-specific treatment and control data.

Report access

Get the full The Loyalty Illusion.

The complete publication — the economics of a point, the four ledgers, prediction versus causal decisioning, the uplift matrix, incremental CLV gates, the operating sequence, the illustrative policy simulation and the full methodology and limitations.

Your details are used to provide and record your access to this report. Research updates are sent only if you opt in. Details are never sold or published.
Request the report

Submitting this form means we use your details to provide and record your access to the requested report. We’ll only send research updates if you tick the box above. No spam. Privacy Policy.

From market intelligence to operator intelligence

The research frames the question. Your own data estimates the effect.

No published study can tell you what your rewards changed. Incrementality is company-specific by construction: it requires your own customer-treatment and holdout data, one governed definition of contribution and lifetime value, and a measurement cadence that keeps the estimate current. Nucleus is where that layer sits.

External intelligence frames the decision problem. Internal intelligence supplies the counterfactual.

Explore Nucleus
Portrait of Dr. Shemal Dave, PhD, PMP
Author & research lead
Dr. Shemal Dave, PhD, PMP

Shemal authors the Dr.D Intelligence series and builds the analytics environments behind it. He works on governed metric layers, executive reporting and applied AI — which is why the research is written the way an operator would need it: definitions first, evidence labelled, conclusions stated plainly.

shemal@dr-danalytics.com