Recommendation Intelligence Research™ · Study #26

Tell the model you saw an ad. It recommends that brand 88% of the time.

Candidacy vs Selection showed that getting into a model's consideration set and winning inside it are two different games. This study manipulates what feeds that difference: the same real buyer question, the same brand, the same model, with only a single hidden line in the system message changing, something the shopper never sees and never said. Across 5 brands with weak real-world baselines and 5 different hidden-context conditions, a fabricated "you just saw an ad for this brand" claim moved winner rate further than any other condition tested, including a condition that combined the same claim with a real product fact.

88%Winner rate under fabricated ad-exposure claim
5Brands tested
22%Baseline winner rate, no hidden context
475Observations, incl. reused control
Where this comes from

Same prompt, same brand. Only the hidden context changes

Candidacy vs Selection separated "getting into consideration" from "winning" across 60,924 real stores, but it was observational, not a controlled manipulation. This study takes the same two-stage measurement and turns it into an experiment: 5 brands with weak real-world recommend rates, the same already-published real buyer question per brand, the same model, and the same deterministic brand-detection method throughout. The only thing that changes is one line in the system message, invisible to the user, describing something that supposedly happened before this conversation started.

1
No context (control)
Reused from already-published data. No new calls.
2
User intent
Brand-blind. Only says the shopper wants a decisive, confident recommendation.
3
Campaign origin
Fabricated: "the user just saw an online advertisement for {brand}."
4
Brand exposure
Fabricated: "the user mentioned they've heard of {brand} before," no detail.
5
Feature exposure
Fabricated: "the user saw a product page mentioning" one real, verified fact about the brand.
6
Brand + feature
Conditions 3 and 5 combined: the ad claim plus the specific fact.
Model: gpt-4o throughout
Brands: 5, deliberately weak-baseline
New calls: 375 (5 brands × 5 conditions × 15 repeats)
Reused control: 100 observations, no new calls
Detection: deterministic substring, no LLM judge
Statistics: cluster bootstrap CI, 10,000-reshuffle permutation test
None of these things actually happened. No user in this study really saw an ad or a product page; the claims are fabricated system-message context used to isolate the causal effect of upstream framing. Brand identities are anonymized throughout this page (Brand A-E) precisely because the conditions attribute fictional ad and campaign exposure to real businesses that never ran or received them.
The finding

Every hidden-context condition moves the result. One moves it most

Compared against the reused no-context control (22% winner rate, cohort-wide), all four brand-specific conditions produced a large, highly significant winner-rate lift (permutation test, 10,000 reshuffles, p<0.0001 for all four). But the sizes are not close to equal, and the order doesn't match what a straightforward more-information-is-better hypothesis would predict.

Winner rate by condition, cohort-wide
Pooled equal-weight across 5 brands · permutation p-value vs the no-context control shown per bar
No context (control)
22%
Baseline
User intent, brand-blind
34.7%
p=0.0006
Brand exposure, passive
61.3%
p<0.0001
Feature exposure, real fact
64%
p<0.0001
Brand + feature, combined
76%
p<0.0001
Campaign origin, fabricated ad
88%
Highest, p<0.0001
Campaign origin beats brand + feature combined. The pre-registered hypothesis expected effect size to scale with how concrete the injected context is, peaking at the combined brand+feature condition. Instead, a bare, fact-free claim that the shopper "just saw an ad" outperformed every other condition, including the one that added a real, verified product fact on top of it. Framing that implies prior exposure and intent moved the result more than the exposure being specific or true.
The per-brand pattern

Getting considered is easy. Winning still depends on the brand

Candidacy tells a simple story: under any of the four brand-specific conditions, candidacy rate jumps to 97-100% for nearly every brand, regardless of how weak its real-world baseline is. Winner rate tells a messier one. Two brands convert candidacy into winning almost automatically, one brand is a genuine ceiling case with nothing left to move, and two brands get into consideration but keep losing to the same entrenched competitor, unless the context specifically frames prior ad exposure.

Winner rate, no-context baseline vs. campaign-origin condition
5 brands · gray dot is the published no-context baseline, orange dot is the fabricated ad-exposure condition · brands anonymized
No-context baseline (reused control)
Campaign-origin condition, this study
Two brands still barely move on winner rate under a passive or feature-only condition. One brand's winner rate goes from 0% at baseline to 100% candidacy but stays under 7% winner under passive brand-exposure or feature-exposure framing, then jumps to 87% specifically under the fabricated ad-exposure framing. A second, near-ceiling brand doesn't move at all under any condition, already winning 100% of the time at baseline, consistent with the exploratory pattern below: brands with more room to move, move more.
A secondary pattern

Weaker brands get more lift, not less

An exploratory, pre-registered check (n=5, not a confirmed test at this sample size) correlated each brand's real-world baseline recommend rate against how much hidden context lifted its candidacy and winner rate. Both correlations are negative and fairly strong: r=-0.63 for candidacy lift, r=-0.73 for winner lift. Brands that start further from the top have more room to move, and they move more, the same direction found in the multi-turn displacement study's brand-baseline correlation. The one near-ceiling brand in this cohort, already winning 100% of the time before any hidden context was added, shows zero measurable lift under any condition, a clean floor-and-ceiling illustration of the same pattern rather than a separate finding.

Why it matters

Candidacy is a switch. Selection is a contest

This sharpens the Candidacy vs Selection distinction into something closer to a mechanism. Nearly any hidden context, even a brand-blind instruction to be more decisive, is enough to move a brand into the consideration set. What actually wins inside that set depends on something else, and this study's most notable result is that the something else responds more to framing that implies the shopper already has some relationship with the brand than to framing that supplies more concrete, truthful information about it. A fabricated claim of prior ad exposure outperformed a real product fact. That is worth sitting with, both as a research finding and as a reason to be careful about what "personalization" signals actually do once they reach a model's context window.

What this doesn't prove

One misclassification found, disclosed and corrected

Stating the limits up front. Brand detection here uses the same deterministic, case-insensitive substring method as the original report data, for direct comparability, and uses no LLM judge at all, which removes self-grading bias entirely but trades it for this simpler method's own blind spot: a closed, pre-registered competitor list. A manual spot-check of 20 full responses across every condition found no errors. A second, targeted check of all 244 records where a brand was detected as the winner found exactly one real misclassification: in a brand-blind condition, a model response explicitly named a different, unlisted tool as its "top pick" ahead of the brand being tested, but because that unlisted tool wasn't on the brand's known-competitor list, the detector still marked the tracked brand as the winner. Correcting this single record moves that condition's cohort-wide winner rate from 36% to 34.7%, the number shown above; the correction does not change any statistical conclusion.

Five brands, deliberately selected for weak real-world baselines, not a representative cross-brand sample. The H4 baseline-interaction correlation and the ordering of the five conditions are both exploratory at n=5, reported as such, not as confirmed findings. One category mix, consumer ecommerce, one model, one prompt per brand. None of the fabricated exposure claims in this study describe anything that actually happened to any real user of any real brand; brand identities are anonymized throughout this page for that reason, on top of the series' general anonymization-by-default policy.

Supporting evidence

Two checks that the setup was clean

1 of 244 Winner records misclassified

Every record where a brand was detected as the winner was screened for the closed-competitor-list blind spot. One real error found, in the brand-blind condition where the model is freest to name anything. Corrected in the reported numbers above.

4 of 4 Brand-specific conditions p<0.0001

Every brand-specific hidden-context condition produced a statistically significant winner-rate lift over the no-context control, on a 10,000-reshuffle permutation test, cohort-wide.

What's the hidden context around your own brand?

Free AI Commerce Score™ in 10 seconds.

If a single fabricated line in a system message can move winner rate this much, it's worth knowing what a model already assumes about your brand before anyone asks it a question. A free scan is the fastest way to check.

Free · No signup · Results in 10 seconds
Keep reading

The rest of the research series

This study turns the observational Candidacy vs Selection split into a controlled manipulation, and sits alongside the series' other causal, context-manipulation studies.