Parse
Work with usPricing
Sign inCheck your brand
Research/Which AI engine changes its answers the most?

Which AI engine changes its answers the most?

Nearly a tie. Ask the same question again and ChatGPT Search keeps 40.5% of its brand list while Google AI Mode keeps 41.5% — and the less stable engine changes month to month.

By Dimitry Apollonsky · August 29, 2026 · 10 min read

Repeat-answer brand-list overlap by engine
  • Google AI Mode41.5% (329,370 pairs)
  • ChatGPT Search40.5% (334,028 pairs)
Mean overlap between consecutive answers to the same prompt on the same engine, May 24 to August 22, 2026. Higher means steadier.
▸Contents
  • Neither engine keeps its brand list: overlap is about 41% on both
  • The #1 pick changed about 45% of the time on both engines
  • An identical brand list came back 5.6% of the time
  • ChatGPT Search replaced the entire list 1.3 times as often
  • The typical repeat answer keeps a quarter to half of the list
  • About 86% of repeat answers add at least one new brand
  • The most volatile engine flipped month to month
  • Waiting longer barely lowers overlap
  • Tech brand lists are the steadiest; apparel and real estate change most
  • A volatile prompt is volatile on both engines
  • One-brand answers are the least repeatable
  • What this study counts, and what it leaves out
  • The GEO takeaway
  • Get the data
  • Sources
  • Related research
Contents
  • Neither engine keeps its brand list: overlap is about 41% on both
  • The #1 pick changed about 45% of the time on both engines
  • An identical brand list came back 5.6% of the time
  • ChatGPT Search replaced the entire list 1.3 times as often
  • The typical repeat answer keeps a quarter to half of the list
  • About 86% of repeat answers add at least one new brand
  • The most volatile engine flipped month to month
  • Waiting longer barely lowers overlap
  • Tech brand lists are the steadiest; apparel and real estate change most
  • A volatile prompt is volatile on both engines
  • One-brand answers are the least repeatable
  • What this study counts, and what it leaves out
  • The GEO takeaway
  • Get the data
  • Sources
  • Related research

We analyzed 663,398 consecutive same-prompt, same-engine answer pairs to 17,039 organic prompts asked on both ChatGPT Search and Google AI Mode from May 24 through August 22, 2026.

40.5% vs 41.5%
repeat-answer brand-list overlap, ChatGPT Search vs Google AI Mode
45.3% / 45.0%
of repeat answers with a #1 pick in both runs changed it, by engine
5.6%
of repeat answers returned an identical brand list
37,025 of 663,398
1.3x
ChatGPT Search replaced the entire list more often
9.3% vs 7.1% of pairs

Neither engine keeps its brand list: overlap is about 41% on both

An answer's brand list is the set of distinct resolved brands it names directly. Overlap is the share of brands the two answers agree on: brands named in both answers divided by all brands named in either. On consecutive answers to the same prompt, mean overlap was 40.5% on ChatGPT Search (334,028 pairs) and 41.5% on Google AI Mode (329,370 pairs).

The gap between engines is 1.0 points. The gap between either engine and a stable answer is nearly 60 points. Which engine changes most is the wrong question: on average, roughly 3 of every 5 brands across the two answers appear in only one of them, on either engine.

Repeat-answer brand-list overlap by engine
  • Google AI Mode41.5% (329,370 pairs)
  • ChatGPT Search40.5% (334,028 pairs)

Takeaway

Both engines rewrite most of the brand list between runs. Treat any single answer as one draw, not the answer.

The #1 pick changed about 45% of the time on both engines

The #1 pick is the single brand an answer places at recommendation position 1. On pairs where both answers had exactly one #1 pick, it stayed the same in 54.7% of 266,788 ChatGPT Search pairs and 55.0% of 268,453 Google AI Mode pairs. It changed 45.3% and 45.0% of the time.

This matches our earlier pooled study, which found the top recommendation changed in 43.6% of consecutive pairs over a shorter window. The engines are indistinguishable on this measure: 0.27 points apart.

54.7%
ChatGPT Search pairs kept the same #1 pick
of 266,788 pairs with a #1 in both answers
55.0%
Google AI Mode pairs kept the same #1 pick
of 268,453 pairs with a #1 in both answers

An identical brand list came back 5.6% of the time

Across both engines, 37,025 of 663,398 repeat answers returned exactly the same brand list: 5.1% on ChatGPT Search and 6.0% on Google AI Mode. That is about 1 identical list in every 18 repeat answers.

SparkToro's 2,961-run study found identical brand lists in fewer than 1% of repeat runs, and the same list in the same order in fewer than 0.1%. Our bar is looser — we compare unordered sets of resolved brands, which merges spelling and naming variants — and the lists still differ 94.4% of the time.

5.6%
of repeat answers returned an identical brand list
37,025 of 663,398 pairs

Takeaway

If a report claims an engine gives the same brand list twice, ask how many runs it checked.

ChatGPT Search replaced the entire list 1.3 times as often

A full swap is a pair whose two answers share no brands at all. ChatGPT Search produced a full swap in 9.3% of pairs (31,194 of 334,028); Google AI Mode in 7.1% (23,434 of 329,370).

This is where the engines differ most. The average overlap is close, but ChatGPT Search reaches the extreme — an answer with zero carryover — 1.3 times as often. Google AI Mode also returned identical lists slightly more often (6.0% vs 5.1%). By tail behavior, ChatGPT Search is the engine that changes its answers the most, by a modest margin.

Share of repeat answers with no shared brands
  • ChatGPT Search9.3% (31,194 of 334,028)
  • Google AI Mode7.1% (23,434 of 329,370)

The typical repeat answer keeps a quarter to half of the list

The most common outcome on both engines is partial overlap: 33.0% of ChatGPT Search pairs and 35.2% of Google AI Mode pairs landed at 25–49% overlap, and about 26% of pairs on each engine landed at 50–74%.

Full agreement and full swap are both tails. The middle of the distribution is an answer that keeps a recognizable core and rotates the rest.

Distribution of repeat-answer overlap
Share of pairs per engine, in percent. ChatGPT Search: 334,028 pairs; Google AI Mode: 329,370.
No shared brands (0%)9.347.11
Identical list (100%)5.136.04
75–99%7.156.27
50–74%26.0626.4
25–49%33.0235.16
1–24%19.3119.01

About 86% of repeat answers add at least one new brand

A repeat answer added at least one brand its predecessor did not name in 85.7% of ChatGPT Search pairs and 85.4% of Google AI Mode pairs. On average, 47.2% of the brands in a ChatGPT Search repeat answer were new, and 45.5% on Google AI Mode.

Answers averaged 5.9 brands on ChatGPT Search and 5.7 on Google AI Mode, so a typical repeat answer carries about 2.8 and 2.6 brands, respectively, that the previous answer did not mention. The change is not brands dropping out of a fixed list; it is a rotating set of near-equal candidate brands.

85.7%
of ChatGPT Search repeat answers added a new brand
85.4%
of Google AI Mode repeat answers added a new brand
47.2% / 45.5%
of a repeat answer's brands were new, by engine

The most volatile engine flipped month to month

On a fixed panel of 13,640 prompts with repeat answers on both engines in all three months, ChatGPT Search was clearly less stable in June (36.1% overlap vs 42.0% for Google AI Mode — almost 6 points). In July the ranking flipped: ChatGPT Search 42.0%, Google AI Mode 40.8%. In August (through the 22nd) they tied: 42.0% vs 42.1%.

A one-month engine-stability ranking would have named a different winner each month. Studies that pick a most-volatile engine from a single window are measuring that window, not the engine.

Repeat-answer overlap by month, fixed 13,640-prompt panel
Mean overlap in percent. August covers the 1st through the 22nd.

Takeaway

Re-run any engine-stability comparison before acting on it. The ranking itself is unstable.

Waiting longer barely lowers overlap

Pairs whose two answers were 0–1 days apart overlapped 39.2% on ChatGPT Search and 43.5% on Google AI Mode. At 8–14 days apart, both engines sat near 37% (37.0% and 37.0%). Pairs 4–7 days apart — the most common spacing — overlapped 41.0% and 41.0%.

If answers drifted over time, longer gaps would score much lower than same-day repeats. They barely do. Most of the change happens between any two runs, immediately; Google AI Mode's same-day advantage (4.3 points) is the one place the engines clearly separate, and it fades within a week.

39.2% / 43.5%
overlap 0–1 days apart, ChatGPT Search / Google AI Mode
41.0% / 41.0%
overlap 4–7 days apart
37.0% / 37.0%
overlap 8–14 days apart

Tech brand lists are the steadiest; apparel and real estate change most

Across 28 industry groups with at least 2,000 pairs per engine and 100 prompts, Information Technology prompts had the steadiest brand lists (46.1% overlap) and Clothing and Apparel the least steady (36.3%), with Real Estate (37.2%) and Consumer Goods (37.5%) close behind.

Google AI Mode was the steadier engine in 17 of the 28 groups and ChatGPT Search in 11. The industry you compete in moves overlap by up to 10 points — a bigger effect than which engine is asked.

Repeat-answer overlap by industry group
Groups with at least 2,000 pairs per engine and 100 prompts; 11,284 of 17,039 matched prompts map to a group.
Information Technology46.1247.6644.5526,578
Apps45.8845.2446.55,073
Software45.4246.6844.1431,637
Gaming45.345.0645.545,507
Internet Services45.0846.4343.7511,464
Media and Entertainment44.946.5343.1610,388
Hardware44.5745.8643.2113,240
Consumer Electronics44.5146.1842.6610,047
Data and Analytics43.3442.5544.1513,147
Sales and Marketing43.142.5443.639,399
Financial Services42.9143.2242.6136,124
Administrative Services42.7842.0143.568,043
Content and Publishing42.4842.4842.485,969
Sports42.0942.2541.929,537
Travel and Tourism4242.3941.636,700
Community and Lifestyle41.9440.3843.549,188
Health Care41.7340.9542.5116,636
Education41.5741.4441.719,628
Artificial Intelligence40.7841.4340.1218,793
Food and Beverage40.6540.0841.2512,081
Transportation40.0639.2940.878,277
Professional Services39.8138.1941.4317,338
Design39.6238.9140.364,593
Commerce and Shopping38.9837.740.2923,137
Manufacturing38.4737.5939.395,376
Consumer Goods37.5437.6237.4613,528
Real Estate37.2435.9338.6211,469
Clothing and Apparel36.3335.9236.794,845

A volatile prompt is volatile on both engines

For the 14,101 prompts with at least 10 pairs on each engine, per-prompt overlap on ChatGPT Search correlates 0.61 with per-prompt overlap on Google AI Mode. 70.8% of these prompts sit on the same side of the median on both engines: 35.4% volatile on both, 35.4% steady on both.

Instability follows the question more than the engine. A question with many near-equal candidate brands is unstable on every engine; switching engines will not settle it.

0.61
correlation of per-prompt overlap across engines
14,101 prompts with 10+ pairs on each engine
70.8%
of prompts on the same side of the median on both engines

One-brand answers are the least repeatable

When the smaller answer in a pair named just one brand, overlap averaged 31.8% on ChatGPT Search and 28.6% on Google AI Mode. When both answers named 10 or more brands, overlap rose to 43.1% and 46.4%.

A short answer looks decisive but is the least likely to come back the same way. Long brand lists keep a stable core and rotate the rest; a one-brand answer has the least carryover of any answer size.

Overlap by shortlist size
Mean overlap in percent by the smaller answer's brand count.
10+ brands43.0646.41
6–9 brands41.6444.34
4–5 brands41.7842.94
2–3 brands38.238.57
1 brand31.8328.62

What this study counts, and what it leaves out

A pair is two immediately adjacent answers to the same organic prompt on the same engine, kept when both name at least one resolved brand; every cross-engine number uses the 17,039 prompts (of 17,560 with pairs) that had pairs on both engines. We computed the headline two independent ways — an array-intersection build and a row-level rebuild of the same pairs — and they agree to four decimals. Removing the matched-prompt restriction moves overlap by at most 0.04 points. Excluding pairs where either answer named a single brand moves it by under 1.5 points (40.9% and 42.4%).

The #1 pick is defined in both answers for 79.9% of ChatGPT Search pairs and 81.5% of Google AI Mode pairs; the rest have no single position-1 brand. Answers are observed on a regular cadence, so a pair is a repeat across runs, not a controlled same-minute re-ask. Brand strings are resolved to root brands, which merges spelling variants but can split unresolved names. The window ends August 22, 2026 because the brand-naming layer of the data thins after that date.

17,039
prompts with repeat answers on both engines
of 17,560 prompts with any pair
79.9% / 81.5%
of pairs had a defined #1 pick in both answers, by engine
40.9% / 42.4%
overlap when both answers name 2+ brands, by engine

The GEO takeaway

Both engines rewrite most of the brand list between runs, the #1 pick changes about 45% of the time, and the monthly most-volatile ranking did not survive the quarter. There is no stable engine to optimize for and no volatile engine to ignore.

Measure your brand's appearance rate: the share of many repeated answers that name you, per engine, over a window of a fixed length. A brand in 8 of 10 runs is visible; a brand in the latest screenshot may not be. Track both engines — brands that drop out of one answer usually come back, and the rotation is how competitors enter answers. Rerun the comparison each quarter with this article's method before believing any engine-stability claim, including ours.

Get the data

Dataset CSVThe metrics behind every figure in this report.

Sources

  1. SparkToro: AIs are highly inconsistent when recommending brands or products · accessed 2026-08-29
  2. Search Engine Journal: AI recommendations change with nearly every query · accessed 2026-08-29
  3. Conductor: Why intent type predicts AI output consistency · accessed 2026-08-29
  4. Schulte, Bleeker, Kaufmann: Don't measure once — measuring visibility in AI search (GEO) · accessed 2026-08-29

Related research

Does AI recommend the same brand when you ask again?
Only about six in ten times. The top recommendation stayed the same in 90,817 of 161,023 consecutive same-prompt, same-engine answer pairs, or 56.4%.
AI citation volatility by industry: one-shot checks miss the signal
Ask the same AI question again and the cited sources usually change. Across repeated runs, ChatGPT answers shared only about 21% of their cited sources, and every industry showed high churn.
Does AI change its mind when comparing brands?
Sometimes. On the same prompt and criterion, one AI engine picked a different brand winner in 383 of 2,438 repeated matchups.
Does AI keep the same recommendation when its sources change?
Almost half the time. The top recommendation stayed the same in 12,243 of 25,883 back-to-back answers to the same prompt where no cited page repeated.
ChatGPT vs Google: same winner, different shortlist
The popular line is that AI engines disagree about who wins. They don't. Across 1,655 buyer categories ChatGPT and Google AI Overviews pick the same #1 brand 93% of the time. What they disagree about is everyone else on the list.
The AI engine report card
Close on which brands to name, far apart on everything else. On the same 17,083 prompts, 7 of 10 report-card metrics split by 2.6 times or more between ChatGPT Search and Google AI Mode.
Does question wording change which brands AI names?
Yes, a lot. Two phrasings of the same buyer question shared 11.7% of named brands on average; the identical phrasing re-asked one to three days later shared 38.0%.

About this research

Dimitry Apollonsky

Founder, Parse

I built Parse to track where AI answers really come from: the sources they cite and the brands they name. DM me on LinkedIn to talk shop.

Measure AI answers as a rate across runs, not a screenshot.

Run a free check against live AI answers — no account needed.

Parse

See where your brand stands in AI recommendations.

Products

  • Brands
  • Markets
  • Integrations
  • Work with us
  • Pricing
  • MCP

Resources

  • Research
  • Methodology
  • Blog

© 2026 Parse. All rights reserved.

LegalPrivacy PolicyTerms of Service