Does AI criticize its top recommendation?
Sometimes. Of 150,193 top-ranked brand recommendations, 5,460 came with a specific negative claim about the brand in the same AI answer.
AI criticized 3.6% of its top recommendations
A specific negative claim appeared with the top-ranked brand in 5,460 of 150,193 analyzed recommendations, or 3.6%. The other 144,733 top recommendations, or 96.4%, had no specific negative claim about that brand. A negative claim is a concrete drawback, limitation, risk, or unfavorable comparison stated in the answer.
First place is not always an unqualified endorsement. Track recommendation rank and the claims around the brand separately so a high rank does not hide a repeated limitation.
Takeaway
ChatGPT Search criticized top picks 5.5 times as often
ChatGPT Search included a negative claim in 4,461 of 67,472 top recommendations, or 6.6%. Google AI Mode did so in 999 of 82,721, or 1.2%. ChatGPT Search's rate was 5.5 times as high, a gap of 5.4 percentage points.
A combined criticism rate hides a large engine difference. Audit the same priority prompts on both engines and keep separate baselines for each.
ChatGPT Search criticized top picks 5.5 times as often
Criticism rate by engine
- ChatGPT Search6.6%4,461 of 67,472
- Google AI Mode1.2%999 of 82,721
Takeaway
Most criticism did not make the whole recommendation negative
Of the 5,460 criticized top recommendations, 4,912, or 90.0%, did not carry negative overall sentiment. Across all 150,193 top recommendations, 560, or 0.37%, had negative overall sentiment; 2,792, or 1.9%, were reluctant; and 148, or 0.10%, explicitly said the brand was not recommended for the need.
A sentiment label alone misses most specific drawbacks. Keep negative claims, overall sentiment, reluctance, and explicit rejection as separate fields because they answer different questions.
- included a specific negative claim
- 3.6%included a specific negative claim5,460 of 150,193
- had negative overall sentiment
- 0.37%had negative overall sentiment560 of 150,193
- were reluctant recommendations
- 1.9%were reluctant recommendations2,792 of 150,193
- explicitly said the top brand was not recommended
- 0.10%explicitly said the top brand was not recommended148 of 150,193
Criticism became more common lower in the recommendation list
Negative claims appeared with 5,460 of 150,193 first-place recommendations, or 3.6%; 17,989 of 259,382 brands ranked second or third, or 6.9%; and 32,388 of 377,610 brands ranked fourth or lower, or 8.6%. The fourth-or-lower rate was 2.4 times the first-place rate.
Rank still carries information about how the answer frames a brand. Compare criticism among similar recommendation positions instead of treating every named brand as equivalent.
Criticism became more common lower in the recommendation list
Criticism rate by recommendation position
- First3.6%5,460 of 150,193
- Second or third6.9%17,989 of 259,382
- Fourth or lower8.6%32,388 of 377,610
Takeaway
Software's top picks drew criticism 2.7 times as often as consumer goods
Among industries with at least 1,000 analyzed top recommendations, Software recorded 414 criticized top picks in 8,448, or 4.9%. Consumer Goods recorded 41 in 2,219, or 1.8%. The gap was 3.1 percentage points, and the Software rate was 2.7 times as high.
Use an industry baseline before treating a brand's rate as unusual. This spread identifies where to review answer language and does not show that industry caused the criticism.
Software's top picks drew criticism 2.7 times as often as consumer goods
Selected industry criticism rates
- Software4.9%414 of 8,448
- Data and Analytics4.5%138 of 3,076
- Information Technology4.4%310 of 7,103
- Financial Services4.1%377 of 9,140
- Consumer Electronics2.4%33 of 1,377
- Travel and Tourism2.1%31 of 1,504
- Sports2.0%35 of 1,793
- Consumer Goods1.8%41 of 2,219
Datadog had the most criticized top recommendations
Among brands with at least 30 criticized top recommendations, Datadog led by count with 213 of 1,605, or 13.3%. Atlassian followed with 139 of 900, or 15.4%. Playwright had the highest rate in the displayed set at 71 of 398, or 17.8%.
These are audit starting points, not brand-quality scores. Prompt mix differs by brand, so compare a brand with its own prompts and recurring limitations before comparing rates across brands.
Takeaway
The result held after narrower eligibility checks
The main result covered 150,193 of 215,354 answers with one unambiguous top-ranked brand. We excluded 65,161 top recommendations without a reviewed statement about that brand and one answer with tied top brands. The rate was 3.6% from June 1 through July 15 and 3.6% when each answer had one buyer-need type, compared with 3.6% in the main result.
The sensitivity checks do not change which recommendations the study could include. The result describes the prompts analyzed here and does not measure recommendation accuracy, brand quality, buyer opinion, or causation.
- main result
- 3.6%main result5,460 of 150,193
- June 1 through July 15
- 3.6%June 1 through July 155,459 of 150,170
- one buyer-need type per answer
- 3.6%one buyer-need type per answer5,425 of 149,786
What marketers should do
The top-ranked brand carried a specific negative claim in 3.6% of analyzed recommendations. ChatGPT Search's rate was 5.5 times Google AI Mode's, and criticism was more common for brands ranked lower.
Track rank, specific drawbacks, overall sentiment, reluctance, and explicit rejection separately. Audit the same priority prompts on both engines. Group repeated limitations by brand and buyer need, then inspect the answer and cited sources before changing positioning or content.
How we measured
In the Parse index, we analyzed 162,788 reviewed statements about 150,193 top-ranked brand recommendations, covering 22,546 brands and 15,858 organic prompts on ChatGPT Search and Google AI Mode from May 24 through July 15, 2026.
- higher criticism rate on ChatGPT Search
- 5.5×higher criticism rate on ChatGPT Search6.6% versus 1.2%
- of criticized top picks were not negative overall
- 90.0%of criticized top picks were not negative overall4,912 of 5,460
- of brands ranked fourth or lower included criticism
- 8.6%of brands ranked fourth or lower included criticism32,388 of 377,610
Get the data
Sources
These are the pages this study used.
- Semrush: What is AI sentiment analysis? A marketer's guide · accessed July 30, 2026
- Semrush: AI Visibility Brand Performance Reports · accessed July 30, 2026
- Ahrefs: How to monitor and win brand mentions in AI answers · accessed July 30, 2026
- The Language Blind Spot: Brand reputation across twelve European languages · accessed July 30, 2026