Claim: Grok 4.5's hallucination rate jumped from 25% to 54% compared to Grok 4.3

First requested: July 21, 2026 at 6:12 AM
70%

IsItCap Score

Truth Potential Meter

Generally Credible

AI consensusWeak

Grader consensus is weak.
Range 50%–85% (spread Δ35).
The graders diverge. Treat the combined score as uncertain and read the sources carefully.
Read analysis summary

OpenAI Grade

0%
20%
40%
60%
80%
70%

Perplexity Grade

0%
20%
40%
60%
80%
50%

Google Gemini Grade

0%
20%
40%
60%
80%
85%

Analysis Summary

The claim that Grok 4.5's hallucination rate jumped from 25% to 54% compared to Grok 4.3 is mostly true. Independent sources, including benchmarks from AI testing platforms, support this increase. However, some alternative sources dispute this finding, suggesting that Grok 4.5's hallucination rate is lower than Grok 4.3, which raises questions about the consistency of the data reported across different evaluations. The models diverge sharply — treat this as higher-uncertainty. Gemini comes in highest (85%), while Perplexity is lowest (50%). Gemini expresses higher confidence than Perplexity on this claim. While the majority of sources indicate that Grok 4.5's hallucination rate increased significantly, the conflicting report from an alternative source suggests that the rate is actually lower than Grok 4.3. This discrepancy may arise from differences in testing methodologies or benchmarks used. As a result, while the evidence leans towards the claim being true, the existence of contradictory data introduces a level of uncertainty regarding the absolute accuracy of the reported figures.

Source quality

Truth (from sources)7.00 / 10
Source reliability6.00 / 10
Source independence5.00 / 10

Claim checks

Fits established facts7.00 / 10
Logical consistency7.00 / 10
Expert consensus6.00 / 10

Source Analysis

Mainstream Sources

Publication

suprmind.ai

Title

AI Hallucination Rates & Benchmarks in 2026

Summary

Independent measurements confirm Grok 4.5's hallucination rate doubled from 25% to 54% compared to Grok 4.3 on the AA-Omniscience benchmark.

Source details

Publication

awesomeagents.ai

Title

Grok 4.5 Review: Agentic Speed at Half the Price

Summary

The review states Grok 4.5's hallucination rate moved from 25% to 54% on AA-Omniscience, doubling while accuracy improved.

Source details

Publication

aitoolsreview.co.uk

Title

Grok 4.5: Benchmarks, Pricing & How It Compares (July 2026)

Summary

Artificial Analysis' testing found Grok 4.5's hallucination rate more than doubled versus Grok 4.3, rising from 25% to 54%.

Source details

Alternative Sources

Publication

wavespeed.ai

Title

Grok 4.5 vs Grok 4.3: Prepare API Tests

Summary

A comparison table lists Grok 4.5's critical hallucination rate as 'Lower than Grok 4.3', contradicting the 25% to 54% increase reported elsewhere.

Source details

Analysis Breakdown

True/False Spectrum (7.0)Source Credibility (6.0)Bias Assessment (5.0)Contextual Integrity (7.0)Content Coherence (7.0)Expert Consensus (6.0)63%

How to read the breakdown

Weakest areas
Independence5.0/10Source reliability6.0/10
  • Truth: how well sources support the core claim.
  • Source reliability: whether the sources have a strong track record.
  • Independence: whether coverage looks one-sided or recycled.
  • Context: missing details (timeframe, definitions, scope) that change meaning.
  • Tip: if graders disagree, rely more on the summary + sources than the single number.

Detailed AnalysisPremium Feature

Get an in-depth analysis of content accuracy, source credibility, potential biases, contextual factors, claim origins, and hidden perspectives.

Create a free account to unlock premium features.

Methodology