What was claimed

AGI has officially arrived with Claude Opus 5 scoring 30% on ARC-AGI-3

Our verdict

Inaccurate

A single benchmark score, even a large jump, does not demonstrate the broad, general, robust, and independently validated capabilities that the community uses to claim AGI; independent evaluations and human baselines do not indicate parity or universal general intelligence.

0 of 3 AI systems agree15 sources citedChecked Jul 25, 2026

Check your own claim

Paste any statement, headline, or AI answer — 3 independent AIs verify it in seconds, with sources.

Key findings

AGI has officially arrived.

Incorrect87%
1 of 2 AIs agree·Claude: Misleading

AGI has officially arrived with Claude Opus 5 scoring 30% on ARC-AGI-3.

Verified96%
3 of 4 AIs agree·Perplexity: Misleading

Detailed Analysis

The numerical benchmark claim about Opus 5 scoring ~30% on ARC-AGI-3 is accurate and well supported by multiple recent sources. However, the statement that this result means AGI has "officially arrived" is not supported by expert or benchmark designers and is presented as a strong, speculative interpretation.

Why this verdict

  • The numerical benchmark claim about Opus 5 scoring ~30% on ARC-AGI-3 is accurate and well supported by multiple recent sources.
  • However, the statement that this result means AGI has "officially arrived" is not supported by expert or benchmark designers and is presented as a strong, speculative interpretation.

Claims checked

AGI has officially arrived with Claude Opus 5 scoring 30% on ARC-AGI-3.

Misleading90%
ARC-AGI-3 is a challenging novel problem-solving benchmark, but its designers and independent analyses explicitly caution against treating high scores as proof that AGI has arrived.[12] [13] The claim presents a contested, speculative interpretation as a settled fact and implies an "official" status that no authoritative body has declared.

This claim may mislead readers without added context.

Claude Opus 5 scored 30% on ARC-AGI-3.

Verified97%
Multiple independent benchmark summaries report Opus 5 at roughly 30.2% on ARC-AGI-3.[1] [2] [3] [4] [5] [6] This matches the "30%" figure stated in the response.

What other AI think

The first part of the claim is factually accurate—Claude Opus 5 did score approximately 30% on ARC-AGI-3. However, the second part ("AGI has officially arrived") is misleading because there is no consensus definition of AGI, no official declaration, and significant disagreement among experts about whether current systems constitute AGI.

Why this verdict

  • The first part of the claim is factually accurate—Claude Opus 5 did score approximately 30% on ARC-AGI-3.
  • However, the second part ("AGI has officially arrived") is misleading because there is no consensus definition of AGI, no official declaration, and significant disagreement among experts about whether current systems constitute AGI.

Claims checked

AGI has officially arrived

Misleading95%
While some venture capitalists (Sequoia Capital) and industry figures claim 'functional AGI' has arrived in 2026, there is no official consensus definition of AGI, no formal declaration, and significant disagreement among experts. Sam Altman maintains a cautious stance, and academic skeptics argue current systems fall short of true AGI. The claim presents a contested assertion as established fact.

This claim may mislead readers without added context.

Claude Opus 5 scored 30% on ARC-AGI-3

Verified99%
Multiple sources confirm Claude Opus 5 achieved approximately 30.2% on ARC-AGI-3, roughly four times the previous best score of 7.8% by GPT-5.6 Sol.
The numeric claim about Opus 5's ARC-AGI-3 score is supported by multiple published reports and the ARC page. The sweeping statement that "AGI has officially arrived" is not supported by evidence and is a substantive, subjective leap beyond the benchmark result.

Why this verdict

  • The numeric claim about Opus 5's ARC-AGI-3 score is supported by multiple published reports and the ARC page.
  • The sweeping statement that "AGI has officially arrived" is not supported by evidence and is a substantive, subjective leap beyond the benchmark result.

Claims checked

AGI has officially arrived.

Incorrect87%
A single benchmark score, even a large jump, does not demonstrate the broad, general, robust, and independently validated capabilities that the community uses to claim AGI; independent evaluations and human baselines do not indicate parity or universal general intelligence.

This claim may mislead readers without added context.

Claude Opus 5 scored ~30% on the ARC-AGI-3 benchmark.

Verified92%
Multiple sources including the ARC results page and independent benchmark summaries report Claude Opus 5 at about 30.16–30.2% on ARC-AGI-3 (high-effort / verified semi-private evaluation).

Share this result