"Claude Opus 5 from @AnthropicAI is the new SOTA on ARC-AGI-3: 30.2% The previous high score (7.8%) was set by GPT-5.6 Sol (Max) Throughout our analysis, we observed novel behavior that allows Opus 5 to solve previously unbeaten environments, outperforming Fable"

In our testing to date, Anthropic’s Fable-class models score approximately 20% on the ARC-AGI-3 Public Demo environments

Claude Opus 5 reaches 30.2%, materially outperforming Fable

Our analysis suggests the gain comes from stronger logical reasoning, which enables more Claude Opus 5 was able to score 100% on 5 previously unbeaten environments

Of these, it was able to beat 4 of them matching or surpassing human level efficiency

Newly beaten environments: ar25, ft09, lp85, r11l, s5i5

6 of the 25 public demo environments have now been solved During our analysis of Opus 5, we observed a new capability previously unseen from frontier models

Opus 5 used advanced logical reasoning to turn ARC-AGI-3 layouts into algebraic notation. On action 23 it described the scene as "4_center = 2×axis − 5_center"

This is the first ARC-AGI-2

Claude Opus 5 scores 90.4% for $2.06/task

This is competitive with previous SOTA performance for slightly higher cost ARC-AGI-1

Claude Opus 5 scores 97.5% for $0.70/task

This is competitive with previous SOTA performance for slightly higher cost — ARC Prize

Source: https://x.com/arcprize/status/2080716561539907928