Anthropic's Opus 5 blows past Fable 5 and GPT-5.6 Sol on the benchmark designed to measure real intelligence
10/10Anthropic's Claude Opus 5 scored 30.2% on the ARC-AGI-3 intelligence benchmark, nearly quadrupling GPT-5.6 Sol's 7.8% score and demonstrating novel reflection capabilities. This marks a significant leap in AI reasoning and real intelligence measurement.
