Anthropic's Claude Opus 5 scored 30.2 percent on ARC-AGI-3, nearly quadrupling GPT-5.6 Sol's previous record of 7.8 percent. The benchmark's developers say the model independently formulated reflection equations, a behavior they had never seen from another model, and attribute to stronger logical reasoning.
The article Anthropic's Opus 5 blows...
🛡️ VERIFIED CYBER INTELLIGENCE ID: #3667315