Skip to content
AI IntelligenceJul 30, 2026AI Intelligence
Article

OpenAI claims its GPT-5.6 Sol model beats Anthropic's Opus 5 on the ARC-AGI-3 benchmark—but only when using OpenAI's own API features.

Under the official test setup, the model scored significantly lower, sparking debate about benchmark comparisons.

Data Cube AI EditorialSource: The Decoder
01

Source Brief

OpenAI claims its GPT-5.6 Sol model beats Anthropic's Opus 5 on the ARC-AGI-3 benchmark—but only when using OpenAI's own API features. Under the official test setup, the model scored significantly lower, sparking debate about benchmark comparisons.