OpenAI's newly released GPT-5.6 Sol has taken the No. 2 spot on the Agent Arena leaderboard, closing in on frontier leader Claude Fable 5 after evaluation across 7.8K real-world agentic sessions. The result marks a measurable step forward for OpenAI in autonomous, long-horizon agent work, though Anthropic's model retains its lead on the signals that best capture implicit user satisfaction.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.