Timeline
Released with 1M token context and 128k output.
Claude quietly removed full thinking traces feature, according to user reports and researcher Ethan Mollick.
Claude Opus 5 released with Fast Mode (2.5x speed) at Opus 4.8 pricing, saving 50% on tokens
Claude Opus 4.8 beats Gemini Pro 5 by 11 points on Fable 5 benchmark
Claude Opus 4.8 achieves 89% task completion and 2.5% harm rate on WorkBench, a dramatic improvement over GPT-4.
Claude Opus 4.8 adds dynamic workflows for agentic coding
Achieved top score on METR time horizon benchmark, handling 90-minute software tasks
Achieved state-of-the-art status on most benchmarks according to preliminary evaluations
Ecosystem
Claude Opus 4.6
Gemini 3 Pro
Benchmarks
Evidence (9 articles)
Anthropic Opus 4.7: 87.6% SWE-Bench, Constrained Cyber Capabilities
Apr 23, 2026Agent Harnessing: The Infrastructure That Makes AI Agents Work
Apr 25, 2026GPT-5.5 Launches: The Super App Strategy, Not the Model
Apr 28, 2026Anthropic Ships Claude Opus 4.7: 80.1 SWE-Bench, 1M Context
May 17, 2026Claude Code's 1M Context Window Is Now GA — And It's Priced Like Regular Context
Mar 13, 2026ByteDance's CUDA Agent: The AI System Outperforming Human Experts in GPU Code Generation
Mar 2, 2026Glass AI IDE Emerges, Claims to Offer Free Access to Claude Opus 4.6, GPT-5.4, and Gemini 3.1 Pro
Mar 25, 2026Google Gemini-SQL2 Hits 80.04% on BIRD, Beating GPT-5.5 by 7 Points
Jun 13, 2026+ 1 more articles