Timeline
Released with 1M token context and 128k output.
Claude quietly removed full thinking traces feature, according to user reports and researcher Ethan Mollick.
GPT-5.6 Sol scored 72.7% on DeepSWE
Claude Opus 5 released with Fast Mode (2.5x speed) at Opus 4.8 pricing, saving 50% on tokens
Claude Opus 4.8 beats Gemini Pro 5 by 11 points on Fable 5 benchmark
Claude Opus 4.8 achieves 89% task completion and 2.5% harm rate on WorkBench, a dramatic improvement over GPT-4.
Claude Opus 4.8 adds dynamic workflows for agentic coding
Study reveals GPT-5 exhibits central tendency bias in clinical scoring, compressing predictions toward scale midpoint.
Evaluation study published on arXiv assessing its clinical reasoning capabilities
Reportedly became available according to user social media post
Ecosystem
Claude Opus 4.6
GPT-5
Benchmarks
Evidence (5 articles)
Beyond the Token Limit: How Claude Opus 4.6's Architectural Breakthrough Enables True Long-Context Reasoning
Feb 15, 2026Traders Bet Claude Opus 4.8 Launch Imminent as Options Spike
Jul 20, 2026Anthropic Ships Claude Opus 4.7: 80.1 SWE-Bench, 1M Context
May 17, 2026OpenAI Bids Farewell to GPT-4o: The End of an Era for Controversial AI
Feb 14, 2026Nebius Makes $275M Bet on AI Agent Search with Tavily Acquisition
Feb 10, 2026