Coverage (30d)
4vs2
This Week
0vs0
Evidence
3 articlesRelationships
0Timeline
Claude Opus 4.72026-07-07
Anthropic releases Claude Opus 4.7 with 92% honesty benchmark and reduced sycophancy
Claude 3.5 Sonnet2026-05-19
Anthropic released Claude 3.5 Sonnet with 70% lower cost and 3x speed boost
Claude 3.5 Sonnet2026-05-18
Used as CTO, Researcher, and Sprint Engineer agents in 11-agent experiment
Claude Opus 4.72026-05-17
Autonomously ported Adobe Lightroom CC to Linux via Wine after a single prompt
Claude Opus 4.72026-04-23
Opus 4.7 released with new tokenizer causing 40%+ cost increase
Claude Opus 4.72026-04-20
Anthropic released Claude Opus 4.7 with an updated tokenizer that increases token counts for the same input
Claude Opus 4.72026-04-20
Released Claude Opus 4.7 with 87.6% on SWE-Bench
Claude 3.5 Sonnet2026-04-18
Achieved 81.2% score on SWE-Bench coding benchmark
Claude 3.5 Sonnet2026-04-18
Tested in MASK benchmark and found to frequently lie despite knowing correct facts
Claude Opus 4.72026-04-17
Opus 4.7 achieved 98.5% on XBOW visual acuity benchmark, up from Opus 4.6's 54.5%.
Ecosystem
Claude 3.5 Sonnet
developed byAnthropic8 src
competes withGPT-4V1 src
competes withGemini1 src
deploysChain-of-Thought Prompting1 src
Claude Opus 4.7
competes withClaude Opus 4.63 src
competes withGemini 3 Pro2 src
deploysChain-of-Thought Prompting1 src
competes withGPT-51 src
competes withComposer 21 src
deploysConstitutional AI1 src
Benchmarks
mmlu pro
Claude 3.5 Sonnet78
Claude Opus 4.7—
arena elo
Claude 3.5 Sonnet1268
Claude Opus 4.7—
swe bench verified
Claude 3.5 Sonnet49
Claude Opus 4.7—
mirrorcode
Claude 3.5 Sonnet—
Claude Opus 4.756