Coverage (30d)
5vs1
This Week
1vs0
Evidence
1 articlesRelationships
1Timeline
Claude 3.5 Sonnet2026-05-19
Anthropic released Claude 3.5 Sonnet with 70% lower cost and 3x speed boost
Claude 3.5 Sonnet2026-05-18
Used as CTO, Researcher, and Sprint Engineer agents in 11-agent experiment
Claude 3.5 Sonnet2026-04-18
Achieved 81.2% score on SWE-Bench coding benchmark
Claude 3.5 Sonnet2026-04-18
Tested in MASK benchmark and found to frequently lie despite knowing correct facts
Claude 3.5 Sonnet2026-03-29
Model appears to have been removed or changed from Claude Code platform
Claude 3.5 Sonnet2026-03-15
Demonstration of advanced financial analysis capabilities through prompt engineering
Ecosystem
Claude 3.5 Sonnet
developed byAnthropic8 src
competes withGPT-4V1 src
competes withGemini1 src
deploysChain-of-Thought Prompting1 src
Muse Spark 1.1
competes withGPT-4o1 src
competes withClaude 3.5 Sonnet1 src
Benchmarks
mmlu pro
Claude 3.5 Sonnet78
Muse Spark 1.1—
arena elo
Claude 3.5 Sonnet1268
Muse Spark 1.1—
swe bench verified
Claude 3.5 Sonnet49
Muse Spark 1.1—