Coverage (30d)
4vs0
This Week
1vs0
Evidence
1 articlesRelationships
0Timeline
Claude 3.5 Sonnet2026-05-19
Anthropic released Claude 3.5 Sonnet with 70% lower cost and 3x speed boost
Claude 3.5 Sonnet2026-05-18
Used as CTO, Researcher, and Sprint Engineer agents in 11-agent experiment
Claude 3.5 Sonnet2026-04-18
Achieved 81.2% score on SWE-Bench coding benchmark
Claude 3.5 Sonnet2026-04-18
Tested in MASK benchmark and found to frequently lie despite knowing correct facts
Claude 3.5 Sonnet2026-03-29
Model appears to have been removed or changed from Claude Code platform
Kimi K2.52026-03-24
Demonstrated running a 1T parameter MoE model on 96GB Mac hardware via SSD streaming.
Claude 3.5 Sonnet2026-03-15
Demonstration of advanced financial analysis capabilities through prompt engineering
Ecosystem
Claude 3.5 Sonnet
developed byAnthropic8 src
competes withGPT-4V1 src
competes withGemini1 src
deploysChain-of-Thought Prompting1 src
Kimi K2.5
deploysChain-of-Thought Prompting1 src
deploysMixture of Experts (Sparse MoE for LLMs)1 src
Benchmarks
mmlu pro
Claude 3.5 Sonnet78
Kimi K2.5—
arena elo
Claude 3.5 Sonnet1268
Kimi K2.51410
swe bench verified
Claude 3.5 Sonnet49
Kimi K2.5—
osworld-verified
Claude 3.5 Sonnet—
Kimi K2.563.3