Claude Sonnet 4.6
Anthropic · Launched Feb 2026
Anthropic's fast mid-tier model; sits right on the human OSWorld-Verified baseline at 72.1%.
Benchmark performance
Other screen-level os control agents
The 3 agents in this category, ranked by peak benchmark.
| Agent | Maker | Launch | Peak | Pricing |
|---|---|---|---|---|
| Claude Mythos Preview | Anthropic | 2026-04 | 86.9 | Research preview |
| Claude Sonnet 4.5 | Anthropic | 2025-09 | 62.9 | Legacy Anthropic API |
Recent coverage
2026-06-30
Claude Sonnet 5 Beats Opus 4.8 on Knowledge Work at Lower Cost
2026-06-15
Anthropic Blocks Claude from Outputting GPL, Apache, 7 Other Licenses
2026-06-08
Anthropic: AI agents fail biology retrieval, miss 261 Ebola sequences
2026-06-04
Ontology-Grounded AI Agent Testing Hits 48.3% Regulatory Coverage vs.
2026-05-14
Anthropic Deprecates Fixed Thinking Budgets, Forces Adaptive Mode
2026-04-23
3 Ways to Switch Claude Code Models Instantly: /model, --flag, and ENV Variables
2026-04-17
Navox Agents: 8 Specialized Claude Code Agents with Human Checkpoints
2026-04-16
Claude Code's Edge: Why Sonnet 4.5 Beats GPT-4o for Multi-File Projects
Quick facts
- Type
- Screen-level OS control
- Maker
- Anthropic
- Launch
- 2026-02-01
- Open source
- No
- Pricing
- $3 / $15 per M tokens
- Benchmarks scored
- 4
- Article mentions
- 27
- Rank in category
- #1 of 3