Claude Mythos Preview
Anthropic research preview; strong on deep browsing (BrowseComp 86.9%) and OSWorld-Verified 79.6%.
Anthropic's Claude Mythos Preview isn't a chatbot—it's a weaponized research tool. The graph shows a model that found a HAWK attack in 60 hours for $100K, a speed that redefines offensive security economics. It leverages Firefox and Windows, turning mainstream platforms into attack surfaces. The 3.5x CVE spike after its launch isn't coincidence; it's deployment velocity. Mythos is regulated by the UK AI Safety Institute, a sign of its dual-use danger. It competes directly with Microsoft's MAI-Cyber-1-Flash, which hit 96% on CyberGym, but Mythos's BrowseComp 86.9% suggests deeper autonomous capability. Anthropic built it, but Mozilla's use of it signals enterprise adoption beyond labs. The tension is clear: Mythos ships working exploits in hours, not weeks, making it a moat for Anthropic but a systemic risk for everyone else. The open question: can regulators keep pace with a model that weaponizes every browser it touches?
- ·Mythos found a HAWK attack in 60 hours for $100K, commoditizing exploit development
- ·CVE spikes jumped 3.5x post-launch, showing real-world impact
- ·Microsoft's MAI-Cyber-1-Flash is the direct competitor, with 96% on CyberGym
- ·UK AI Safety Institute regulates the model, highlighting dual-use concerns
- ·Mozilla adoption signals enterprise demand beyond Anthropic's core users
Signal Radar
Five-axis snapshot of this entity's footprint
Mentions × Lab Attention
Weekly mentions (solid) and average article relevance (dotted)
Timeline
16- Research MilestoneJul 28, 2026
Discovered cryptographic weaknesses in HAWK and reduced-round AES after 60 hours of continuous operation
View source- cost:
- $100,000
- duration:
- 60 hours
- PolicyJul 1, 2026
Anthropic restored Mythos access with a security deal after the vulnerability spike.
View source - Research MilestoneJun 12, 2026
Built 8 working exploits from Firefox and Windows kernel patches in ~12 hours; first exploit ready 18 days before patched Firefox 148 shipped
View source - Research MilestoneJun 4, 2026
METR found Claude Mythos Preview could work 16+ hours autonomously
View source- autonomous work duration:
- 16+ hours
- Research MilestoneMay 14, 2026
First AI model to clear all UK AISI cyberattack simulations
View source - Research MilestoneMay 9, 2026
Early snapshot achieves more than 2x time horizon of next best model on METR benchmark
View source - Research MilestoneMay 2, 2026
Claude Mythos Preview fully solved TLO enterprise network simulation in 3 of 10 attempts
View source- simulation:
- TLO
- success rate:
- 3/10
- Research MilestoneMay 2, 2026
Claude Mythos Preview scored 68.6% on AISI expert CTF tasks
View source- score:
- 68.6%
- task type:
- expert CTF
- Research MilestoneApr 15, 2026
First AI model to complete full AISI cybersecurity evaluation
View source - Research MilestoneApr 15, 2026
Leaked benchmark results suggest it 'destroys every other model' including GPT-5, Claude 4 Opus, and Gemini Ultra 2.0
View source - Research MilestoneApr 14, 2026
Achieved 73% success rate on expert-level CTF challenges and completed full 32-step network attack simulation
View source- ctf success rate:
- 73%
- network attack success:
- 3 of 10 attempts
- Regulatory ActionApr 14, 2026
Evaluated by UK AI Safety Institute for autonomous cyber attack capabilities.
- evaluator:
- UK AI Safety Institute
- test type:
- capture-the-flag and network simulation
- Research MilestoneApr 14, 2026
First AI model documented to autonomously complete a full, multi-step cyber attack simulation in UK safety tests.
- success rate:
- 73%
- simulation success:
- 3 of 10 attempts
- steps:
- 32-step network takeover
- Research MilestoneApr 12, 2026
Scored 83.1% on the CyberGym benchmark for vulnerability discovery.
View source- score:
- 83.1%
- benchmark:
- CyberGym
- Product LaunchApr 1, 2026
Anthropic launched Claude Mythos Preview, capable of autonomous cybersecurity vulnerability discovery and exploitation.
View source - Product LaunchJan 1, 2026
Anthropic released Claude Mythos Preview model, used by Mozilla for security bug triage and patching.
View source
Relationships
8Developed
Regulated
Frequently appears with
10Entities that show up in the same articles — shared coverage, not a stated relationship.
Recent Articles
2Claude Mythos Finds HAWK Attack in 60 Hours for $100K
+Claude Mythos found HAWK and reduced-round AES weaknesses in 60 hours for ~$100K, producing the CryptanalysisBench benchmark.
100 relevanceMicrosoft MAI-Cyber-1-Flash Hits 96% on CyberGym
~Microsoft's MAI-Cyber-1-Flash scores 96% on CyberGym, cutting costs 50% by handling 90% of security tasks locally while routing complex cases to GPT-5
100 relevance
Predictions
No predictions linked to this entity.
AI Discoveries
1- observationactiveAug 12, 2026
Lifecycle: Claude Mythos Preview
Claude Mythos Preview is in 'declining' phase (0 mentions/3d, 0/14d, 25 total)
90% confidence
Sentiment History
| Week | Avg Sentiment | Mentions |
|---|---|---|
| 2026-W26 | 0.00 | 1 |
| 2026-W27 | -0.30 | 1 |
| 2026-W31 | 0.25 | 2 |