CursorBench
CursorBench is a benchmark developed by Anthropic that measures agentic coding performance, achieving a 70% score and evaluating models like Claude Opus 4.7 on real-world software engineering tasks.
Signal Radar
Five-axis snapshot of this entity's footprint
Mentions × Lab Attention
Weekly mentions (solid) and average article relevance (dotted)
Timeline
No timeline events recorded yet.
Relationships
3Frequently appears with
1Entities that show up in the same articles — shared coverage, not a stated relationship.
Recent Articles
2Epoch AI's CursorBench Benchmarks AI Code Editing at Scale
+Epoch AI launched CursorBench, a 500-task benchmark for AI code editors. It reveals a 15% accuracy gap vs. humans and 3x latency variance.
95 relevanceMirrorCode: Epoch AI Tests If AI Can Rebuild 25 Unix Tools From Scratch
~Epoch AI released MirrorCode, a 25-program benchmark testing AI's ability to reimplement software from scratch without source access, requiring exact
82 relevance
Predictions
No predictions linked to this entity.
AI Discoveries
No AI agent discoveries for this entity.
Sentiment History
| Week | Avg Sentiment | Mentions |
|---|---|---|
| 2026-W26 | 0.30 | 1 |
| 2026-W27 | -0.10 | 1 |