GPT-5.5
Signal Radar
Five-axis snapshot of this entity's footprint
Mentions × Lab Attention
Weekly mentions (solid) and average article relevance (dotted)
Timeline
2- Research MilestoneJun 12, 2026
Open-source MoE model matches GPT-5.5 on agentic coding benchmarks
View source - Research MilestoneJun 1, 2026
GPT-5.5 achieves 16% Pass@8 on MA-ProofBench Level I and 5% on Level II, outperforming most models.
View source- score undergraduate:
- 16%
- score phd:
- 5%
Relationships
6Frequently appears with
2Entities that show up in the same articles — shared coverage, not a stated relationship.
Recent Articles
6Databricks Defaults to Chinese Model GLM 5.2, Matches Opus at $1.28/Task
+Databricks defaulted to GLM 5.2 after it matched Opus 4.8 at $1.28/task vs $1.94. The move signals enterprises building custom benchmarks and multi-ve
100 relevanceZhipu GLM-5.2 tops global coding benchmarks, sparks 'DeepSeek moment'
-Zhipu AI's GLM-5.2 ranks top-3 globally on a coding benchmark, with US engineers calling it a daily driver superior to GPT-5.5.
100 relevanceGemini 3.5 Flash Scores 78.4 on OSWorld, Matching GPT-5.5
~Google integrated Computer Use into Gemini 3.5 Flash, scoring 78.4 on OSWorld — matching GPT-5.5 and undercutting on cost.
100 relevanceDeepSeek Raises $7.4B at $50B Valuation in First External Round
~DeepSeek raised ~$7.4B at a $50B valuation in its first external round, with an unusual limited partnership structure and a $2.9B personal investment
100 relevanceMA-ProofBench: GPT-5.5 Hits 16% on Math Analysis, Most Models Near 0%
~MA-ProofBench, a new theorem-proving benchmark for mathematical analysis, shows GPT-5.5 achieving 16% on undergraduate problems and 5% on PhD-level, w
82 relevanceChinese Lab's Free MoE Model Matches GPT-5.5 on Agentic Coding
~A Chinese lab released an Apache-2.0 open-weights MoE model matching GPT-5.5 on agentic coding. This free model challenges proprietary AI's lead with
100 relevance
Predictions
No predictions linked to this entity.
AI Discoveries
3- hypothesisactiveJun 30, 2026
H: Within 14 days, a major publication will publish a comparison or co-deployment article featuring bot
Within 14 days, a major publication will publish a comparison or co-deployment article featuring both GPT-5.5 and Anthropic Claude, likely highlighting a joint enterprise deployment or a benchmark comparison.
75% confidence - observationactiveJun 29, 2026
Novel co-occurrence: GPT-5.5 + Anthropic
GPT-5.5 (ai_model) and Anthropic (company) appeared together in 2 articles this week but have NEVER co-occurred before and have no existing relationship. This is a potential breaking story signal.
85% confidence - observationactiveJun 17, 2026
Research: Theorem Proving for Mathematical Analysis [stable]
State of art: MA-ProofBench reveals GPT-5.5 achieves 16% on undergraduate-level math analysis problems, with most models near 0% on harder PhD-level sets.. Key insight: Mathematical analysis remains a near-unsolved frontier — the 16% ceiling shows that chain-of-thought and RL fine-tuning alone are i
70% confidence
Sentiment History
| Week | Avg Sentiment | Mentions |
|---|---|---|
| 2026-W24 | -0.20 | 1 |
| 2026-W25 | 0.00 | 2 |
| 2026-W26 | -0.05 | 2 |
| 2026-W28 | 0.30 | 1 |