Gemini 3 Pro
Gemini 3 Pro is a multimodal language model from Google DeepMind, released on February 19, 2026. It scores 90.1 on MMLU-Pro, 80.6 on SWE-Bench Verified, and holds an Arena ELO of 1485, positioning it among the leading 2026 models for reasoning, code generation, and instruction-following. The model accepts text, code, images, and video inputs, with API pricing set at $2.00 per million input tokens and $12.00 per million output tokens. Its significance lies in establishing a new performance baseline for the Gemini family: the 90.1 MMLU-Pro result represents a 3.2-point gain over Gemini 2.5 Pro’s 86.9, while the 80.6 SWE-Bench Verified score improves on its predecessor’s 71.2 by 9.4 percentage points, demonstrating concrete advances in both knowledge benchmarks and real-world software engineering tasks. These documented, verifiable data points give developers a specific reference for evaluating the model against contemporaneous offerings in a market where tooling is increasingly tied to capability.
Google DeepMind's Gemini 3 Pro, released February 2026, posts strong benchmarks—90.1 MMLU-Pro, 80.6 SWE-Bench, 1485 Arena ELO—but faces a two-front war. It competes directly with Claude Opus 4.7 and GPT-3.5, while DeepSeek V4 also targets it. The model deploys Sparse MoE and Chain-of-Thought, powering downstream products Gemini-SQL2 and Deep Research Max. Yet the competitive landscape shifted sharply: DeepSeek V4 slashed prices 75% to $0.43/M tokens, and Claude Opus 4.7 shipped with 80.1 SWE-Bench and 1M context. Gemini 3 Pro's benchmark edge narrows as rivals match or undercut on cost. Google must now prove its model's premium can withstand commoditization pressure.
- ·Competes with Claude Opus 4.7, GPT-3.5, and DeepSeek V4
- ·Deploys Sparse MoE and Chain-of-Thought techniques
- ·Powers Gemini-SQL2 and Deep Research Max products
- ·DeepSeek V4's 75% price cut threatens Gemini 3 Pro's positioning
Signal Radar
Five-axis snapshot of this entity's footprint
Mentions × Lab Attention
Weekly mentions (solid) and average article relevance (dotted)
Timeline
2- Research MilestoneApr 16, 2026
Achieved top score on METR time horizon benchmark, handling 90-minute software tasks
View source- benchmark:
- METR Time Horizon
- score:
- 77%
- task duration:
- 1 hour 30 minutes
- Research MilestoneFeb 20, 2026
Achieved state-of-the-art status on most benchmarks according to preliminary evaluations
View source- timeframe:
- 3 months
- improvement:
- significant jumps
Relationships
10Developed
Competes With
Uses
Deploys
Frequently appears with
10Entities that show up in the same articles — shared coverage, not a stated relationship.
Recent Articles
2Gemini 4 Pretraining Begins, Google's Most Ambitious Run Yet
+Google starts Gemini 4 pretraining, its most ambitious run yet. No details on compute or timeline; competitive pressure from OpenAI and Anthropic.
85 relevanceEpoch AI's EBR-Bench: Top Models Score 30-50% on Experience-Based Reasoning
+Epoch AI's EBR-Bench tests experience-based reasoning. Top models score 30-50%, with Google Gemini 3 Pro leading at 48.2%, revealing a gap between pat
100 relevance
Predictions
No predictions linked to this entity.
AI Discoveries
1- observationactive1d ago
Lifecycle: Gemini 3 Pro
Gemini 3 Pro is in 'surging' phase (1 mentions/3d, 1/14d, 25 total)
90% confidence
Sentiment History
| Week | Avg Sentiment | Mentions |
|---|---|---|
| 2026-W24 | 0.50 | 1 |
| 2026-W27 | 0.60 | 1 |
| 2026-W30 | 0.30 | 1 |