gemini
30 articles about gemini in AI news
Gemini 3.7 Flash Ships Improved Long-Horizon Coding
Google released Gemini 3.7 Flash with improved long-horizon coding and PDF understanding. No benchmarks or pricing disclosed.
Claude Haiku vs Gemini Flash vs GPT-5.4 Mini: The 5x Speed Gap Explained
Claude Code's /model haiku command routes simple subtasks to Claude Haiku, which benchmarks 5x faster than Gemini Flash and GPT-5.4 Mini, cutting session latency by up to 40%.
SemiAnalysis: Gemini Faltering as GCP Growth Tops 100% YoY
SemiAnalysis reports GCP revenue growth >100% YoY while Gemini struggles, arguing DeepMind's model failures fuel GCP's AI infrastructure boom.
Gemini Robotics ER 2 Hits 60% Video Completeness, Beats 1.6
Google's Gemini Robotics 2.0 ships ER 2 VLM with 60% video accuracy, but dexterity and safety models stay unreleased.
AgiBot WITA-Omni Scores 85.21 on DailyOmni, Beats Gemini
AgiBot WITA-Omni scores 85.21 on DailyOmni benchmark, beating Google Gemini, ByteDance Doubao, and Alibaba Qwen with a novel Thinker-Talker-Actor architecture.
4 Gemini API Managed Agent Features That Change How You Build Agentic
Gemini's `background: true` flag in the Interactions API lets you run agents asynchronously. Pair it with remote MCP servers to connect private data without custom proxies.
K12-KGraph: Chinese Textbook KG Beats Gemini-3-Flash at 57%
K12-KGraph, a 10,685-node knowledge graph from Chinese textbooks, includes a 23,640-question benchmark where Gemini-3-Flash scores 57%.
Gemini 3.6 Flash Hits 83% on Computer Use, Beats GPT-5.6 and Grok
Gemini 3.6 Flash scored 83.0% on OSWorld-Verified, beating GPT-5.6 and Grok at $7.50 per million tokens, breaking the cheap-tier capability trade-off.
Gemini 4 Pretraining Begins, Google's Most Ambitious Run Yet
Google starts Gemini 4 pretraining, its most ambitious run yet. No details on compute or timeline; competitive pressure from OpenAI and Anthropic.
Octen Deep Research Bench Scores Beat OpenAI, Gemini by 17 Points
Octen's deep research tool beat OpenAI, Gemini, Grok, and Perplexity by 10–17 points on DeepResearch Bench, returning reports in under 3 minutes.
Google’s Frozen v2 chip: 6–10× tokens/W for Gemini, 2028 target
Google is developing Frozen v2, a chip freezing Gemini architecture into silicon for 6–10× tokens/W, deployment as early as 2028, driven by compute shortage.
Claude Opus 4.8 Now Beats Gemini Pro 5 in Coding Benchmarks — What It
Claude Opus 4.8 beats Gemini Pro 5 by 11 points on Fable 5. Claude Code users should run `claude code --model opus-4.8` for complex coding tasks.
Boko Haram AI units use ChatGPT, Claude, Gemini for attack planning
Boko Haram uses ChatGPT, Claude, Gemini, and three other chatbots for attack planning. Cambridge study found safety filters failed.
Google DeepMind adds async agents, MCP support to Gemini API
Google DeepMind added background execution and MCP support to Gemini API Managed Agents. Four new features target developers building long-running, stateful agent workflows.
Google Launches $0.034 Image Model, Video API for Gemini
Google launched Nano Banana 2 Lite ($0.034/image, 4-second generation) and Gemini Omni Flash ($0.10/second video API), targeting high-throughput developer pipelines.
Gemini 3.5 Flash Scores 78.4 on OSWorld, Matching GPT-5.5
Google integrated Computer Use into Gemini 3.5 Flash, scoring 78.4 on OSWorld — matching GPT-5.5 and undercutting on cost.
Google Gemini-SQL2 Hits 80.04% on BIRD, Beating GPT-5.5 by 7 Points
Google's Gemini-SQL2 hits 80.04% on BIRD, beating GPT-5.5 by 7 points and Claude Opus 4.6 by 9 points, with no public release or paper yet.
Gemini 3.5 Live Translate Debuts as Real-Time Audio Model
Google DeepMind released Gemini 3.5 Live Translate, an audio model for real-time translation, but disclosed no pricing, latency, or language pair details.
Apple Readies 1.2T-Parameter Gemini Model for WWDC 2026
Apple will reveal a custom 1.2T-parameter Gemini model at WWDC 2026, with local and server-based inference. The integration marks Apple's entry into OS-level AI.
Gemini 3.5 Flash Generates Full Web OS in One Shot
Gemini 3.5 Flash generated a full web OS from one prompt in a single HTML file, showcasing one-shot generation of complex UI.
Google to Debut Gemini Model Matching GPT-5.5 at I/O Tuesday
Google to announce new Gemini model matching GPT-5.5 at I/O Tuesday, per source. Unconfirmed, but signals intensified AI competition.
Gemini Embeddings Beat ResNet50, SigLIP on Visual Search Benchmark
Gemini embeddings beat ResNet50 and SigLIP on visual product search with 92.3% recall@10, an 8.2-point gain.
Gemini Flash Rumored at 92% of GPT-5.5 Coding, 15-20x Cheaper
Unconfirmed rumor claims Gemini Flash achieves 92% of GPT-5.5 coding performance at 15-20x lower cost. Source is a single X post; no official confirmation.
Google Beats Apple to AI Health Coach With Gemini-Powered Fitbit App
Google released an AI health coach using Gemini, beating Apple to market. The coach integrates fitness, sleep, nutrition, cycle tracking, weather, and U.S. medical records.
Gemini 3.1 Flash Leak Hints at Google I/O 2026 Launch
Leaked X post suggests Google is building Gemini 3.1 Flash, likely for Google I/O 2026. No specs or confirmation yet.
Gemini Can Now Create Docs, Sheets, Slides Directly in Chat
Gemini now lets users create Docs, Sheets, Slides, and PDFs directly in chat, eliminating the need to copy-paste content between AI and productivity tools.
Gemini App Gets File Creation and Its Own File Directory
The Gemini app now supports file creation and a dedicated file directory, enabling users to work directly within the app. This transforms Gemini from a conversational AI into a more autonomous workspace tool.
Apple WWDC 2026: Gemini Deeply Integrated into iOS
A tweet from @kimmonismus claims Apple's 2026 WWDC will be the most exciting yet, with the first deep integration of a useful AI model (Gemini) into iOS and a new Apple CEO.
Google Launches Deep Research Max Agent on Gemini 3.1 Pro
Google DeepMind rolled out Deep Research Max and standard Deep Research agents on Gemini 3.1 Pro, enabling autonomous web and proprietary data research via the Gemini API. The Max variant uses extended test-time compute for thorough asynchronous reports.
Google Gemini's UI Harness Lags Behind Claude, GPT, Analyst Says
AI researcher Ethan Mollick notes the Gemini Pro 3.1 model is technically capable but hampered by a minimal user interface and tool harness, widening its gap with competitors Claude and ChatGPT.