30 articles about google in AI news
All 8 'Attention Is All You Need' Authors Have Left Google
All eight Transformer paper authors have left Google. SemiAnalysis sees a strategic pivot to TPU infrastructure monetization.
Google's 7-Year-Old TPUs Still Run at 100% Utilization, Says Vahdat
Google's 7-8 year old TPUs still run at 100% utilization per Amin Vahdat, citing Jevons Paradox. Efficiency gains drive demand, keeping all chips busy.
Google's TPUv8i Starts Software Bring-Up on g3 Codebase
SemiAnalysis reports Google's TPUv8i entered software bring-up on g3 and public stacks, signaling accelerated TPU software externalization. No specs disclosed.
Hyperscalers Commit ~$2T to AI Hardware; Google Leads at $811B
Hyperscalers hold ~$2T in AI hardware commitments; Google leads at $811B while Apple trails at $57B. Memory becomes strategic asset.
Google Open-Sources TPU Raiden, Its NIXL Equivalent for KV Cache
Google open-sourced TPU Raiden, its KV cache transfer library equivalent to NVIDIA NIXL, signaling deeper externalization of its TPU stack.
Google Cloud Adds MCP Support to Vertex AI
Google Cloud's MCP support in Vertex AI lets Claude Code users query BigQuery and GCS directly. Set up MCP servers to replace custom scripts with a standardized protocol.
Epoch AI: Google's Colossus 1 Training Compute Hits 1e26 FLOP
Google's Colossus 1 used 1e26 FLOP at $4.6B, per Epoch AI. It is the largest known training run, signaling a new capital scale.
Google Ships 3 Flash Models as 3.5 Pro Remains Missing
Google shipped three Gemini Flash models but 3.5 Pro remains delayed. Efficiency gains don't close the frontier gap with OpenAI and Anthropic.
Gemini 4 Pretraining Begins, Google's Most Ambitious Run Yet
Google starts Gemini 4 pretraining, its most ambitious run yet. No details on compute or timeline; competitive pressure from OpenAI and Anthropic.
Michaels Launches 'Ask Mike' AI-Powered Shopping Assistant Built on Google Cloud
Michaels launched 'Ask Mike,' an AI shopping assistant on Google Cloud using Gemini models. The tool helps customers find products and get project ideas, potentially reducing search friction in craft retail.
Google’s Frozen v2 chip: 6–10× tokens/W for Gemini, 2028 target
Google is developing Frozen v2, a chip freezing Gemini architecture into silicon for 6–10× tokens/W, deployment as early as 2028, driven by compute shortage.
Google Chooses Intel EMIB-T for 9th-Gen TPUs, Breaking TSMC's CoWoS Monopoly
Google picks Intel EMIB-T for 9th-gen TPU, breaking TSMC CoWoS monopoly. Move signals architectural bet on power integrity and reticle-free scaling.
Google alone ships full any-to-any multimodal models
Mollick notes Google alone ships full any-to-any multimodal models; OpenAI and Anthropic lag. This gives Google a structural advantage in agentic workflows.
Google DeepMind adds async agents, MCP support to Gemini API
Google DeepMind added background execution and MCP support to Gemini API Managed Agents. Four new features target developers building long-running, stateful agent workflows.
Vibe coding leaves terminal; Google Cloud MCP server goes live
Google Cloud ships first major cloud MCP server, enabling AI agents to directly access Vertex AI, BigQuery, and Cloud Storage. Move validates MCP as standard for AI-to-infrastructure communication.
Whering Secures $7M from eBay Ventures and Google AI Futures Fund
Whering raised $7M from eBay Ventures and Google AI Futures Fund, reaching 10M users. The funding will scale AI-powered wardrobe tech for personalized, sustainable fashion.
Kering Deploys AI-Powered Sustainable Sourcing Assistant on Google Cloud
Kering launched a Sustainable Sourcing Assistant on Google Cloud's Vertex AI. The tool helps luxury brands like Gucci and Saint Laurent evaluate materials for environmental and social impact, advancing sustainability in procurement.
Google TPU Humufish Drops TSMC CoWoS for Intel EMIB-T
Google's next TPU Humufish uses Intel EMIB-T packaging instead of TSMC CoWoS, breaking the industry default for AI accelerators.
Google Launches $0.034 Image Model, Video API for Gemini
Google launched Nano Banana 2 Lite ($0.034/image, 4-second generation) and Gemini Omni Flash ($0.10/second video API), targeting high-throughput developer pipelines.
Google ADK Go 2.0 Adds Graph Engine, Human-in-Loop for Agents
Google released ADK Go 2.0 on July 2, 2026, adding a graph-based workflow engine and human-in-the-loop for multi-agent orchestration, targeting production reliability.
Google Cloud Joins MCP: How to Connect Claude Code to BigQuery
Google Cloud's MCP server lets Claude Code query BigQuery and manage GCS directly. Install it with `claude mcp add google-cloud` and authenticate.
Gap Inc. Partners with Google Cloud
Gap Inc. announced a multi-partner AI initiative with Google Cloud, Zeta Global, and Publicis Sapient to modernize its marketing, focusing on personalized experiences across owned channels for brands like Old Navy and Athleta.
Google DeepMind loses its third senior AI researcher in months as Nobel laureate John Jumper joins Anthropic
Nobel laureate John Jumper, DeepMind Director and VP Engineering Fellow who co-created AlphaFold, has left Google after nine years for Anthropic. The move follows Noam Shazeer's exit to OpenAI two days earlier — less than two years after Google paid $2.7B to reacquire him — and David Silver's Januar
Noam Shazeer leaves Google for OpenAI after $2.7B Character.AI return
Noam Shazeer left Google for OpenAI months after returning via a $2.7B Character.AI deal, marking the second major AI talent move this year.
Google Gemini-SQL2 Hits 80.04% on BIRD, Beating GPT-5.5 by 7 Points
Google's Gemini-SQL2 hits 80.04% on BIRD, beating GPT-5.5 by 7 points and Claude Opus 4.6 by 9 points, with no public release or paper yet.
Google-Backed Tapestry Claims 811 Grid Apps Processed in Under an Hour
Google-backed Tapestry processed 811 PJM interconnection applications in under an hour, claiming to slash a months-long process to minutes.
Google Open-Sources DiffusionGemma, 26B Model Hits 1K Tokens/Sec on H100
Google open-sourced DiffusionGemma, a 26B-parameter diffusion text model hitting 1,000 tokens/sec on H100 — 4x faster than autoregressive models, but with lower quality.
Google Books Intel for 3M+ TPUs in 2028 as TSMC CoWoS Hits Capacity Wall
Google booked Intel to package 3M+ TPUs in 2028 as TSMC CoWoS capacity caps out. SK hynix tests HBM on Intel EMIB, potentially unlocking Nvidia's Feynman architecture.
Google Titan: A New Architecture That Could Dethrone Transformers
Google's Titan architecture claims to surpass Transformers on long-context tasks via neural long-term memory, achieving 1.2x-2.5x speedups on benchmarks.
Google to Pay SpaceX $920M/Month for xAI Compute Capacity
Google commits $11B/year to SpaceX for compute at xAI data centers, potentially adding $1T to SpaceX's valuation.