Timeline
Next-gen AI rack system delayed to 2028 due to manufacturing snags
Collaborated with Hugging Face to release open-source robot models for physical AI
Cerebras and Flex announce 7x production expansion of CS-3 at Milpitas facility
Renting back GPU capacity from neoclouds due to softening demand
US export controls target NVIDIA A100 and H100 chips
NVIDIA claims Blackwell inference stack cut DeepSeek V4 token costs 5x in one month
Vera Rubin NVL72 cloud rollout expands to Europe with H2 2026 deployments
Disclosed three mechanical innovations for wafer-scale chip cooling: vertical power delivery, flexible interposers, direct-impingement cooling
Announced CS4 wafer-scale chip staying on 5nm due to SRAM scaling flattening.
Cerebras published benchmark results for Kimi K2.6 on CS-3, claiming 981 tokens/sec and 6.7× speedup over GPU cloud.
Ecosystem
Cerebras Systems
Nvidia
Evidence (13 articles)
China's Open-Source AI Surge: How Local Models Are Redefining Global Competition
Feb 12, 2026OpenAI Unleashes Real-Time Coding Revolution with GPT-5.3-Codex-Spark
Feb 12, 2026Cerebras Challenges Nvidia Inference Monopoly with Wafer-Scale Edge
May 20, 2026Cerebras Claims Performance Parity With Nvidia H100 on AI Training
Jun 13, 2026OpenAI's $100 Billion Horizon: How ChatGPT's Explosive Growth Is Reshaping the AI Industry
Feb 9, 2026Cerebras, Flex Expand CS-3 Production 7x at Milpitas Facility
Jul 9, 2026Nvidia's Next-Gen AI Rack Delayed to 2028, SemiAnalysis Says
Jul 6, 2026Inference shift opens door for AI chip startups to challenge Nvidia
May 3, 2026+ 5 more articles