Skip to content
gentic.news — AI News Intelligence Platform
Connecting to the Living Graph…

semianalysis

30 articles about semianalysis in AI news

SemiAnalysis: Gemini Faltering as GCP Growth Tops 100% YoY

SemiAnalysis reports GCP revenue growth >100% YoY while Gemini struggles, arguing DeepMind's model failures fuel GCP's AI infrastructure boom.

85% relevant

SemiAnalysis: 15GW+ New DC Builds Skip Gensets, UPS

SemiAnalysis tracks 15GW+ of new datacenter capacity without gensets or UPS, signaling a structural shift toward grid reliance over on-site backup power.

85% relevant

SemiAnalysis Tests Qwen3.8-Max-Preview, 2.4T Params

SemiAnalysis tested Qwen3.8-Max-Preview, a 2.4T-param model, per a tweet. No results disclosed, but independent eval is notable.

85% relevant

SemiAnalysis Runs Coding Agents on Its Own Research Workflow

SemiAnalysis is using coding agents internally for data collection, charting, and drafting. No metrics disclosed, but signals production shift.

72% relevant

Nvidia's Next-Gen AI Rack Delayed to 2028, SemiAnalysis Says

Nvidia's next-gen AI rack delayed to 2028 on manufacturing snags per SemiAnalysis. Delay benefits AMD and custom silicon rivals.

95% relevant

China's Etch Tool Localization Outpaces Deposition, SemiAnalysis Says

China's etch imports fell 18% YTD while deposition rose 3%, per @SemiAnalysis_, signaling faster domestic etch localization with supply-chain risks for Western tool makers.

85% relevant

SemiAnalysis: US behind-the-meter datacenter could hit 40GW+ by 2028

SemiAnalysis projects 40GW+ behind-the-meter datacenter capacity by 2028 as grid delays push 50%+ of new builds off-grid.

95% relevant

SemiAnalysis Launches Mythos AI Research Platform

SemiAnalysis launched Mythos, a proprietary AI research platform for semiconductor and AI industry analysis, announced via Twitter on March 5, 2026.

97% relevant

SemiAnalysis: Pretraining Dead for All but Frontier Labs

@SemiAnalysis_ declares pretraining dead for non-frontier labs, citing 'Pretrainitis' as vanity-driven waste. Prompt engineering offers higher ROI.

85% relevant

Anthropic Opus 4.8 Cuts Bug-Finding Cost by 5x, SemiAnalysis Finds

Anthropic's Opus 4.8 + ultracode mode cuts severe bug-finding cost to ~1/5, per preliminary SemiAnalysis experiments with wide error bars.

100% relevant

SemiAnalysis Calls Jensen ComputeX Keynote 'F Tier' Over No AI DC News

SemiAnalysis rated Jensen Huang's ComputeX keynote 'F Tier' for no AI datacenter news and revealed a delayed NVIDIA ARM chip with broken video output.

82% relevant

SemiAnalysis: N3 chip demand far outstrips current consensus estimates

SemiAnalysis argues N3 chip demand far exceeds consensus accelerator models, implying a structural silicon shortage not priced by markets.

89% relevant

SemiAnalysis: Perplexity Slack Bot Beats Claude in Internal Trial

SemiAnalysis found Perplexity's Slack bot beats Claude in internal trial. 96% token budget goes to Anthropic, but usage may shift.

75% relevant

Cerebras Understates On-Chip SRAM by 8x, SemiAnalysis Notes

Cerebras understates on-chip SRAM by 8x per SemiAnalysis, a rare under-specification in chip marketing.

75% relevant

SemiAnalysis: NVIDIA's Customer Data Drives Disaggregated Inference, LPU Surpasses GPU

SemiAnalysis states NVIDIA's direct customer feedback is leading the industry toward disaggregated inference architectures. In this model, specialized LPUs can outperform GPUs for specific pipeline tasks.

85% relevant

Buffett Invests in Google After SemiAnalysis TPU Deep Dive

Berkshire Hathaway invested in Google in Q3 2025, after Buffett studied TPU v5p architecture. He compared it to railroads, citing 8,960 chips and 4.8 Tbps links.

85% relevant

Claude Tool Use: Fable 5 Beats Opus 4.8 at 1.00 Calls

SemiAnalysis analyzed 2.27M Claude responses, finding Fable 5 averages 1.00 tool calls per response versus 0.76 for Opus 4.8. The Opus line shows a downward trend.

93% relevant

KV Cache Offload Makes Storage the New AI Bottleneck

Storage, driven by KV cache offload and rising SSD costs, is now the primary AI bottleneck per Supermicro and SemiAnalysis.

87% relevant

Kirin 9030 metal pitch 32.5nm beats Intel 18A by 10%

Kirin 9030 metal pitch measured 32.5nm, beating Intel 18A by ~10%, achieved without EUV, per SemiAnalysis.

100% relevant

Meta's Superintelligence Compute Ramp Spans 2000km Across Data Centers

Meta's superintelligence compute ramp spans 2000km+ with an RL startup, per SemiAnalysis, marking the most aggressive AI infrastructure build.

100% relevant

Unitree Claims Fastest Iteration Cycle in Global Robotics

@SemiAnalysis_ claims China's Unitree will dominate global robotics due to fastest iteration cycle. No data on iteration time or funding disclosed.

85% relevant

ERCOT datacenter requests exceed grid capacity by 5x

ERCOT datacenter requests far exceed grid underwriting capacity, per @SemiAnalysis_, revealing grid approval as a binding constraint on AI infrastructure buildout.

87% relevant

Cerebras CS4 Stays on 5nm as SRAM Scaling Flattens

Cerebras CS4 stays on 5nm due to SRAM scaling flattening, per @SemiAnalysis_. 3nm offers no density gain, so the chip prioritizes yield and cost.

85% relevant

Median Coding Agent Hits 96k Input Tokens, Rewriting Inference Economics

SemiAnalysis found median coding agent uses 96k input tokens from 432k requests, shifting inference cost focus from output to context.

95% relevant

Vibe-Coding Bottleneck: CPU Box Rental Gets Harder

SemiAnalysis flags that vibe-coding wave makes cheap CPU box rentals less routine, bottlenecking developers who need quick cloud compute for AI prototyping.

75% relevant

Datacenter Developers Flee City Zoning for Unincorporated County Land

Datacenter developers are siting projects on unincorporated county land to avoid city zoning delays, redrawing the AI infrastructure map per @SemiAnalysis_.

100% relevant

NVIDIA Vera Rubin VR NVL72: Value Extraction Engine Arrives

NVIDIA's Vera Rubin VR NVL72 shifts from value vendor to value extractor, targeting TCO. SemiAnalysis argues this overturns prior pricing paradigm.

95% relevant

CPU Demand Flipping the AI Narrative as Datacenter Growth Shifts

A new analysis from SemiAnalysis indicates CPU demand is rising in AI datacenters, reversing a narrative of GPU-only dominance. This shift signals changing workload patterns and infrastructure priorities.

100% relevant

Claude Code Turns Are 75% Reading, 219 Sessions Show

Red Hat's analysis of 219 Claude Code sessions shows median turns are ~75% reading. This reframes optimization toward context management.

91% relevant

SambaNova SN50 MVP Runs MiniMax M2.7, But Batch Size Limit Looms

SambaNova's SN50 MVP runs MiniMax M2.7 but is stuck at batch size 2, highlighting software maturity issues for frontier models.

72% relevant