nvidia
30 articles about nvidia in AI news
NVIDIA Opens NIXL Library to AMD After Decade-Long Block
NVIDIA accepted AMD PRs into NIXL after 10-year block, brokered by SemiAnalysis over 24 months. The move could reshape AI inference infrastructure competition.
First Nvidia H200s Reach China as Beijing Loosens Import Block
ByteDance and Tencent each received ~10,000 Nvidia H200s, the first meaningful China shipments since December. Beijing still gates purchases and wants most of the 100,000-unit allowance in power-starved Hong Kong.
OpenAI Leases 8GW Ohio Site; Nvidia Backs $105B
OpenAI leased 8GW Ohio site; Nvidia backs $105B residual value. WSJ: $3T off-balance-sheet AI commitments.
NVIDIA Releases 550B-Param Nemotron Chat Teacher
NVIDIA released a 550B-param Nemotron Chat Teacher on Hugging Face for multi-turn chat, tone-sensitive writing, and distillation. No benchmarks disclosed.
Nvidia to Acquire Groq; Leftover Biz Buys Blackwell Clusters
Nvidia reportedly acquiring Groq; leftover Groq rents B300/GB300/Rubin GPUs, many Blackwell-only, per SemiAnalysis.
Nvidia, Wall Street Giants Eye $500B AI Infrastructure Fund
Nvidia is negotiating a $500B AI infrastructure funding package with Apollo, Blackstone, BlackRock, Brookfield, Goldman and KKR, per the FT. The deal may be announced Monday.
SemiAnalysis: Can TileRT Software Match Cerebras on NVIDIA GPUs?
SemiAnalysis is testing TileRT InferenceX, software claiming batch-1 ultra-high interactivity on NVIDIA GPUs, targeting Cerebras, Groq LPU, and SambaNova. No benchmarks disclosed yet.
Nvidia to Invest Up to $3B in Stargate Power Developer Lancium
Nvidia to invest up to $3B in Lancium, Stargate's power developer, per The Information. Deal signals Nvidia's move into energy infrastructure.
US Probes Chinese Firms' Remote Access to Nvidia GPUs
BIS probes Chinese firms' remote access to Nvidia chips, a potential export-control gap. Smuggling is covered; remote rental may not be.
NVIDIA Vera CPU Claims 3.67x Storage Speed vs x86
NVIDIA's Vera Arm CPU claims 3.67x faster storage processing than x86, targeting Intel and AMD via BlueField-4 STX.
Musk: SpaceX to Build Exclusively With Nvidia, Touts Space AI Servers
Musk says SpaceX will build exclusively with Nvidia and floats space-based AI servers, extending Nvidia's reach into aerospace. Contract terms undisclosed.
NVIDIA Releases Nemotron VoiceChat, First Open Full-Duplex Speech Model
NVIDIA released Nemotron VoiceChat, claiming the first open full-duplex speech model with tool calling and barge-in. The move targets real-time voice agents, challenging proprietary APIs.
Nvidia-OpenAI Talks Hit $250B; NVL72 Ships 96GB HBM3e
Nvidia in talks to invest $250B in OpenAI, six times Microsoft's stake. NVL72 now ships 96GB HBM3e; $25B bond planned.
Safe Superintelligence Partners Nvidia for 10x Compute Scale-Up
SSI partners with Nvidia for 10x compute scale; Nvidia also invests. Details on investment size and timeline undisclosed, raising questions about the startup's capital needs.
Nvidia Vera Rubin Shifts AI Strategy Beyond Raw GPU Speed
Nvidia's Vera Rubin architecture pivots from raw GPU FLOPS to system-level AI infrastructure, targeting memory bandwidth and interconnect bottlenecks that constrain large-scale model training.
Nvidia Weighs $250B Guarantee for OpenAI's Ohio Campus
Nvidia may guarantee $250B for OpenAI's Ohio data center lease, with a $350B chip financing deal, per @tomshardware. Unconfirmed but signals massive AI infrastructure funding.
NVIDIA's Molt: 9.2K-Line RL Framework Scales to 1T-Parameter MoE Models
NVIDIA released Molt, a 9.2K-line PyTorch RL framework scaling to 1T-parameter MoE models via vLLM, targeting agentic tasks with fully-async rollout.
Nvidia, SK Group Announce $500B AI Infrastructure Partnership
Nvidia and SK Group announced a $500B partnership for HBM4 memory supply and a 2 GW AI data center in South Korea, locking in SK Hynix as Nvidia's primary memory supplier through 2030.
Jensen Huang: DeepSeek, Kimi open models boost Nvidia sales
Jensen Huang says Chinese open models DeepSeek and Kimi boost Nvidia GPU demand, not threaten it. Market misunderstood their impact twice.
Nvidia Invests $1B in Naver, Expands SK Group Pact for Korea AI Hub
Nvidia invests $1B in Naver for an AI data center in South Korea and expands its SK Group accord. The deals lock in Asian supply chains and data center capacity.
Nvidia, Meta, Mistral Warn US Against Broad Open-Weight AI Restrictions
Nvidia, Meta, Mistral sign letter urging US against broad open-weight AI restrictions, arguing distillation is legitimate. Comes as White House weighs response to Chinese AI model distillation.
Nvidia Ships Hundreds of Thousands of Grace Standalone Servers
Nvidia shipped hundreds of thousands of Grace standalone servers. The CPU pivot targets agentic AI workloads shifting hardware balance.
Nvidia Vera CPU Hits SPECrate 2026: 1.7× AMD Epyc 9755
Nvidia's Vera CPU scored 1.7× SPECrate integer 2026 vs AMD Epyc 9755. First custom core for agentic AI, H2 2026 release.
zAI Completes 1-Gigawatt AI Data Center Without Nvidia Chips
zAI built a 1GW AI data center in China with no Nvidia chips, using only domestic silicon. It supports frontier GLM model development and has begun operations.
Alibaba Open-Sources SAIL Stack to Break Nvidia CUDA Lock-In
Alibaba T-Head open-sourced SAIL stack for Zhenwu chips at WAIC, targeting Nvidia CUDA dominance with 7-day migration claim.
NVIDIA's #1 RTEB Embedding Model Skips Token Generation Entirely
NVIDIA's RTEB hit #1 on MTEB by skipping token generation. The cheapest reasoning token is the unused one.
Japan to Buy 27,500 Nvidia Rubin Chips for Robot AI
Japan to buy 27,500 Nvidia Rubin chips for sovereign robot AI model, signaling strategic push in embodied intelligence.
Nvidia Vows 'Giant Amounts' of Vera Rubin as Blackwell Delays Bite
Nvidia CEO Huang pledges 'giant amounts' of Vera Rubin chips, asserting roadmap intact amid Blackwell delays. Japan's $2B+ Rubin factory anchors real demand.
US Allows ZTE to Buy Nvidia H200 AI Chips, Joining Alibaba, Tencent
US authorized ZTE to buy Nvidia H200 AI chips, joining Alibaba, Tencent, ByteDance in accessing Hopper architecture under targeted export controls.
Nvidia Cuts Asia Partner List by Half to Curb AI Chip Smuggling
Nvidia cut authorized Asia customers by half, sending inspectors to data centers. Move follows US pressure to curb AI chip smuggling of H100/Blackwell.