Skip to content
gentic.news — AI News Intelligence Platform
Connecting to the Living Graph…

Listen to today's AI briefing

Daily podcast — 5 min, AI-narrated summary of top stories

Rack of Nvidia Grace standalone servers in a data center, with cooling pipes and cables, indicating high-density…

Nvidia Ships Hundreds of Thousands of Grace Standalone Servers

Nvidia shipped hundreds of thousands of Grace standalone servers. The CPU pivot targets agentic AI workloads shifting hardware balance.

·16h ago·3 min read··8 views·AI-Generated·Report error
Share:
Source: tomshardware.comvia tomshardwareSingle Source
How many Grace standalone servers has Nvidia shipped?

Nvidia shipped hundreds of thousands of Grace standalone servers, per VP Ian Buck, as CPUs gain prominence in agentic AI data centers. The company previously disclosed 2.5 million Grace CPUs shipped total.

TL;DR

Nvidia shipped hundreds of thousands of Grace standalone servers. · Grace CPUs target data-rich backend workloads, not web servers. · Vera CPU uses monolithic die, custom cores, and high-bandwidth fabric.

Nvidia shipped hundreds of thousands of Grace standalone servers, VP Ian Buck revealed. The GPU giant pivots messaging as agentic AI workloads shift CPU-to-GPU ratios toward one-to-one.

Key facts

  • Nvidia shipped hundreds of thousands of Grace standalone servers.
  • Over 2.5 million Grace CPUs shipped total, per May disclosure.
  • Grace uses 72 Arm Neoverse V2 cores with SCF fabric.
  • Vera features 88 monolithic cores with 3.4 TB/s bandwidth.
  • Agentic AI shifts CPU-to-GPU ratio toward one-to-one.

Nvidia wants you to know it's a CPU company too. Ian Buck, vice president of hyperscale and high-performance computing and inventor of CUDA, said the company has "shipped... let's put it in the hundreds of thousands of Grace standalone servers" [According to Tom's Hardware]. In May, Nvidia disclosed it had shipped over 2.5 million Grace CPUs total, and announced a partnership with Meta to deploy standalone Grace servers in February. Buck's comments suggest even larger scale, as Nvidia competes with Intel and AMD in data center CPUs.

Grace uses 72 stock Arm Neoverse V2 cores, differentiated by Nvidia's Scalable Coherency Fabric (SCF). Buck said the CPUs are "being deployed for the backend, data-rich operations, like the data processing," not cheap web servers. This positions Grace as an on-ramp for Nvidia's next-generation Vera CPU, which features Nvidia's first custom core design, Olympus, and a monolithic 88-core die with 3.4 TB/s fabric bandwidth.

Key Takeaways

  • Nvidia shipped hundreds of thousands of Grace standalone servers.
  • The CPU pivot targets agentic AI workloads shifting hardware balance.

Agentic AI Reshapes Hardware Balance

Evolving agentic AI workloads have changed the hardware balance, shifting from as many as eight GPUs per CPU toward a one-to-one ratio in some cases. This trend has wiped around $1 trillion from Nvidia's market cap since its peak earlier this year, as investors rally behind CPU makers like Intel. Nvidia's Vera CPU, architected specifically for agentic workloads, aims to ride that train.

Vera's monolithic design contrasts sharply with Intel and AMD's chiplet approaches, which trade latency and coherency for core density. "One of the reasons we don't have 128 cores is because we've dedicated so much of the die area toward the fabric," Buck said. The 3.4 TB/s internal bandwidth allows every core to communicate efficiently, a design choice optimized for data-intensive agentic AI tasks.

Competitive Landscape Heats Up

AMD is expected to launch its Zen 6 Venice CPUs this week, intensifying the battle. Meanwhile, Nvidia's Vera Rubin NVL72 rack system faced delays to 2028, per recent reports. The company's CPU push comes amid broader AI infrastructure buildout, with hyperscaler off-balance-sheet debt hitting $1.65 trillion across five tech giants, as previously reported.

An Nvidia Vera CPU

The question is whether Nvidia's monolithic bet pays off against Intel and AMD's chiplet ecosystems. Grace cracked the door; Vera represents Nvidia's big entrance.

What to watch

Watch AMD's Zen 6 Venice launch this week and Nvidia's Vera Rubin NVL72 delivery timeline, especially whether enterprise adoption of standalone Grace servers accelerates beyond the hundreds of thousands mark.

Nvidia Vera CPU


Source: tomshardware.com


Sources cited in this article

  1. May
  2. CPU
  3. Nvidia
Source: gentic.news · · author= · citation.json

AI-assisted reporting. Generated by gentic.news from 4 verified sources, fact-checked against the Living Graph of 4,300+ entities. Edited by Ala SMITH.

Following this story?

Get a weekly digest with AI predictions, trends, and analysis — free.

AI Analysis

Nvidia's CPU play is a strategic hedge against market cap erosion and changing workload profiles. The company's monolithic die approach for Vera is a bet that coherency and bandwidth matter more than core count for agentic AI. This contrasts sharply with Intel and AMD's chiplet strategy, which optimizes for density at the cost of latency. The timing is notable. With $1 trillion in market cap wiped since peak and hyperscaler debt ballooning, Nvidia needs to diversify beyond GPUs. Grace's deployment in data-rich backend operations suggests Nvidia is targeting a specific niche rather than competing head-on with Intel's broad x86 ecosystem. The Meta partnership validates this approach. However, the monolithic design limits core counts—Vera has 88 cores versus AMD's potential 128+ in chiplet designs. If agentic AI workloads scale to require more parallel threads, Nvidia's fabric advantage may not compensate. The real test will be whether Vera's custom Olympus cores and SCF deliver enough performance uplift to justify the architectural divergence.
Compare side-by-side
Nvidia vs Meta
Enjoyed this article?
Share:

AI Toolslive

Five one-click lenses on this article. Cached for 24h.

Pick a tool above to generate an instant lens on this article.

Related Articles

From the lab

The framework underneath this story

Every article on this site sits on top of one engine and one framework — both built by the lab.

More in Products & Launches

View all