hpc
30 articles about hpc in AI news
Bull Delivers HPC Infrastructure to Power Mimer AI Factory
Bull, a subsidiary of Atos, has supplied the core HPC infrastructure for Mimer's new AI factory. This facility is dedicated to training and developing large language models for the European market.
FRAGATA: A Hybrid RAG System for Semantic Search Over 20 Years of HPC
A new paper details FRAGATA, a system enabling semantic search over two decades of technical support tickets at a supercomputing center. It uses hybrid retrieval-augmented generation (RAG) to find relevant past incidents despite typos, language, or wording differences, showing a qualitative improvement over the legacy search.
NVIDIA Vera Rubin: One Rack Matches TOP500, 35 EU Labs Deploy
NVIDIA's Vera Rubin NVL72 delivers TOP500-class performance in a single rack, with 35 European labs deploying the system for AI and HPC.
HPE Slingshot Leads Supercomputer Interconnects; China's 1.2 Exaflops
HPE's Slingshot tops supercomputer interconnects, but China's 1.2 exaflops machine steals the show, signaling a tightening race in HPC.
AMD-Cerebras Disaggregated Inference: 5× T/s/W, Prompt vs. Decode Split
AMD and Cerebras launched a disaggregated inference platform splitting prompt processing on Helios from decode on WSE, claiming up to 5× T/s/W.
AMD to Supply Anthropic with 2GW of MI450 GPUs, Invest Up to $5B
AMD to supply Anthropic with 2GW MI450 GPUs in H1 2027 and invest up to $5B, challenging Nvidia's AI hardware dominance.
Nvidia Vera CPU Hits SPECrate 2026: 1.7× AMD Epyc 9755
Nvidia's Vera CPU scored 1.7× SPECrate integer 2026 vs AMD Epyc 9755. First custom core for agentic AI, H2 2026 release.
Google Chooses Intel EMIB-T for 9th-Gen TPUs, Breaking TSMC's CoWoS Monopoly
Google picks Intel EMIB-T for 9th-gen TPU, breaking TSMC CoWoS monopoly. Move signals architectural bet on power integrity and reticle-free scaling.
Chinese AI Firms Raise $20B in Hong Kong Amid US Chip Curbs
Chinese tech firms raised $20B in Hong Kong for AI and semiconductor expansion, driven by US chip export controls.
Cerebras, Flex Expand CS-3 Production 7x at Milpitas Facility
Cerebras and Flex expand CS-3 production 7x at Milpitas facility. The partnership keeps wafer-scale AI manufacturing in the U.S. as Nvidia faces delays.
California Gov. Newsom Partners Anthropic for State AI Tools
California partners with Anthropic for state AI tools targeting tax, health, DMV services. No cost or timeline disclosed; deal tests AI safety branding in public sector.
NHN Cloud Tops Korean TOP500 with FactoryX GPU Clusters
NHN Cloud tops Korean TOP500 with FactoryX GPU clusters delivering 1.2 exaflops, marking first domestic cloud provider to lead the list.
Anthropic Launches Claude Tag as Multiplayer Slack Agent Ahead of IPO
Anthropic released Claude Tag, a Slack-native agent for teams. The tool already approves 65% of internal code changes as Anthropic pushes enterprise adoption ahead of its 2026 IPO.
VIAVI Ships First Ultra Ethernet Validation Tool for AI Data Centers
VIAVI launched the first Ultra Ethernet validation tool for AI data centers, supporting 800GE/1.6TE links. The tool enables certification of low-latency, lossless transport critical for distributed AI training.
JUPITER Exascale Maps Brain at Cellular Scale on 4,096 Grace Hopper Nodes
JUPITER, Europe's first exascale supercomputer, trained CytoNet brain model on 6.5 PB in 5 days and runs climate, 6G, and quantum simulations.
LANL Taps NVIDIA Vera CPUs for 7x Agentic AI Speed on Scientific Workloads
LANL selects NVIDIA Vera CPUs for three supercomputers, claiming 7x performance on agentic AI workloads over x86. Systems Mission, Vision, Veritas deploy by 2027.
AWS Beats Cloud Rivals to NVIDIA Blackwell with EC2 G7 — 4.6x AI Inference Gain Over G6
AWS launched EC2 G7 instances on June 19, 2026, becoming the first major cloud to offer NVIDIA RTX PRO 4500 Blackwell Server Edition GPUs. The instances claim 4.6x AI inference performance over G6, backed by 700 Gbps EFA networking and 32 GB GDDR7 per GPU. The move arrives the same week AWS confirme
Lansing AI data center petition hits 20,000 signatures
Over 20,000 signatures oppose a Lansing AI data center tied to Nvidia's Vera Rubin build-out, signaling growing local resistance.
NVIDIA, GENCI Launch AI Factory France Compute Access for Startups
NVIDIA and GENCI launched AI Factory France at VivaTech, giving European startups free access to AI supercomputers. The program includes compute, tools, and expert support for NVIDIA Inception members.
AI Infrastructure Hit $300B in 2025, Forecast to Exceed $520B by 2030
AI infrastructure spending hit $300B in 2025, up 60.1% YoY, and is forecast to exceed $520B by 2030, with sovereign AI emerging as the fastest-growing segment.
Intel Omni-Path Resurfaces as InfiniBand Rival for DoE Supercomputers
Intel's Omni-Path interconnect, revived by Cornelis Networks, will connect DoE supercomputers at 400Gbps as an InfiniBand alternative.
IREN Buys Nostrum Group, Adds 490MW Spanish Grid Power for AI Cloud
IREN acquired Nostrum Group, adding 490MW grid power in Spain for AI cloud expansion.
Schneider Electric & Foxconn Partner on AI Data Center Infrastructure
Schneider Electric and Foxconn announced a strategic collaboration to co-develop next-gen AI data center infrastructure, including reference architectures and modular power/cooling skids. Production begins later this year.
KKR Launches Helix Digital Infrastructure with $10B for AI Buildout
KKR launched Helix Digital Infrastructure with over $10B in commitments from KKR, KIA, Nvidia, and Vistra to bundle AI data centers, power, and connectivity for hyperscalers.
Trillion Labs Builds Industrial World Models on NVIDIA Omnibus
Trillion Labs announced Industrial World Models for AI Factories using NVIDIA Omniverse and Nemotron to optimize data centers and power plants.
JPMorgan, OQC, AMD Build First Quantum AI Data Center for Finance
JPMorgan, OQC, and AMD are building a dedicated quantum AI data center for financial workflows, moving from remote-access demos to enterprise-grade infrastructure. No budget or timeline disclosed.
Liquid Cooling Hits 15kW: CoolIT Coldplate Quadruples Capacity for AI
CoolIT demoed a 15kW single-phase coldplate, quadrupling capacity, while Vertiv, Accelsius, and LiquidStack launched products targeting scalable AI cooling deployment.
Liquid Cooling Crosses 50% by 2027? Rack Densities Force Shift
AI-driven rack densities are pushing liquid cooling adoption past 50% in new hyperscale builds by 2027, though cost and expertise remain barriers.
Ayar Labs Joins NVIDIA NVLink Fusion Ecosystem for Co-Packaged Optics
Ayar Labs joined NVIDIA's NVLink Fusion ecosystem to bring co-packaged optics to AI factories, following its $500M Series E and alongside Lightmatter's similar move.
Google and Blackstone Launch TPU Venture, Challenging Nvidia Dominance
Google and Blackstone launched a TPU venture, financing AI infrastructure outside the hyperscale cloud model. Enterprise buyers get a standalone alternative to Nvidia-dominated GPU clusters.