scale ai
30 articles about scale ai in AI news
DeepSeek Builds Gigawatt-Scale AI Data Center in Inner Mongolia
DeepSeek is building a gigawatt-scale AI data center in Inner Mongolia, per Bloomberg. The project marks a strategic pivot from efficiency to raw compute scale.
Microsoft to Deploy AMD Helios Rack-Scale AI at Scale on Azure
Microsoft will deploy AMD's Helios rack-scale AI accelerator at scale on Azure, powered by MI455X GPUs and Epyc Venice CPUs. The move diversifies Azure's AI silicon beyond Nvidia.
Upscale AI Raises $500M for AI-Native Networking Silicon
Upscale AI raised $500M for AI networking silicon, with Google Cloud as a strategic partner. The deal targets GPU cluster interconnect bottlenecks.
Upscale AI Raises $190M for AI Networking Infrastructure
Upscale AI raised $190M to expand AI networking infrastructure, addressing the bottleneck of 100K+ GPU clusters.
Box Elder County to Vote on Hyperscale AI Data Center After Delay
Box Elder County votes on hyperscale AI data center after delay. Decision tests local government balance between infrastructure demand and resource constraints.
Meta to Release First LLM Built Under Scale AI's Alexandr Wang
Meta is set to release its first large language model developed under the technical leadership of Scale AI founder Alexandr Wang. While not fully open initially, the company plans to eventually open-source versions of the new model family.
Analysis: Meta's AI Investment Strategy Questioned as Scale AI Acquihire and Data Center Spend Top $700B
An analysis estimates Meta's total AI investment at ~$700B, including a ~$14.3M Scale AI acquihire and over $600B in data centers. The post questions why this has not yielded a competitive upcoming model against Chinese open-source labs.
Tulip and Salesfloor Merge to Scale AI-Powered Retail Engagement
Tulip, a mobile retail platform, and Salesfloor, a clienteling and virtual selling solution, have announced a merger. The combined entity aims to scale AI-powered customer engagement for retailers, focusing on unifying in-store and online experiences.
Perplexity's pplx-embed: The Bidirectional Breakthrough Transforming Web-Scale AI Retrieval
Perplexity has launched pplx-embed, a new family of multilingual embedding models that set state-of-the-art benchmarks for web-scale retrieval. Built on Qwen3 architecture with bidirectional attention, these models specifically address the noise and complexity of real-world web data.
Throughput Optimization as a Strategic Lever in Large-Scale AI Systems
A new arXiv paper argues that optimizing data pipeline and memory throughput is now a strategic necessity for training large AI models, citing specific innovations like OVERLORD and ZeRO-Offload that deliver measurable efficiency gains.
Nvidia Bets Big on Thinking Machines Lab with Gigawatt-Scale AI Partnership
Nvidia has formed a strategic partnership with Thinking Machines Lab, led by former OpenAI CTO Mira Murati, committing to deploy at least one gigawatt of next-generation Vera Rubin systems. The multiyear deal includes significant investment and aims to accelerate frontier AI development while expanding access to customizable models.
Whering Secures $7M from eBay Ventures and Google AI Futures Fund
Whering raised $7M from eBay Ventures and Google AI Futures Fund, reaching 10M users. The funding will scale AI-powered wardrobe tech for personalized, sustainable fashion.
Nebius Claims First NVIDIA GB300 Exemplar Cloud for Training
Nebius becomes first cloud provider validated as NVIDIA Exemplar Cloud on GB300 for training, targeting hyperscale AI workloads.
Aehr Test Systems Lands $41M AI Chip Order; H2 Bookings Top $92M
Aehr Test Systems received a record $41 million production order from a key hyperscale AI customer. Total bookings for the second half of its fiscal year exceeded $92 million, highlighting surging demand for semiconductor test and burn-in equipment.
Cloudflare Agent Cloud Integrates OpenAI GPT-5.4 & Codex for Enterprise AI
Cloudflare has integrated OpenAI's GPT-5.4 and Codex models into its Agent Cloud platform. This allows enterprises to build, deploy, and scale AI agents for production workflows with built-in security and performance.
AI System Claims 100x Energy Efficiency Gain with Higher Accuracy
A new AI system reportedly uses 100 times less energy than current models while achieving higher accuracy. If validated, this could significantly reduce the operational costs and environmental impact of large-scale AI deployment.
Google's AI Infrastructure Strategy: What Retail Leaders Should Watch in 2026
Google's evolving AI infrastructure and compute strategy, including data center investments and model compression techniques, will directly impact how retail brands deploy and scale AI applications by 2026. The company's focus on efficiency and real-time capabilities signals a shift toward more accessible, powerful retail AI tools.
AI Agents Get a Memory Upgrade: New Research Tackles Long-Horizon Task Challenges
Researchers have developed new methods to scale AI agent memory for complex, long-horizon tasks. The breakthrough addresses one of the biggest limitations in current agent systems—their inability to retain and utilize information over extended sequences of actions.
Cerebras, Flex Expand CS-3 Production 7x at Milpitas Facility
Cerebras and Flex expand CS-3 production 7x at Milpitas facility. The partnership keeps wafer-scale AI manufacturing in the U.S. as Nvidia faces delays.
Realty Income Launches $6B Data Center JV with Cloud Capital
Realty Income formed a $6B data center JV with Cloud Capital and an institutional investor, signaling REITs are normalizing hyperscale AI infrastructure as core assets.
Google's Virgo Network Links 134,000 TPU v8 Chips with 47 Pbps Fabric
Google unveiled its Virgo networking stack for TPU v8, capable of linking 134,000 chips in a single fabric with 47 petabits/sec of bi-sectional bandwidth. This represents a massive scale-up in interconnect technology for large-scale AI model training.
TradeBeyond: Why Supply Chain Traceability Fails at Scale — and the Fix
TradeBeyond's Just Style piece argues traceability fails at scale due to fragmented data. The fix: a unified, interoperable platform consolidating supplier, material, and compliance data into one source of truth, enabling brands to move beyond pilots.
Nscale Acquires Anyscale, Adding Ray Creator to AI Cloud Stack
Nscale acquires Anyscale, adding Ray's software layer to its full-stack AI cloud. The ~200-person team joins, with terms undisclosed.
Hyperscalers' AI Data Center Spend Traps Them in a Vicious Cycle
Ed Zitron argues hyperscalers' AI data center spending traps them in a cycle where more spend leads to more losses, as competitive pressure forces unsustainable capital outlays.
Crusoe and ON.energy to Deploy 5 GW of AI UPS at Hyperscale Campuses
Crusoe and ON.energy will deploy 5 GW of gas-fired AI UPS with carbon capture at hyperscale campuses, bypassing grid constraints for AI training clusters.
AGCO scales employee-built AI agents with Microsoft Copilot Studio
AGCO scaled employee-built AI agents using Microsoft Copilot Studio, growing from 3 agents to 500+ use cases. This shows how low-code tools can democratize AI in enterprise settings.
Scale-Across: Cloud Giants Link Datacenters for Million-Accelerator AI Clusters
Cloud providers are linking multiple datacenters for million-accelerator AI clusters, a new 'scale-across' paradigm.
Vibe Coding Fails: Why AI-Generated Code Breaks at Scale
Vibe coding fails because AI-generated code lacks architectural coherence, test coverage, and security validation, breaking at scale beyond 1,000 lines.
AI Data Center Scale Doubles Every 7 Months, Epoch Finds
Epoch AI finds AI data center scale doubles every 7 months, driven by Google, Microsoft, and Amazon investments. This accelerates beyond the earlier 12-month cycle, raising training cost projections to $10 billion by 2028.
JUPITER Exascale Maps Brain at Cellular Scale on 4,096 Grace Hopper Nodes
JUPITER, Europe's first exascale supercomputer, trained CytoNet brain model on 6.5 PB in 5 days and runs climate, 6G, and quantum simulations.