Reports
AI-generated structured vendor updates
AMD Secures 2.5GW Capacity with Core Scientific, Shifts to AI Infrastructure Provider
AMD signs a 15-year agreement with Core Scientific to secure up to 2.5GW of data center capacity, starting with 529MW across five US states. AMD will directly lease 377MW to deploy Instinct GPUs and EPYC CPUs, marking a strategic shift from chip vendor to integrated infrastructure provider, directly competing with NVIDIA's DGX Cloud.
NVIDIA通知合作伙伴GPU套件全线涨价
...
AMD Unveils 6th Gen EPYC Venice, MI400 GPUs, Helios Rack for AI Inference
AMD launches 6th Gen EPYC Venice (2nm) and MI400 series GPUs, with MI455X claiming 34x token throughput improvement. The Helios rack solution integrates 72 MI455X GPUs with 18 EPYC CPUs via Pensando networking and ROCm software, offering 30% more inference tokens per dollar than competitors. Adopted by OpenAI, Meta, and others.
NVIDIA Invests $5B in SSI, Opens Vera Rubin Platform to Lock In AI Safety Research
NVIDIA makes a major equity investment in Safe Superintelligence (SSI) and provides access to its next-generation Vera Rubin GPU platform. The partnership goes beyond hardware sales, giving NVIDIA rare access to SSI's confidential research, with insights feeding back into NVIDIA's platform roadmap, marking a strategic shift from hardware vendor to deep research partner.
NVIDIA Vera Rubin NVL72 Enters Production, 10x Energy Efficiency Reshapes AI Infrastructure
NVIDIA has announced full production of the Vera Rubin NVL72 platform, with first shipments to CoreWeave, Microsoft, Amazon, and Oracle. Each rack integrates 72 Rubin GPUs and 36 Vera CPUs, delivering 10x improvement in energy efficiency and inference cost over Blackwell for agentic AI workloads, marking a new era of rack-scale AI infrastructure.
NVIDIA GPU套装全线涨价涉及GDDR7与GDDR6显存
...
AMD发布第六代EPYC Venice处理器与Helios机架级AI解决方案
...
Microsoft Azure Integrates AMD Helios Rack-Scale AI Platform, Breaks NVIDIA GPU Monopoly
Microsoft Azure announces deployment of AMD Helios rack-scale AI platform in H2 2026. Rack integrates 72 MI455X GPUs, 18 Venice CPUs, liquid cooling, delivering 2.9 Exaflops FP4 inference. This signals a major industry shift away from NVIDIA GPU monopoly towards multi-vendor heterogeneous AI infrastructure.
AMD Launches Helios Rack-Scale AI Platform with MI400 GPUs, Targeting Inference TCO
At Advancing AI 2026, AMD unveiled the Helios rackscale platform integrating 72 MI455X GPUs and 18 EPYC Venice CPUs per rack, delivering 2.9 exaflops FP4 inference and 31TB HBM4 memory. The MI430X offers 288 TFLOPS FP64 for HPC. AMD claims up to 30% more inference tokens per dollar vs. competitors.
Microsoft Launches MAI Models, Slashes GPU Costs 89%, Reducing OpenAI Dependency
Microsoft unveiled MAI-Image-2.5-Pro and MAI-Voice-2-Flash on Azure Foundry, achieving 96.8% text rendering accuracy at 8K and reducing GPU costs by 84-89% vs GPT. Integrated across Bing, PowerPoint, and Dynamics 365, it marks a strategic shift from OpenAI dependency. Also, NVIDIA Jetson heads to the moon for edge AI.
AMD Helios Enters Production: 12-Stack HBM4 Outmuscles NVIDIA, UALoE Opens AI Network
AMD's second-generation Helios rack-scale AI server enters full production, featuring 72 MI455X GPUs with Samsung's exclusive 12-stack HBM4 (31TB per rack). Compared to NVIDIA's 8-stack design, it offers 50% more memory and 30% lower token cost. Microsoft Azure commits to large-scale deployment, solidifying hyperscaler dual-vendor strategy.
Anthropic partners with SpaceX for 220K GPUs, shifting AI compute ecosystem beyond hyperscalers
Anthropic signs a deal with SpaceX for 300MW capacity and 220,000 NVIDIA GPUs at Colossus 1 data center, with exploration of orbital AI compute. This diversifies Anthropic's compute sources beyond hyperscalers and doubles Claude Code usage limits.
Meta Transforms into AI Compute Landlord with $10B Anthropic Lease Deal
Meta is in early talks with Anthropic for a $10 billion compute lease deal, marking its strategic shift from internal AI infrastructure user to commercial landlord. This move monetizes Meta's massive capex and creates an inter-cloud leasing model, fundamentally reshaping the AI compute supply chain.
NVIDIA and SK Group Lock HBM4 Supply and Launch Sovereign AI Factory Model with $500B+ Deal
NVIDIA and SK Group announced a $500B+ AI partnership including a 2GW AI factory using Vera Rubin and HBM4, long-term HBM4 supply lock, and a $1B NVIDIA investment in Naver (with $9B from Brookfield). Samsung and Broadcom signed a $200B deal. This signals a new era of sovereign AI infrastructure and supply chain deep-locking.
AMD and Cerebras Unveil Disaggregated AI Inference with Wafer-Scale Engine
AMD and Cerebras launch a disaggregated AI inference solution combining the Helios Rackscale system (6th-gen EPYC Venice CPUs + up to 72 Instinct MI455X GPUs) with the Cerebras WSE-3 (4 trillion transistors) via Infinity Fabric, targeting ultra-low latency and high throughput for AI inference, challenging traditional GPU clusters.
AMD launches world's first 2nm GPU MI455X and Zen 6 EPYC, targets NVIDIA and Intel
At Advancing AI 2026, AMD announced 46% data center CPU market share and launched the world's first 2nm GPU, Instinct MI455X, with CDNA architecture and HBM4 memory. The 6th-gen EPYC Venice (Zen 6) was also unveiled, targeting AI workloads.
NVIDIA Prepay $1.5B to Amkor for US 2nm/3nm/HBM4 Advanced Packaging
NVIDIA prepays $1.5B to Amkor to expand advanced packaging capacity in Arizona, covering 2nm/3nm/HBM4. This onshores AI chip packaging, complementing Wistron's system integration, to reduce reliance on Taiwan and secure supply for Vera Rubin and Blackwell Ultra.
NVIDIA and Wistron Launch US-Based GB300 Production, Shifting AI Manufacturing Landscape
NVIDIA partnered with Wistron to open a $700M factory in Texas, achieving L6 integration of the GB300 Grace Blackwell Ultra system. The first US-made unit contains 1.5M parts, weighs 2 tons, and costs $4M. This milestone accelerates US AI capex localization from 5% to 30%, reshaping the global AI supply chain.
Tiered AI Chip Market Emerges as US Allows H200 Exports to China with 25% Levy
The US Commerce Department approved NVIDIA H200 exports to China with a 25% sales tax, while maintaining a ban on Blackwell. This formalizes a tiered AI chip market, making H200 the best available imported chip for China, but the performance gap and added tax burden increase deployment costs and complexity for Chinese AI infrastructure.
AMD发布Helios机架级AI平台与MI455X GPU
...