Reports
AI-generated structured vendor updates
AMD Secures 2.5GW Capacity with Core Scientific, Shifts to AI Infrastructure Provider
AMD signs a 15-year agreement with Core Scientific to secure up to 2.5GW of data center capacity, starting with 529MW across five US states. AMD will directly lease 377MW to deploy Instinct GPUs and EPYC CPUs, marking a strategic shift from chip vendor to integrated infrastructure provider, directly competing with NVIDIA's DGX Cloud.
HPE采用第六代AMD EPYC处理器交付业界最高密度AI科学计算超算
...
AMD Unveils 6th Gen EPYC Venice, MI400 GPUs, Helios Rack for AI Inference
AMD launches 6th Gen EPYC Venice (2nm) and MI400 series GPUs, with MI455X claiming 34x token throughput improvement. The Helios rack solution integrates 72 MI455X GPUs with 18 EPYC CPUs via Pensando networking and ROCm software, offering 30% more inference tokens per dollar than competitors. Adopted by OpenAI, Meta, and others.
AMD发布第六代EPYC Venice处理器与Helios机架级AI解决方案
...
AMD Launches Helios Rack-Scale AI Platform with MI400 GPUs, Targeting Inference TCO
At Advancing AI 2026, AMD unveiled the Helios rackscale platform integrating 72 MI455X GPUs and 18 EPYC Venice CPUs per rack, delivering 2.9 exaflops FP4 inference and 31TB HBM4 memory. The MI430X offers 288 TFLOPS FP64 for HPC. AMD claims up to 30% more inference tokens per dollar vs. competitors.
AMD Helios Enters Production: 12-Stack HBM4 Outmuscles NVIDIA, UALoE Opens AI Network
AMD's second-generation Helios rack-scale AI server enters full production, featuring 72 MI455X GPUs with Samsung's exclusive 12-stack HBM4 (31TB per rack). Compared to NVIDIA's 8-stack design, it offers 50% more memory and 30% lower token cost. Microsoft Azure commits to large-scale deployment, solidifying hyperscaler dual-vendor strategy.
AMD and Cerebras Unveil Disaggregated AI Inference with Wafer-Scale Engine
AMD and Cerebras launch a disaggregated AI inference solution combining the Helios Rackscale system (6th-gen EPYC Venice CPUs + up to 72 Instinct MI455X GPUs) with the Cerebras WSE-3 (4 trillion transistors) via Infinity Fabric, targeting ultra-low latency and high throughput for AI inference, challenging traditional GPU clusters.
AMD launches world's first 2nm GPU MI455X and Zen 6 EPYC, targets NVIDIA and Intel
At Advancing AI 2026, AMD announced 46% data center CPU market share and launched the world's first 2nm GPU, Instinct MI455X, with CDNA architecture and HBM4 memory. The 6th-gen EPYC Venice (Zen 6) was also unveiled, targeting AI workloads.
Intel Accelerates 14A to 2027H2 Risk Production, 18A Yield Exceeds Target by 25%, Capex Raised to $20B
Intel announced 14A process acceleration to risk production in 2027H2 with HVM in 2028, 18A yield exceeding target by 25% supporting Panther Lake volume, and raising 2026 capex to $20B, signaling an aggressive foundry push. Advanced packaging EMIB-T becomes a profit pillar.
AMD发布Helios机架级AI平台与MI455X GPU
...
AMD Helios Rack Challenges NVIDIA NVLink with Open UALoE Interconnect
At Advancing AI 2026, AMD launched the Helios rack with 72 MI455X GPUs, 18 Venice EPYC CPUs, and Pensando networking, claiming 30% higher inference token/$ vs NVIDIA NVL72. It introduced UALoE open interconnect to break NVLink lock-in, partnering with Cerebras, Cisco, and major AI firms.
AMD Helios Full-Stack AI Server Deployed on Azure, UALoE Open Standard Challenges NVLink
AMD and Microsoft Azure announce large-scale deployment of Helios full-stack AI servers, featuring 72 MI455X GPUs, 18 EPYC Venice CPUs, and Pensando DPUs, with 31TB HBM4 and 1.7PB/s memory bandwidth. Two new Azure VM series target Agentic AI and semiconductor design, marking Microsoft's multi-vendor strategy and challenging NVIDIA's dominance.
Microsoft Azure Deploys AMD Helios Rack with MI455X GPUs, Launches Three New VM Families
Microsoft Azure announces the deployment of AMD Helios rack-scale AI platform, featuring 72 Instinct MI455X GPUs, 31TB HBM4 memory, and 1.4PB/s bandwidth per rack. Three new VM families target AI inference, data engineering, and HPC, powered by 6th-gen EPYC Venice CPUs and Pensando DPUs.
AMD与Anthropic达成战略合作:部署2GW MI450 GPU并投资50亿美元
...
AMD Invests $5B in Anthropic, Secures 2GW MI450 Deployment, Reshaping AI Compute Ecosystem
AMD and Anthropic announce a strategic partnership: Anthropic will deploy up to 2GW of AMD Instinct MI450 GPUs, with AMD investing up to $5B in Anthropic. They will collaborate on ROCm optimization and Claude workload tuning, marking AMD's transition from chip vendor to AI ecosystem investor and accelerating multi-sourcing in AI compute.
AMD Unveils Zen 6 Venice, MI455X, and Helios Rack-Level Design to Challenge NVIDIA
At Advancing AI 2026, AMD launched Zen 6 EPYC Venice (2nm, up to 256 cores) and MI455X (CDNA5, 432GB HBM4, 40 PFLOPS FP4), along with Helios rack reference design (2.9 exaFLOPS FP4 per rack), claiming a 1000x AI performance roadmap, with major commitments from Meta, OpenAI, and others.
Microsoft Azure Deploys AMD Helios Rack with MI455X GPUs, Breaking NVIDIA's Cloud AI Monopoly
Microsoft Azure officially adopts AMD Helios rack-scale AI infrastructure, featuring 72 MI455X GPUs (432GB HBM4, 19.6TB/s), Venice EPYC CPUs, and Pensando DPUs. Three new instances (ND MI455X v7, HDv2, HXv2) are launched, marking Azure's shift from exclusive NVIDIA dependency to a multi-vendor AI strategy.
AMD and HPE Launch Helios Open AI Infrastructure to Rival NVIDIA Ecosystem
AMD and HPE expand partnership to launch Helios, an open-stack AI infrastructure platform integrating EPYC CPUs, Instinct MI455X GPUs, Pensando networking, and ROCm software. Each rack delivers up to 2.9 exaFLOPS FP4, built on OCP principles with Juniper switches, targeting simplified deployment and energy efficiency.
AMD Confirms Zen 6 EPYC Venice: First 2nm Server CPU Launching July 2026
AMD confirms Zen 6 EPYC Venice launch at Advancing AI 2026 (July 22-23). As the first 2nm server CPU, it features triple-core hybrid architecture, up to 192 cores, ~29% single-thread and ~22% multi-thread gains, targeting AI inference and tight CPU-GPU synergy via Infinity Fabric.
AMD's Experimental Topological Ghost Protocol Boosts MI300X Inference 10x
AMD introduces experimental Topological Ghost Protocol (TGP) on MI300X GPUs, achieving 431 tokens/sec with 100% success in high-concurrency inference, 10x improvement over standard vLLM. TGP uses KV-cache recycling and segmented state management, still experimental but potentially redefining AI inference benchmarks.