Filter

×
Active Filters Clear All
Keyword: GPU ×
363 Total Reports
1/19 Page
AMD Other 2026-07-29

AMD Secures 2.5GW Capacity with Core Scientific, Shifts to AI Infrastructure Provider

AMD signs a 15-year agreement with Core Scientific to secure up to 2.5GW of data center capacity, starting with 529MW across five US states. AMD will directly lease 377MW to deploy Instinct GPUs and EPYC CPUs, marking a strategic shift from chip vendor to integrated infrastructure provider, directly competing with NVIDIA's DGX Cloud.

NVIDIA Other 2026-07-28

NVIDIA通知合作伙伴GPU套件全线涨价

...

AMD Other 2026-07-28

AMD Unveils 6th Gen EPYC Venice, MI400 GPUs, Helios Rack for AI Inference

AMD launches 6th Gen EPYC Venice (2nm) and MI400 series GPUs, with MI455X claiming 34x token throughput improvement. The Helios rack solution integrates 72 MI455X GPUs with 18 EPYC CPUs via Pensando networking and ROCm software, offering 30% more inference tokens per dollar than competitors. Adopted by OpenAI, Meta, and others.

NVIDIA Other 2026-07-28

NVIDIA Invests $5B in SSI, Opens Vera Rubin Platform to Lock In AI Safety Research

NVIDIA makes a major equity investment in Safe Superintelligence (SSI) and provides access to its next-generation Vera Rubin GPU platform. The partnership goes beyond hardware sales, giving NVIDIA rare access to SSI's confidential research, with insights feeding back into NVIDIA's platform roadmap, marking a strategic shift from hardware vendor to deep research partner.

NVIDIA Other 2026-07-28

NVIDIA Vera Rubin NVL72 Enters Production, 10x Energy Efficiency Reshapes AI Infrastructure

NVIDIA has announced full production of the Vera Rubin NVL72 platform, with first shipments to CoreWeave, Microsoft, Amazon, and Oracle. Each rack integrates 72 Rubin GPUs and 36 Vera CPUs, delivering 10x improvement in energy efficiency and inference cost over Blackwell for agentic AI workloads, marking a new era of rack-scale AI infrastructure.

NVIDIA Other 2026-07-27

NVIDIA GPU套装全线涨价涉及GDDR7与GDDR6显存

...

AMD Other 2026-07-27

AMD发布第六代EPYC Venice处理器与Helios机架级AI解决方案

...

Microsoft Azure Other 2026-07-26

Microsoft Azure Integrates AMD Helios Rack-Scale AI Platform, Breaks NVIDIA GPU Monopoly

Microsoft Azure announces deployment of AMD Helios rack-scale AI platform in H2 2026. Rack integrates 72 MI455X GPUs, 18 Venice CPUs, liquid cooling, delivering 2.9 Exaflops FP4 inference. This signals a major industry shift away from NVIDIA GPU monopoly towards multi-vendor heterogeneous AI infrastructure.

AMD Other 2026-07-26

AMD Launches Helios Rack-Scale AI Platform with MI400 GPUs, Targeting Inference TCO

At Advancing AI 2026, AMD unveiled the Helios rackscale platform integrating 72 MI455X GPUs and 18 EPYC Venice CPUs per rack, delivering 2.9 exaflops FP4 inference and 31TB HBM4 memory. The MI430X offers 288 TFLOPS FP64 for HPC. AMD claims up to 30% more inference tokens per dollar vs. competitors.

Microsoft Other 2026-07-26

Microsoft Launches MAI Models, Slashes GPU Costs 89%, Reducing OpenAI Dependency

Microsoft unveiled MAI-Image-2.5-Pro and MAI-Voice-2-Flash on Azure Foundry, achieving 96.8% text rendering accuracy at 8K and reducing GPU costs by 84-89% vs GPT. Integrated across Bing, PowerPoint, and Dynamics 365, it marks a strategic shift from OpenAI dependency. Also, NVIDIA Jetson heads to the moon for edge AI.

AMD Other 2026-07-26

AMD Helios Enters Production: 12-Stack HBM4 Outmuscles NVIDIA, UALoE Opens AI Network

AMD's second-generation Helios rack-scale AI server enters full production, featuring 72 MI455X GPUs with Samsung's exclusive 12-stack HBM4 (31TB per rack). Compared to NVIDIA's 8-stack design, it offers 50% more memory and 30% lower token cost. Microsoft Azure commits to large-scale deployment, solidifying hyperscaler dual-vendor strategy.

Anthropic Other 2026-07-26

Anthropic partners with SpaceX for 220K GPUs, shifting AI compute ecosystem beyond hyperscalers

Anthropic signs a deal with SpaceX for 300MW capacity and 220,000 NVIDIA GPUs at Colossus 1 data center, with exploration of orbital AI compute. This diversifies Anthropic's compute sources beyond hyperscalers and doubles Claude Code usage limits.

Meta Other 2026-07-26

Meta Transforms into AI Compute Landlord with $10B Anthropic Lease Deal

Meta is in early talks with Anthropic for a $10 billion compute lease deal, marking its strategic shift from internal AI infrastructure user to commercial landlord. This move monetizes Meta's massive capex and creates an inter-cloud leasing model, fundamentally reshaping the AI compute supply chain.

NVIDIA Other 2026-07-26

NVIDIA and SK Group Lock HBM4 Supply and Launch Sovereign AI Factory Model with $500B+ Deal

NVIDIA and SK Group announced a $500B+ AI partnership including a 2GW AI factory using Vera Rubin and HBM4, long-term HBM4 supply lock, and a $1B NVIDIA investment in Naver (with $9B from Brookfield). Samsung and Broadcom signed a $200B deal. This signals a new era of sovereign AI infrastructure and supply chain deep-locking.

AMD Other 2026-07-25

AMD and Cerebras Unveil Disaggregated AI Inference with Wafer-Scale Engine

AMD and Cerebras launch a disaggregated AI inference solution combining the Helios Rackscale system (6th-gen EPYC Venice CPUs + up to 72 Instinct MI455X GPUs) with the Cerebras WSE-3 (4 trillion transistors) via Infinity Fabric, targeting ultra-low latency and high throughput for AI inference, challenging traditional GPU clusters.

AMD Other 2026-07-25

AMD launches world's first 2nm GPU MI455X and Zen 6 EPYC, targets NVIDIA and Intel

At Advancing AI 2026, AMD announced 46% data center CPU market share and launched the world's first 2nm GPU, Instinct MI455X, with CDNA architecture and HBM4 memory. The 6th-gen EPYC Venice (Zen 6) was also unveiled, targeting AI workloads.

NVIDIA Other 2026-07-25

NVIDIA Prepay $1.5B to Amkor for US 2nm/3nm/HBM4 Advanced Packaging

NVIDIA prepays $1.5B to Amkor to expand advanced packaging capacity in Arizona, covering 2nm/3nm/HBM4. This onshores AI chip packaging, complementing Wistron's system integration, to reduce reliance on Taiwan and secure supply for Vera Rubin and Blackwell Ultra.

NVIDIA Other 2026-07-25

NVIDIA and Wistron Launch US-Based GB300 Production, Shifting AI Manufacturing Landscape

NVIDIA partnered with Wistron to open a $700M factory in Texas, achieving L6 integration of the GB300 Grace Blackwell Ultra system. The first US-made unit contains 1.5M parts, weighs 2 tons, and costs $4M. This milestone accelerates US AI capex localization from 5% to 30%, reshaping the global AI supply chain.

NVIDIA Other 2026-07-25

Tiered AI Chip Market Emerges as US Allows H200 Exports to China with 25% Levy

The US Commerce Department approved NVIDIA H200 exports to China with a 25% sales tax, while maintaining a ban on Blackwell. This formalizes a tiered AI chip market, making H200 the best available imported chip for China, but the performance gap and added tax burden increase deployment costs and complexity for Chinese AI infrastructure.

AMD Other 2026-07-24

AMD发布Helios机架级AI平台与MI455X GPU

...