Reports
AI-generated structured vendor updates
Meta launches Meta Compute to sell idle AI capacity, partners with BlackRock on $14B data center
Meta announces Meta Compute, selling idle AI capacity externally, and partners with BlackRock on a $14B data center. Meta is in talks to lease $10B in compute to Anthropic. With 2026 capex raised to $125-145B, Meta transforms from a social platform to an AI infrastructure provider.
Microsoft signs $130B+ data center leases, locking in AI compute capacity
Microsoft signed over $130 billion in new data center leases in Q4 2026, securing capacity for Azure, Copilot, and OpenAI. Total lease obligations for unoperated facilities reached $329.1B, signaling long-term bets on AI demand. Azure annual revenue exceeded $100B for the first time, with Copilot paid seats surpassing 30 million.
Google Guarantees Anthropic Data Center, Deploys TPUs to Challenge NVIDIA's GPU Dominance
Anthropic is building a $15B AI data center in Texas, backed by Google's financial guarantee in exchange for 20% equity. The facility will deploy Google's TPU chips, co-designed with Broadcom, signaling a shift in AI compute landscape and challenging NVIDIA's GPU dominance.
NVIDIA Guarantees $250B for OpenAI Data Center, Shifts from Chip Vendor to AI Financier
NVIDIA plans to provide $250B in guarantees for OpenAI's 10GW data center and $350B in chip financing. This shifts NVIDIA from a GPU vendor to an AI infrastructure credit intermediary, locking OpenAI into NVIDIA's compute ecosystem and creating a closed-loop model.
Microsoft Takes Over OpenAI's Arctic Project, Deploys 30K Vera Rubin GPUs
Microsoft partners with Nscale to take over OpenAI's shelved Arctic data center in Narvik, deploying 30,000 NVIDIA Vera Rubin chips. This signals OpenAI's retrenchment and Microsoft's aggressive resource locking, highlighting the Arctic as a new AI compute frontier.
AMD Secures 2.5GW Capacity with Core Scientific, Shifts to AI Infrastructure Provider
AMD signs a 15-year agreement with Core Scientific to secure up to 2.5GW of data center capacity, starting with 529MW across five US states. AMD will directly lease 377MW to deploy Instinct GPUs and EPYC CPUs, marking a strategic shift from chip vendor to integrated infrastructure provider, directly competing with NVIDIA's DGX Cloud.
Microsoft Launches MAI Models, Slashes GPU Costs 89%, Reducing OpenAI Dependency
Microsoft unveiled MAI-Image-2.5-Pro and MAI-Voice-2-Flash on Azure Foundry, achieving 96.8% text rendering accuracy at 8K and reducing GPU costs by 84-89% vs GPT. Integrated across Bing, PowerPoint, and Dynamics 365, it marks a strategic shift from OpenAI dependency. Also, NVIDIA Jetson heads to the moon for edge AI.
AMD Helios Enters Production: 12-Stack HBM4 Outmuscles NVIDIA, UALoE Opens AI Network
AMD's second-generation Helios rack-scale AI server enters full production, featuring 72 MI455X GPUs with Samsung's exclusive 12-stack HBM4 (31TB per rack). Compared to NVIDIA's 8-stack design, it offers 50% more memory and 30% lower token cost. Microsoft Azure commits to large-scale deployment, solidifying hyperscaler dual-vendor strategy.
Anthropic partners with SpaceX for 220K GPUs, shifting AI compute ecosystem beyond hyperscalers
Anthropic signs a deal with SpaceX for 300MW capacity and 220,000 NVIDIA GPUs at Colossus 1 data center, with exploration of orbital AI compute. This diversifies Anthropic's compute sources beyond hyperscalers and doubles Claude Code usage limits.
NVIDIA and Wistron Launch US-Based GB300 Production, Shifting AI Manufacturing Landscape
NVIDIA partnered with Wistron to open a $700M factory in Texas, achieving L6 integration of the GB300 Grace Blackwell Ultra system. The first US-made unit contains 1.5M parts, weighs 2 tons, and costs $4M. This milestone accelerates US AI capex localization from 5% to 30%, reshaping the global AI supply chain.
OpenAI Launches Project Camellia: $20B Self-Built AI Data Center, Shift from Cloud Renting to Ownership
OpenAI announces Project Camellia, a $20B self-built AI data center in Georgia with 3.2GW power, marking a shift from cloud renting to self-control. It hires key xAI Colossus team members and raises compute spending forecast to $750B.
AMD Invests $5B in Anthropic, Secures 2GW MI450 Deployment, Reshaping AI Compute Ecosystem
AMD and Anthropic announce a strategic partnership: Anthropic will deploy up to 2GW of AMD Instinct MI450 GPUs, with AMD investing up to $5B in Anthropic. They will collaborate on ROCm optimization and Claude workload tuning, marking AMD's transition from chip vendor to AI ecosystem investor and accelerating multi-sourcing in AI compute.
NVIDIA-OpenAI $100B Partnership: 10GW Vera Rubin AI Factories Reshape Ecosystem
NVIDIA and OpenAI announce a strategic partnership to deploy at least 10GW of NVIDIA systems using the Vera Rubin platform (Rubin GPU, Vera CPU, HBM4, NVLink 6). NVIDIA will invest up to $100B. First facilities go online in H2 2026, powering OpenAI's next-gen models, marking the era of multi-GW AI factories.
Microsoft and Mistral Partner to Build Sovereign AI Infrastructure for Regulated European Industries
Microsoft and Mistral expand their partnership with a multi-billion dollar deal. Mistral gains thousands of NVIDIA Vera Rubin GPUs and integrates its Medium 3.5 and OCR 4 models into Microsoft Foundry and Copilot Studio, offering cloud, connected, and offline deployment modes for European regulated industries under EU AI Act.
Google's Frozen v2 Chip Hardwires Gemini Architecture for 6-10x TPU Efficiency, Set for 2028
Google is developing Frozen v2, a dedicated AI chip that hardwires the Gemini model architecture into silicon for 6-10x energy efficiency per token over current TPUs. It is a new product line, planned for 2028, with weight update flexibility but a frozen architecture. This validates the industry shift from general-purpose GPUs to dedicated ASICs for AI inference.
NVIDIA Invests $2B in CoreWeave, Debuts 'Compute Central Bank' Model
NVIDIA invests $2 billion in CoreWeave and launches the AI Compute Partner Program, featuring credit enhancement, revenue sharing, and GPU buyback. This transforms NVIDIA from a hardware vendor into a 'compute central bank', tightening control over the AI cloud leasing ecosystem and squeezing intermediaries.
Google DeepMind AlphaEvolve GA: AI Self-Evolution for Data Center and Algorithm Optimization
On July 19, 2026, Google DeepMind announced the GA of AlphaEvolve, a Gemini-based multi-agent evolution system for algorithmic discovery, mathematical discovery, and data center efficiency optimization, already used in Borg and Orca, aiming to reduce Capex in massive AI compute investments.
Huawei Ascend 950 Supernode: Self-Developed HCCS Interconnect for Sovereign AI Compute Ecosystem
Huawei unveiled the Ascend 950 Supernode at WAIC, integrating 32 self-developed Ascend 950 AI processors with HCCS high-speed interconnect, achieving 2.5x compute density and supporting trillion-parameter model training, offering a sovereign AI compute alternative free from overseas supply chains.
NVIDIA Vera Rubin Platform and Dynamo 1.0 Disaggregate Inference, Shift Focus to Intelligence per Dollar
NVIDIA unveils Vera Rubin platform with a 7-chip stack (Vera CPU, Rubin GPU, NVLink 6, etc.) and Dynamo 1.0 inference disaggregation. A single NVL72 rack packs 72 GPUs/36 CPUs with 1.6 PB/s bandwidth, achieving up to 7x inference performance. The new 'intelligence per dollar' metric signals a shift from training to inference cost competition.
Meta to lease AI compute to Anthropic, signaling infrastructure monetization push
Meta is in talks to lease AI compute capacity to Anthropic, aiming to monetize its massive infrastructure investment. This marks Meta's shift from internal consumer to external provider, potentially reshaping the AI compute market and intensifying competition with cloud providers.