Reports
AI-generated structured vendor updates
Cisco and NVIDIA Extend Secure AI Factory with Network-Security Integration
Cisco and NVIDIA deepen collaboration on Secure AI Factory, extending AI deployment from core to edge. Launch high-performance switches with NVIDIA Spectrum and expand security enforcement to DPU level with AI guardrails integration.
NVIDIA Releases AI Factory Reference Design and Digital Twin Blueprint
NVIDIA unveiled Vera Rubin DSX AI factory reference design and Omniverse DSX digital twin blueprint, built on Spectrum-X Ethernet, Quantum-X800 InfiniBand and BlueField-3 DPU. The architecture connects real-world sensors with digital twins for continuous AI model training and optimization, extending AI computing from data centers to physical world automation.
Intel Xeon 6 Selected as Host CPU for NVIDIA DGX Rubin, Enhancing AI Inference Infrastructure
Intel Xeon 6 is chosen as host CPU for NVIDIA DGX Rubin NVL8 AI system, delivering 3x memory bandwidth and full-path confidential computing. This collaboration highlights CPU's architectural role in data orchestration and security for AI inference workloads.
Samsung Unveils HBM4E and Hybrid Copper Bonding for AI Infrastructure
Samsung announced HBM4 mass production and showcased next-gen HBM4E with 4TB/s bandwidth at GTC 2026. Hybrid copper bonding enables 16+ layers with 20% lower thermal resistance. Also launched SOCAMM2 memory and PCIe 6.0 SSD for NVIDIA AI infrastructure.
HPE Deploys Sovereign AI Factories with NVIDIA at National Labs
HPE announces a collaboration with NVIDIA to deploy liquid-cooled sovereign AI systems at Argonne National Laboratory in the U.S. and HLRS in Germany. This move aims to provide government and research institutions with AI infrastructure that meets data sovereignty and compliance requirements, accelerating the deployment and scaling of their AI initiatives.
Cisco Accelerates AI Data Center Deployment with Certified Refurbished Equipment
Cisco introduces a certified refurbished equipment program, offering rigorously tested hardware with full warranty and performance matching new products to accelerate AI-ready data center deployment. The solution reduces deployment time by up to 80% while optimizing capital efficiency and promoting sustainability.
NVIDIA and Thinking Machines Lab Form Gigawatt-Scale AI Infrastructure Partnership
NVIDIA and Thinking Machines Lab announced deployment of at least one gigawatt of next-gen Vera Rubin systems for cutting-edge AI model training. This collaboration sets a new benchmark for hyperscale AI compute demand, signaling a move towards gigawatt-scale AI infrastructure.
Meta Accelerates Custom AI Chip Roadmap with Focus on Inference Optimization
Meta plans to launch four generations of MTIA AI chips in two years, adopting an 'inference-first' design strategy optimized for generative AI tasks. Built on PyTorch and open standards, the chips enable seamless data center deployment, targeting improved compute efficiency and cost control.
NVIDIA Partners with Thinking Machines Lab for Gigawatt-Scale AI Infrastructure
NVIDIA and Thinking Machines Lab form a multi-year partnership to deploy at least 1 GW of next-gen Vera Rubin systems for cutting-edge AI model training and scalable customized AI platforms. The collaboration includes co-designing training and inference systems and expanding access to advanced AI and open-source models for enterprises and research institutions.
NVIDIA Partners with Coherent on Data Center Optical Interconnect Tech
NVIDIA partners with photonics specialist Coherent to develop next-gen data center optical interconnect technology. The collaboration targets high-performance, high-density, low-power optical solutions for AI and HPC workloads, addressing bandwidth and efficiency bottlenecks. This strengthens NVIDIA's system-level optimization in AI infrastructure hardware ecosystems.
NVIDIA and Coherent Collaborate on Data Center Optical Interconnect Technology
NVIDIA and optical technology provider Coherent have formed a strategic partnership to develop next-generation data center optical interconnect solutions. The collaboration combines NVIDIA's AI computing expertise with Coherent's photonics technology to deliver higher bandwidth and lower latency interconnects for AI clusters and HPC.
Palo Alto Networks Advocates Service Provider Shift to Secure AI Factory
Palo Alto Networks proposes service providers transform into 'secure AI factories' by building integrated platforms for AI development, deployment, governance, and security. The platform emphasizes embedded security layers for proactive protection against model poisoning and data leaks, repositioning security from cost to business enabler.
Samsung and NVIDIA Complete Multi-Cell AI-RAN Test with Chip-Level Integration
Samsung validated integrated vRAN software with NVIDIA's accelerated computing platform in real network environments, demonstrating AI algorithm optimization for wireless physical layer performance. The collaboration extends to chip-level architecture using unified processors to optimize CPU-GPU connectivity for spectrum efficiency.
AMD Partners with TCS to Deploy Helios AI Rack Architecture in India
AMD partners with Tata Consultancy Services to introduce the Helios rack-scale AI architecture in India, built on Instinct MI300 accelerators for large-scale AI training and inference workloads. The solution is delivered as complete racks, scalable to thousands of nodes, optimized for generative AI and HPC. The collaboration leverages TCS's integration services in cloud, AI, and cybersecurity for end-to-end AI solutions.
AWS Launches Inferentia2 Chip for Generative AI Infrastructure Optimization
AWS launched second-gen Inferentia2 AI inference chip, designed for Transformer models with 4x performance boost and support for 175B parameter models. Integrated into EC2 Inf2 instances with UltraClusters architecture for large-scale deployment, offering 40% better cost-performance and 50% lower power consumption than GPU instances.
Microsoft partners with Starlink to advance AI-ready community digital access strategy
Microsoft partners with Starlink to provide digital access tools for rural areas via low-earth orbit satellite connectivity. The company evolves its digital access strategy from coverage to adoption and empowerment, building systematic solutions including reliable energy, affordable devices and AI tools. This move aims to support global AI economy development and construct AI-ready community infrastructure.
Apple Expands US AI Server Manufacturing and Mac Production Lines
Apple relocates Mac mini production to Houston for the first time and expands AI server manufacturing, producing core components like logic boards locally. It also invests in an advanced manufacturing center to provide technical training for US manufacturing skills enhancement.
Cisco Expands AI Security Architecture and Launches Partner Incentive Program
Cisco launched new solutions for AI agent security, expanding AI Defense to protect AI application supply chain and model integrity, and introducing SASE for Agentic AI with automated detection and access control. The company also added AgenticOps autonomous remediation in Security Cloud Control and enhanced identity security with Duo for Active Directory.
Meta and AMD Form 6GW AI Infrastructure Strategic Partnership
Meta announced a multi-year strategic partnership with AMD to deploy up to 6GW of AMD Instinct GPU computing capacity. The collaboration involves multi-generational integration of AMD GPUs, EPYC CPUs, and jointly developed Helios rack architecture, supporting Meta's diversified computing strategy. First deployments are scheduled for late 2026.
Intel Partners with SambaNova to Expand AI Inference Infrastructure
Intel announces multi-year strategic partnership with SambaNova to develop AI inference solutions based on Xeon processor infrastructure. The collaboration integrates Intel's compute, networking, storage hardware with SambaNova's AI platform, offering rack-scale inference options for heterogeneous data centers. Intel confirms this doesn't alter its independent GPU roadmap and will continue investing in edge-to-cloud AI products.