Reports
AI-generated structured vendor updates
ARM Builds Its First Chip in 35 Years: AGI CPU Targets AI Data Centers, Meta First Customer
ARM announces its first in-house CPU in 35 years, the AGI CPU, targeting AI and data center workloads. Meta is the launch customer. Built on TSMC's 3nm process, the chip focuses on performance-per-watt, directly challenging x86 dominance and fundamentally restructuring ARM's business model from IP licensor to merchant silicon vendor.
Meta Partners with Arm to Develop New AI Data Center CPUs
Meta partners with Arm to co-develop data center CPUs optimized for AI workloads. The first product, the Arm AGI CPU, aims to boost rack performance density for large-scale AI deployments. It will be available through Arm's ecosystem, with board designs to be open-sourced via the Open Compute Project.
ARM Launches AGI CPU for Agentic AI Infrastructure Era
ARM introduces the Arm AGI CPU, its first silicon product, designed for agentic AI infrastructure on Neoverse. Optimized for massively parallel workloads, it supports 272 cores per blade in a 1OU design, delivering 8160 cores per rack and over 2x performance vs. x86 systems.
ARM Launches AGI CPU Silicon for AI Infrastructure Market
ARM introduced its first production AGI CPU silicon in March 2026, marking a strategic shift from IP licensing to full silicon solutions provider. Designed for next-gen AI infrastructure, this move may reshape the data center processor ecosystem.
Arm Neoverse Reshapes Control Layer in AI Infrastructure
ARM introduces Neoverse infrastructure CPU cores optimized for cloud, AI, and HPC workloads, adopted by NVIDIA, AWS, Microsoft, and Google for their AI platforms, delivering performance gains and energy efficiency. This architecture enables high-density AI workload deployment in cloud and edge environments with enhanced multi-tenant security.
HPE Enhances AI Security Architecture for Adoption Risks
HPE introduces SRX400 Series Firewalls, expanded hybrid mesh security, and AI governance capabilities to secure AI adoption. Features include AI app visibility, prompt-level inspection, and identity-based protection to mitigate data exposure risks.
NVIDIA Donates GPU Dynamic Resource Allocation Driver to Kubernetes Community
NVIDIA donated its GPU Dynamic Resource Allocation (DRA) driver to the CNCF, making it an upstream Kubernetes project. This move aims to shift the core control point of GPU orchestration from proprietary vendor layers to the open-source community, and drive standardization in collaboration with major cloud providers.
NVIDIA IGX Thor: 8x Edge AI Compute with ConnectX-7 Network Lock-In
NVIDIA launches IGX Thor edge AI platform with Blackwell GPU, up to 5,581 FP4 TFLOPS, dual 200GbE RDMA via ConnectX-7, and ISO 26262 safety. Pin-compatible with Jetson Thor and 10-year lifecycle enable seamless migration, but create vendor lock-in through proprietary networking and GPU dependencies.
ARM and NVIDIA Drive Localization Revolution in AI Workstations
ARM and NVIDIA jointly launch DGX Spark AI workstations based on GB10 Grace Blackwell chips, with eight major OEMs releasing products simultaneously. The solution features unified memory architecture supporting 200B parameter models locally, with third-party tests showing 41% faster rendering and 3.2x AI processing speed versus x86 alternatives, enabling seamless cloud-to-edge toolchain migration.
NVIDIA Launches OpenShell, Establishing Runtime Sandbox for Secure Autonomous AI Agents
NVIDIA introduces OpenShell, an open-source project designed as a secure-by-design runtime for autonomous AI agents. It employs a "browser tab" model, isolating agent operations from policy enforcement at the system level to prevent policy overrides and data leaks. NVIDIA is collaborating with key security vendors to establish a unified policy layer for enterprise AI agents.
Cisco Extends Zero Trust Security to AI Agent Ecosystem
At RSA 2026, Cisco introduced security innovations for AI agents, extending Zero Trust Access with agent discovery in Identity Intelligence, agentic IAM in Duo, and MCP enforcement in Secure Access SSE. It launched AI Defense: Explorer Edition for self-serve testing and DefenseClaw open source framework to automate security deployment.
SK Hynix Jumps to TSMC 3nm for HBM4E Logic Die to Counter Samsung's 4nm Lead
SK Hynix plans to use TSMC's 3nm process for the logic die in its 7th-gen HBM4E, a leap from the 12nm used in HBM4. This aims to reverse the performance gap with Samsung (which used 4nm logic in HBM4) and deliver higher bandwidth and power efficiency for next-gen AI chips like NVIDIA's Vera Rubin Ultra.
Meta Integrates AI Support Assistant with Content Moderation, Reducing Third-Party Reliance
Meta launched an AI support assistant and deployed advanced AI content moderation systems to enhance user experience and platform safety. This signals a strategic shift from relying on third-party vendors to strengthening internal AI systems, with plans to deeply integrate AI into core operations.
Google Data Center Demand Response Signs 1GW, Building Grid Flexibility
Google integrates multiple utilities through long-term energy contracts to achieve 1GW data center demand response capability. The technology regulates energy consumption by limiting or shifting ML workloads to balance grid supply and demand. This transforms data centers from power consumers to grid flexibility assets.
AMD and NAVER Cloud Collaborate on Sovereign AI Infrastructure in Korea
AMD and NAVER Cloud announced a strategic collaboration to accelerate sovereign AI infrastructure in Korea. NAVER Cloud will expand deployment of AMD EPYC "Venice" CPUs and gain early access to next-gen Instinct MI455X GPUs, with joint optimization of AI services and software stacks on AMD platforms.
AMD and Samsung Deepen Collaboration, Locking HBM4 Supply and Exploring Foundry Partnership
AMD and Samsung signed an MOU, designating Samsung as the primary HBM4 supplier for the next-gen Instinct MI455X GPU and collaborating on DDR5 memory optimized for 6th Gen EPYC CPUs. The companies will also explore opportunities for Samsung to provide foundry services for future AMD products.
NVIDIA RTX Workstations Directly Connect to Apple Vision Pro for Enterprise XR
NVIDIA's CloudXR SDK 6.0 enables native direct connection between RTX-accelerated workstations and Apple Vision Pro, eliminating traditional streaming servers. Integrated with Omniverse and OpenUSD workflows, it reduces deployment complexity for enterprise XR applications.
NVIDIA and Telecom Operators Build AI Grids to Redistribute AI Inference
NVIDIA is partnering with global telecom operators like AT&T and Comcast to transform existing distributed network sites into 'AI Grids' for edge AI inference. This initiative aims to deploy AI compute closer to users and data, reducing latency and cost per token. It represents a strategic shift for telcos from being data carriers to distributed AI computing platforms.
HPE Launches AI Grid with NVIDIA to Unify Distributed Inference Clusters
HPE announced the AI Grid at NVIDIA GTC, an end-to-end solution built on NVIDIA's reference architecture to securely connect distributed AI factories and inference clusters into a single intelligent system. It enables service providers to deploy and operate thousands of edge inference sites, meeting the predictable, low-latency infrastructure requirements of AI-native applications.
HPE Report Shows Attackers' AI-Driven Business Models
HPE Threat Labs report reveals cyber adversaries adopting business-like operations with automation and generative AI to scale attacks. Based on 2025 global threat analysis, it underscores the need for AI-integrated defenses and zero trust.