Reports
AI-generated structured vendor updates
Trend Micro Exposes Azure DNS Design Flaw Enabling Cloud Infrastructure Takeover
Trend Micro's TrendAI™ research team disclosed a security vulnerability "by design" in the Azure cloud platform. DNS records of deleted Azure resources may persist, allowing attackers to exploit these lingering DNS names to hijack trusted endpoints and compromise dependent systems, highlighting a critical but often overlooked trust inheritance risk in cloud infrastructure.
ReflectionAI Secures $6.3B SpaceX Compute Deal, Open-Source AI Breaks Hardware Lock-in
Open-source AI startup ReflectionAI signs a $6.3B deal with SpaceXAI to lease NVIDIA GB300 compute at Colossus 2 for training open-weight frontier models. This gives open-source labs parity with closed-source giants but creates deep dependency on NVIDIA's proprietary hardware.
NVIDIA Acquires Groq LPU: Inference Architecture Shift from HBM to On-Chip SRAM
NVIDIA signs ~$20B licensing deal with Groq for LPU tech, featuring 230MB on-chip SRAM at 80TB/s bandwidth. This targets Transformer inference decode, replacing HBM bottlenecks with ultra-low latency on-chip storage, potentially reshaping the AI inference chip landscape.
Intel Lands Google TPU Package Order: Foundry Pivot Gains Traction, TSMC Still Core
Intel secured a multi-million unit order for Google TPU packaging using its EMIB-T technology, marking its largest external AI chip deal. However, analysts caution the order is primarily for packaging, not wafer fabrication, with TSMC retaining the core manufacturing role.
Microsoft GitHub Leases AWS Capacity: AI Demand Forces Cross-Cloud Collaboration, Shattering Vendor Lock-In
Microsoft's GitHub, facing a 14x surge in AI-driven code commits, is renting compute capacity from rival AWS. This reveals that no single cloud provider can meet AI infrastructure demand, breaking traditional cloud competition and heralding cross-cloud hybrid deployment as the new norm.
Google Gemini 3.5 Flash Turns Search into AI-First Answer Engine, Shifting Control from Links to Summaries
Google transforms Search into an AI-first answer engine powered by Gemini 3.5 Flash, with redesigned search bar, AI-generated summary pages, and proactive monitoring. Model improvements include 1M context, 65K output tokens, and multi-agent orchestration via Antigravity, enabling complex task automation.
Google TurboQuant: 6x KV Cache Compression, AI Inference Memory Cost Inflection Point
Google releases TurboQuant, a two-stage KV cache compression algorithm (PolarQuant + QJL) achieving 6x memory reduction (3-bit quantization) and 8x attention speedup with no measurable accuracy loss. The announcement triggered a sell-off in memory stocks (Micron -3%, Western Digital -4.7%), signaling a potential structural shift in AI inference memory demand.