Filter

1575 Total Reports
1/79 Page
NVIDIA Other 2026-07-22

NVIDIA and Wistron Open US Factory for GB300 and Vera Rubin AI Superchips

Wistron opens its first US manufacturing facility in Fort Worth, producing NVIDIA GB300 Grace Blackwell Ultra and Vera Rubin superchips. The $700M plant aims for tens of thousands of boards monthly, marking NVIDIA's strategic shift to domestic AI hardware production.

NVIDIA Other 2026-07-21

NVIDIA Vera Rubin Platform Specs Revealed: 10x Tokens per Watt, Monolithic CPU+GPU Design

NVIDIA unveiled Vera Rubin platform specs with a monolithic design pairing 2 Rubin GPUs with 1 Vera CPU, flagship NVL72 integrating 36 CPUs and 72 GPUs. Claims 10x tokens per watt and 3x memory bandwidth over Grace Blackwell. Vera CPU sold standalone. First customers: Microsoft, OpenAI, Oracle. Mass production H2 2026. Performance claims await independent verification.

Microsoft Other 2026-07-21

Microsoft Invests Billions in Mistral AI, Integrates Models into Azure for Sovereign AI

Microsoft and Mistral AI announce a multi-billion dollar partnership, with Microsoft investing in Mistral's European data center capacity and integrating Mistral's Medium 3.5 and OCR 4 models into Azure Foundry. The deal directly responds to US export controls on Anthropic, offering regulated European industries a sovereign AI alternative, signaling a shift from centralized US AI to localized infrastructure.

NVIDIA Other 2026-07-21

NVIDIA Spectrum-6 102.4Tbps Switch Goes Commercial, Cisco Adoption Confirms Bandwidth Inflection

NVIDIA announces Spectrum-6 102.4Tbps Ethernet switch for AI factories, doubling bandwidth with CPO and liquid cooling. Cisco confirms adoption in N9100 series, while Broadcom launches Tomahawk 6, signaling a terabit Ethernet race for AI infrastructure.

Cloudflare Other 2026-07-21

Cloudflare Precursor GA: Continuous Behavior Verification Replaces Static CAPTCHA for Bot Defense

Cloudflare announces GA of Precursor, a continuous behavior verification engine on its global edge network. It analyzes entire user sessions (mouse movements, typing rhythms) to detect sophisticated bots, replacing static CAPTCHA. Zero-code deployment, privacy-first, it protects billions of interactions daily, addressing the new norm where bots account for 57% of web traffic.

OpenAI Other 2026-07-21

OpenAI GPT-5.6: Three-Layer Routing Shifts Control, Multi-Agent Parallelism Locks Workflows

OpenAI launches GPT-5.6 with Soul/Terra/Luna three-layer model routing, enabling automatic model selection and tool orchestration. New ChatGPT Work, Ultra Mode multi-agent parallelism, and a four-layer security framework shift AI from Q&A to autonomous task execution, consolidating OpenAI's control over AI workflow orchestration.

NVIDIA Other 2026-07-21

NVIDIA Open-Sources Cosmos 3 Edge 4B World Model for Real-Time Robot Control at 15Hz on Jetson Thor

NVIDIA open-sources Cosmos 3 Edge, a 4B parameter world action model for edge robotics. It achieves 15Hz real-time inference with 32 actions per inference on the Jetson Thor module. This extends NVIDIA's physical AI stack from training to real-time deployment, enabling end-to-end robot control at the edge.

Microsoft Other 2026-07-21

Microsoft's Project Perception Automates Vulnerability Remediation with Multi-Model AI Orchestration

In response to competitors like Anthropic and Palo Alto, Microsoft's Project Perception leverages multi-model AI orchestration to automate vulnerability discovery and remediation. This product signifies a transition from manual security operations to autonomous self-healing systems, with control shifting from security analysts to an AI-driven platform.

NVIDIA Other 2026-07-20

NVIDIA Expands Agent Toolkit with Omniverse Libraries for Physical AI Simulation

At SIGGRAPH 2026, NVIDIA announced an expansion to its Agent Toolkit, adding Omniverse libraries that enable AI agents to build and simulate 3D worlds. The company also open-sourced Cosmos 3 Edge, a 4B-parameter world action model, completing its physical AI ecosystem from training to edge deployment.

Apple Other 2026-07-20

Apple Engages PrismML for 1-bit Quantization, Enabling 15x Memory Reduction for On-Device AI

Apple is evaluating PrismML's native 1-bit model compression technology, reducing model size to 1/14, memory usage by 90%, and boosting inference speed by 8x. The Bonsai 27B model can run on iPhone 15, marking a breakthrough in on-device AI that could reshape the mobile AI landscape.

NVIDIA Other 2026-07-20

NVIDIA Vera Rubin at BMS: Mission Control and BioNeMo Shift the AI Factory Control Plane

BMS deploys NVIDIA DGX SuperPOD with Vera CPU and Rubin GPU, managed by Mission Control and leveraging BioNeMo Agent Toolkit for drug discovery. This signals NVIDIA's shift from hardware vendor to AI factory control plane provider, potentially locking enterprises into its ecosystem.

Google Other 2026-07-20

Google's Frozen v2 Chip Hardwires Gemini Architecture for 6-10x TPU Efficiency, Set for 2028

Google is developing Frozen v2, a dedicated AI chip that hardwires the Gemini model architecture into silicon for 6-10x energy efficiency per token over current TPUs. It is a new product line, planned for 2028, with weight update flexibility but a frozen architecture. This validates the industry shift from general-purpose GPUs to dedicated ASICs for AI inference.

AMD Other 2026-07-20

Microsoft Azure Deploys AMD Helios Rack with MI455X GPUs, Breaking NVIDIA's Cloud AI Monopoly

Microsoft Azure officially adopts AMD Helios rack-scale AI infrastructure, featuring 72 MI455X GPUs (432GB HBM4, 19.6TB/s), Venice EPYC CPUs, and Pensando DPUs. Three new instances (ND MI455X v7, HDv2, HXv2) are launched, marking Azure's shift from exclusive NVIDIA dependency to a multi-vendor AI strategy.

NVIDIA Other 2026-07-20

NVIDIA Invests $2B in CoreWeave, Debuts 'Compute Central Bank' Model

NVIDIA invests $2 billion in CoreWeave and launches the AI Compute Partner Program, featuring credit enhancement, revenue sharing, and GPU buyback. This transforms NVIDIA from a hardware vendor into a 'compute central bank', tightening control over the AI cloud leasing ecosystem and squeezing intermediaries.

CrowdStrike Other 2026-07-20

Beyond the Model: Harnessing Frontier AI for Stronger Defense

...

Intel Other 2026-07-20

Intel Foundry 18A Yields Jump to 85%+; EMIB Packaging Hits 98%, Challenging TSMC N2

Intel Foundry 18A yields surged from 65% to 85%+ in a single quarter, approaching TSMC N2's 90%. EMIB advanced packaging yields reached 90-98%, turning a former bottleneck into a selling point. NVIDIA, AMD, Apple signed on but mostly as secondary suppliers.

Meta Other 2026-07-20

Meta Launches Muse Spark 1.1 API at 25% Competitor Price, Ends Open-Source Era

Meta releases Muse Spark 1.1, a multimodal reasoning model with 1M token context window, and launches its first paid API at 25% of competitors' price. This ends the Llama open-source era, signaling a strategic shift to proprietary API monetization and aggressive market share capture.

Anthropic Other 2026-07-20

Anthropic Extends Claude Cowork Unified Interface to Web and Mobile, Compliance Gap Looms

Anthropic launches Claude Cowork unified interface on Web and Mobile for Max users, merging chat and task execution with local file access and cross-device continuity. However, Cowork activities are not captured in audit logs or Compliance API, creating a significant governance gap vs. Microsoft Copilot.

Microsoft Other 2026-07-20

Microsoft July Patch Tuesday Hits Record 622 CVEs, AI Infrastructure Vulnerabilities Emerge as New Attack Surface

Microsoft's July 2026 Patch Tuesday addresses a record 622 CVEs, including three critical AI vulnerabilities: Copilot RCE (CVSS 9.6), Azure OpenAI EoP (CVSS 9.9), and M365 Copilot EoP (CVSS 9.3). The attack surface expands from OS to AI service infrastructure, signaling an era of AI-driven vulnerability inflation.

Google Cloud Other 2026-07-20

Google DeepMind AlphaEvolve GA: AI Self-Evolution for Data Center and Algorithm Optimization

On July 19, 2026, Google DeepMind announced the GA of AlphaEvolve, a Gemini-based multi-agent evolution system for algorithmic discovery, mathematical discovery, and data center efficiency optimization, already used in Borg and Orca, aiming to reduce Capex in massive AI compute investments.