Filter

×
Active Filters Clear All
Keyword: AI ×
1487 Total Reports
9/75 Page
Huawei Other 2026-07-17

Huawei Atlas 950 SuperPoD & 灵衢2.0: A Systemic Pivot in China's AI Compute from Chip to Cluster

At WAIC 2026, Huawei publicly demonstrated the Atlas 950 SuperPoD, a 1024-ascend NPU card cluster, and unveiled the 灵衢2.0 high-speed interconnect protocol. This signals a strategic shift in China's AI infrastructure from single-chip to system-level leadership, creating a closed-loop ecosystem that directly challenges NVIDIA's NVL series dominance.

TSMC Other 2026-07-17

TSMC Pledges $100B More for 6 US Fabs, Localizing 3nm for AI Chip Supply Chain

TSMC announces an additional $100B investment in Arizona, bringing total US commitment to $265B, with plans for 6 fabs focused on 3nm and beyond. This move localizes advanced process for AI chip demand from NVIDIA, Apple, AMD, reshaping global semiconductor supply chain. Q2 net profit surged 77% YoY, FY capex raised to $60-64B.

Huawei Other 2026-07-17

Huawei Ascend 950 SuperPoD: 1024 NPUs with 256TB Unified Memory Redefines AI Compute

Huawei unveiled the Ascend 950 SuperPoD at WAIC 2026, featuring 1024 NPUs per cabinet with 256TB unified memory and 3μs latency. This system-level innovation compensates for process limitations, scaling to 500,000 NPUs for trillion-parameter model training, shifting the compute race from single-chip to system efficiency.

Microsoft Azure Other 2026-07-16

Microsoft Azure Cuts 200-400 Jobs in China Amid Cloud Growth, Reshapes Geopolitical Compliance

Microsoft Azure is cutting 200-400 jobs in China even as its cloud business grows 40% YoY, offering some employees relocation to Canada. This signals a strategic shift to move compliance and operational control out of China, reshaping enterprise multi-cloud and data sovereignty decisions.

Cloudflare Other 2026-07-16

Cloudflare Launches AI Payment Gateway, Revives HTTP 402 for Bot Monetization

Cloudflare introduces an AI content payment gateway leveraging HTTP 402 and stablecoin settlements to enforce payments at the edge, enabling websites to charge AI crawlers for access. This system aims to monetize bot traffic and end free scraping.

NVIDIA Other 2026-07-16

Huang Denies Vera Rubin Delay; NVIDIA Defends AI Compute Throne

Jensen Huang officially denies rumors of a delay for the Vera Rubin platform, stating it is already in production and on track for mass deployment. This move aims to quell market anxiety over NVIDIA's product roadmap and solidify its leadership in AI training and inference chips.

AMD Other 2026-07-16

AMD与OpenAI达成6GW算力供应历史性协议 1.6亿认股权证可获10%股权 股价盘前涨35%

...

NVIDIA Other 2026-07-16

NVIDIA Debuts T3000/T2000 Modules and Cosmos 3 Edge, Builds Sovereign AI Ecosystem in Japan

NVIDIA unveils T3000/T2000 compute modules (Thor architecture) and Cosmos 3 Edge world model, signs Japan Noetra alliance for 13,750 Vera CPUs + 27,500 Rubin GPUs (140MW). Sovereign AI revenue triples to $30B+ in FY2026, accelerating the physical AI ecosystem.

CrowdStrike Other 2026-07-16

CrowdStrike Buys XM Cyber IP: Attack Path Management to Power Predictive Security on Falcon

CrowdStrike acquires all IP assets of XM Cyber (45+ patents and source code) while leaving XM Cyber as an independent licensee. This integrates attack path management into the Falcon platform, shifting from reactive EDR to predictive security, and strengthens partnership with European retail giant Schwarz Digits for market expansion.

Microsoft Other 2026-07-16

Microsoft Replaces OpenAI/Anthropic with In-House MAI Models to Cut Costs and Reduce Dependency

Microsoft has started replacing OpenAI and Anthropic AI calls in Excel and Outlook with its in-house MAI models, handling tens of thousands of prompts weekly. The move aims to cut costs and reduce dependency on Anthropic, signaling a strategic shift toward internal AI models and impacting the AI vendor ecosystem.

NVIDIA Other 2026-07-16

NVIDIA-Nokia Alliance Redefines RAN Ecosystem with GPU-Based AI Acceleration

NVIDIA and Nokia are jointly developing AI-powered RAN technology, using NVIDIA GPUs to accelerate baseband processing and AI algorithms for beamforming and spectrum optimization. Targeting commercial deployment by 2027 and 2x spectral efficiency by 2028, this partnership marks a fundamental shift from dedicated RAN hardware to GPU-based, software-defined AI networks.

Other Other 2026-07-16

IBM Unveils 0.7nm Nanostack 3D Transistor, Doubling Density and Extending Moore’s Law

IBM debuts the world’s first sub-1nm chip technology at 0.7nm node using a nanostack 3D transistor architecture, packing nearly 100 billion transistors with double the density of its 2nm node. It delivers 50% performance gain or 70% energy efficiency improvement, and 40% SRAM scaling for AI workloads.

NVIDIA Other 2026-07-16

测试情报-NVIDIA AI chip news

...

NVIDIA Other 2026-07-16

NVIDIA Jetson Thor T3000/T2000: Blackwell GPU Crashes Edge AI Cost Barrier

NVIDIA unveils Jetson Thor T3000 and T2000 modules. The T3000 packs a Blackwell GPU and 8-core Neoverse CPU, delivering 865 FP4 TFLOPS at half the power of the T5000. New Jetson Agent Skills automate memory optimization, aiming to scale deployment of humanoid robots and edge AI.

NVIDIA Other 2026-07-16

NVIDIA CUDA 13.3 Introduces clmad for Hardware-Accelerated Carryless Multiplication on GPUs

NVIDIA CUDA 13.3 adds the clmad hardware instruction for carryless multiply-accumulate on Ampere+ GPUs. GHASH throughput reaches 6.3 TB/s on B200, up to 18.8x faster than bitsliced. Sum-check protocol accelerates 3-13x. The instruction also benefits CRC, Reed-Solomon, and post-quantum cryptography.

Qualcomm Other 2026-07-15

Qualcomm Negotiates Custom Chips with ByteDance, Shifts to Data Center Ecosystem

Qualcomm is in talks with ByteDance to develop custom chips, including VPU, AI components, and CPUs, leveraging AlphaWave Semi's interconnect tech. This marks Qualcomm's strategic shift from smartphones to data center custom silicon, with its Dragonfly portfolio featuring C1000 CPU, HBC, and AI300 accelerators.

Intel Other 2026-07-15

Intel 18A Yield Hits 85%, Secures Orders from NVIDIA, OpenAI, Reshaping Foundry Landscape

Intel reports 18A process yield improvement to 85%, from 65% last quarter, nearing TSMC N2's 90%. Secured foundry deals with NVIDIA, AMD, OpenAI, etc. EMIB advanced packaging yield reaches 98%, used in NVIDIA Feynman, Google TPU. This marks a strategic inflection in AI chip manufacturing.

NVIDIA Other 2026-07-15

NVIDIA and Nokia Launch Commercial AI-RAN: GPU-Defined RAN Replaces Purpose-Built Hardware

NVIDIA and Nokia announce the first commercial AI-RAN platform, built on Nokia's anyRAN software and NVIDIA's Aerial AI-RAN stack. It achieves over 20% spectrum efficiency gain via AI-driven radio innovations, targeting 100%+ by 2028. The platform aims to shift RAN from purpose-built hardware to a software-defined, GPU-based compute model.

Google Other 2026-07-15

Google Deeply Integrates Gemini Enterprise Telemetry with BigQuery for AI Governance

Google Cloud enables streaming Gemini Enterprise app telemetry (prompts, responses, activity logs) into BigQuery for real-time analysis. Leveraging BigQuery's AI capabilities (Conversational Analytics, auto-schema), it automates auditing, compliance, and insights for large-scale AI deployments, driving data-driven AI observability.

Apple Other 2026-07-15

Apple in Talks with PrismML to Compress Qwen 27B Model 15x for On-Device AI

Apple is negotiating with AI startup PrismML to deploy a compressed version of Alibaba's Qwen 27B parameter model on iPhone. PrismML's compression technology reduces memory usage by 15x, enabling 27B models to run locally with 10GB VRAM, shifting Apple's AI strategy from cloud-dependent to on-device inference.