Filter

×
Active Filters Clear All
Keyword: 中国 ×
60 Total Reports
2/3 Page
NVIDIA Other 2026-07-21

NVIDIA Vera Rubin Platform Specs Revealed: 10x Tokens per Watt, Monolithic CPU+GPU Design

NVIDIA unveiled Vera Rubin platform specs with a monolithic design pairing 2 Rubin GPUs with 1 Vera CPU, flagship NVL72 integrating 36 CPUs and 72 GPUs. Claims 10x tokens per watt and 3x memory bandwidth over Grace Blackwell. Vera CPU sold standalone. First customers: Microsoft, OpenAI, Oracle. Mass production H2 2026. Performance claims await independent verification.

NVIDIA Other 2026-07-21

NVIDIA Open-Sources Cosmos 3 Edge 4B World Model for Real-Time Robot Control at 15Hz on Jetson Thor

NVIDIA open-sources Cosmos 3 Edge, a 4B parameter world action model for edge robotics. It achieves 15Hz real-time inference with 32 actions per inference on the Jetson Thor module. This extends NVIDIA's physical AI stack from training to real-time deployment, enabling end-to-end robot control at the edge.

Other Other 2026-07-20

Moonshot AI Launches Kimi K3: 2.8T Parameter Open-Source MoE Model at $3/$15

Moonshot AI unveils Kimi K3, a 2.8T parameter open-source MoE model with 896 experts (16 active), native vision, and 1M context. API pricing at $3/$15 input/output per million tokens undercuts rivals. Open weights release July 27. GPU capacity exhausted within 2 days. Arena score 1679 tops Fable 5.

Other Other 2026-07-19

Alibaba Launches 2.4T Parameter Qwen3.8-Max MoE Model with 0.2x Pricing

Alibaba released Qwen3.8-Max-Preview, a 2.4 trillion parameter MoE multimodal model with 1M context window. It launched Qoder platform and Token Plan with aggressive discounts up to 0.2x, significantly reducing inference cost. The company claims it is second only to Anthropic Fable 5, marking China's AI entry into dual-track of parameter arms race and open-source competition.

Other Other 2026-07-18

PPIO Launches Agentic Cloud, Intelligent Model Gateway Becomes New Control Point

PPIO unveiled Agentic Cloud and Intelligent Model Gateway at WAIC 2026, targeting AI agent workloads with semantic routing and cost-aware scheduling. With over 1.2 trillion daily tokens and sub-200ms sandbox cold start, it signals the emergence of dedicated agent infrastructure.

Huawei Other 2026-07-17

Huawei Atlas 950 SuperPoD & 灵衢2.0: A Systemic Pivot in China's AI Compute from Chip to Cluster

At WAIC 2026, Huawei publicly demonstrated the Atlas 950 SuperPoD, a 1024-ascend NPU card cluster, and unveiled the 灵衢2.0 high-speed interconnect protocol. This signals a strategic shift in China's AI infrastructure from single-chip to system-level leadership, creating a closed-loop ecosystem that directly challenges NVIDIA's NVL series dominance.

Huawei Other 2026-07-17

Huawei Ascend 950 SuperPoD: 1024 NPUs with 256TB Unified Memory Redefines AI Compute

Huawei unveiled the Ascend 950 SuperPoD at WAIC 2026, featuring 1024 NPUs per cabinet with 256TB unified memory and 3μs latency. This system-level innovation compensates for process limitations, scaling to 500,000 NPUs for trillion-parameter model training, shifting the compute race from single-chip to system efficiency.

Microsoft Other 2026-07-16

Microsoft Replaces OpenAI/Anthropic with In-House MAI Models to Cut Costs and Reduce Dependency

Microsoft has started replacing OpenAI and Anthropic AI calls in Excel and Outlook with its in-house MAI models, handling tens of thousands of prompts weekly. The move aims to cut costs and reduce dependency on Anthropic, signaling a strategic shift toward internal AI models and impacting the AI vendor ecosystem.

Qualcomm Other 2026-07-15

Qualcomm Negotiates Custom Chips with ByteDance, Shifts to Data Center Ecosystem

Qualcomm is in talks with ByteDance to develop custom chips, including VPU, AI components, and CPUs, leveraging AlphaWave Semi's interconnect tech. This marks Qualcomm's strategic shift from smartphones to data center custom silicon, with its Dragonfly portfolio featuring C1000 CPU, HBC, and AI300 accelerators.

Other Other 2026-07-15

New York Enacts First Statewide AI Data Center Moratorium, Signaling Regulatory Paradigm Shift

New York State has signed an executive order imposing a one-year moratorium on AI hyperscale data centers over 50MW, effective immediately. This first statewide ban in the U.S., with 11+ states considering similar laws, signals a regulatory paradigm shift from 'build fast' to 'build steady' in AI infrastructure.

Cisco Other 2026-07-15

Cisco, G42, AMD Deploy 1GW AI Cluster in UAE, Pushing GPU Diversification and Full-Stack Integration

Cisco, G42, and AMD partner to deploy a large-scale AI cluster in the UAE based on AMD MI350X GPUs, integrating Cisco's full-stack secure AI infrastructure (UCS servers, Nexus 9K switches, Firepower firewalls). This marks Cisco's transformation into a full-stack AI infrastructure integrator and positions AMD as a second GPU supplier for US-allied nations, locking out Chinese vendors in the UAE market.

NVIDIA Other 2026-07-14

NVIDIA Halves Asian AI Chip Customers, Whitelist Regime Reshapes Supply Chain

NVIDIA slashes its authorized AI chip customer list in Asia by more than half, establishing a whitelist regime in Singapore, Malaysia, and Japan. Customers must submit detailed business proofs and end-use declarations. This move, aimed at preventing illegal diversions to China, will reshape the global AI chip supply chain and force enterprises to reassess procurement strategies.

Other Other 2026-07-13

JCET's $1.4B AI Packaging Capex Reshapes Advanced Packaging Supply Chain

JCET announces $1.4B capex for AI advanced packaging in 2026, targeting Chiplet, HBM, 2.5D/3D. This marks Chinese OSAT's major entry into high-end packaging, complementing TSMC's CoWoS expansion and shifting AI packaging from Taiwan-centric to cross-strait division, supporting local hyperscalers' computing needs.

Huawei Other 2026-07-10

Huawei Ascend 10K-Card Cluster Goes Live, UnifiedBus Protocol Pools All Resources

Huawei launched an Ascend 10,000-card AI cluster in Shaoguan, Guangdong, and showcased the Atlas 950 SuperPoD with its proprietary UnifiedBus interconnect supporting 8,192 NPUs at 16.3 PB/s. Huawei Cloud also entered the Gartner 2026 Cloud AI Infrastructure Leaders quadrant, reinforcing its push for a self-contained AI ecosystem.

OpenAI Other 2026-07-09

OpenAI Reopens with GPT-oss Models: Apache 2.0 License Hides Cloud Offload Control

OpenAI launches GPT-oss-120b and GPT-oss-20b under Apache 2.0 license, capable of running on a single 80GB GPU. However, a built-in cloud offload mechanism routes complex queries to proprietary models, masking a strategic control point shift behind the open-source facade.

Anthropic Other 2026-07-07

Anthropic Admits Covert Tagging in Claude Code: Full-Chain Account Locking Signals Control Shift

Anthropic confirms a covert account tagging system in Claude Code to detect and block unauthorized resale and model distillation, causing full-chain lockout for users in certain regions. This defensive move against IP leakage shifts access control from users to the vendor, risking collateral damage on legitimate customers.

Microsoft Other 2026-07-07

AI Giants Bet $10B on Forward Deployed Engineers: Control Shifts from Models to Engineering

Microsoft, OpenAI, Anthropic, and AWS collectively announced nearly $10B investment in Forward Deployed Engineer (FDE) model. Model interchangeability is now assumed; scarce resource moves from model parameters to engineering capability of embedding AI into business processes. This signals a fundamental paradigm shift in enterprise AI deployment.

Huawei Other 2026-07-06

Huawei Unveils Tao's Law V2: Kirin 2026 Boosts AI Inference 40% on Same Node

Huawei's He Tingbo releases Tao's Law V2, detailing Kirin 2026 metrics: 238 MTr/mm² transistor density (+55%), 41% power reduction at iso-performance, and 40% SRAM frequency increase. Without EUV lithography, co-optimization of architecture, circuit, and process delivers equivalent performance gains, proving system-level optimization as a viable alternative to Moore's Law scaling.

Google Cloud Other 2026-07-06

Google Cloud Launches Blackwell GPU Confidential VM & Open-Source Prompt Encryption SDK, Redefining AI Security

Google Cloud upgrades its confidential computing portfolio with Blackwell GPU-based confidential VMs (Confidential G4 VMs preview), open-source Prompt Encryption SDK, and enhanced Confidential Space featuring Intel Trust Authority and Hopper GPU support, addressing TEE vulnerability CVE-2026-33697 to bolster AI inference and cross-organization training security.

Microsoft Other 2026-07-03

Microsoft Launches $2.5B AI Deployment Unit, Slashes China R&D

Microsoft establishes Microsoft Frontier Company with $2.5B and 6,000 staff to focus on enterprise AI deployment. Simultaneously cuts 200-400 Azure R&D roles in Beijing and Shanghai, signaling retreat from China.