Filter

×
Active Filters Clear All
Keyword: 算力 ×
179 Total Reports
2/9 Page
OpenAI Other 2026-07-24

OpenAI Launches Presence Agent Platform, Shifts Control Plane to Lock Enterprise AI Deployment

OpenAI launches Presence, an enterprise agent deployment platform integrating model inference, permissions, policies, evaluation, and escalation tools, shifting the control plane from models to the platform. ChatGPT Health is now fully available to US users 18+, integrating Apple Health, accelerating consumer AI agent adoption.

Google Other 2026-07-24

Google Begins Gemini 4 Pre-training with 4M+ Context and Monthly Releases

Alphabet confirms start of largest pre-training run for Gemini 4, featuring 4M+ context and native multimodality with near-monthly releases. 2026 capex raised to $195-205B, Google Cloud Q2 up 82%, signaling full-stack AI acceleration.

OpenAI Other 2026-07-24

OpenAI Launches Project Camellia: $20B Self-Built AI Data Center, Shift from Cloud Renting to Ownership

OpenAI announces Project Camellia, a $20B self-built AI data center in Georgia with 3.2GW power, marking a shift from cloud renting to self-control. It hires key xAI Colossus team members and raises compute spending forecast to $750B.

Microsoft Other 2026-07-24

AMD Helios Full-Stack AI Server Deployed on Azure, UALoE Open Standard Challenges NVLink

AMD and Microsoft Azure announce large-scale deployment of Helios full-stack AI servers, featuring 72 MI455X GPUs, 18 EPYC Venice CPUs, and Pensando DPUs, with 31TB HBM4 and 1.7PB/s memory bandwidth. Two new Azure VM series target Agentic AI and semiconductor design, marking Microsoft's multi-vendor strategy and challenging NVIDIA's dominance.

AMD Other 2026-07-23

AMD Invests $5B in Anthropic, Secures 2GW MI450 Deployment, Reshaping AI Compute Ecosystem

AMD and Anthropic announce a strategic partnership: Anthropic will deploy up to 2GW of AMD Instinct MI450 GPUs, with AMD investing up to $5B in Anthropic. They will collaborate on ROCm optimization and Claude workload tuning, marking AMD's transition from chip vendor to AI ecosystem investor and accelerating multi-sourcing in AI compute.

NVIDIA Other 2026-07-23

NVIDIA-OpenAI $100B Partnership: 10GW Vera Rubin AI Factories Reshape Ecosystem

NVIDIA and OpenAI announce a strategic partnership to deploy at least 10GW of NVIDIA systems using the Vera Rubin platform (Rubin GPU, Vera CPU, HBM4, NVLink 6). NVIDIA will invest up to $100B. First facilities go online in H2 2026, powering OpenAI's next-gen models, marking the era of multi-GW AI factories.

AMD Other 2026-07-23

AMD Unveils Zen 6 Venice, MI455X, and Helios Rack-Level Design to Challenge NVIDIA

At Advancing AI 2026, AMD launched Zen 6 EPYC Venice (2nm, up to 256 cores) and MI455X (CDNA5, 432GB HBM4, 40 PFLOPS FP4), along with Helios rack reference design (2.9 exaFLOPS FP4 per rack), claiming a 1000x AI performance roadmap, with major commitments from Meta, OpenAI, and others.

Huawei Other 2026-07-22

华为昇腾950超节点亮相WAIC 2026,单柜64卡支持8192卡高速互联

...

Microsoft Other 2026-07-22

Microsoft and Mistral Partner to Build Sovereign AI Infrastructure for Regulated European Industries

Microsoft and Mistral expand their partnership with a multi-billion dollar deal. Mistral gains thousands of NVIDIA Vera Rubin GPUs and integrates its Medium 3.5 and OCR 4 models into Microsoft Foundry and Copilot Studio, offering cloud, connected, and offline deployment modes for European regulated industries under EU AI Act.

Microsoft Other 2026-07-21

Microsoft Invests Billions in Mistral AI, Integrates Models into Azure for Sovereign AI

Microsoft and Mistral AI announce a multi-billion dollar partnership, with Microsoft investing in Mistral's European data center capacity and integrating Mistral's Medium 3.5 and OCR 4 models into Azure Foundry. The deal directly responds to US export controls on Anthropic, offering regulated European industries a sovereign AI alternative, signaling a shift from centralized US AI to localized infrastructure.

NVIDIA Other 2026-07-21

NVIDIA Open-Sources Cosmos 3 Edge 4B World Model for Real-Time Robot Control at 15Hz on Jetson Thor

NVIDIA open-sources Cosmos 3 Edge, a 4B parameter world action model for edge robotics. It achieves 15Hz real-time inference with 32 actions per inference on the Jetson Thor module. This extends NVIDIA's physical AI stack from training to real-time deployment, enabling end-to-end robot control at the edge.

Google Other 2026-07-20

Google's Frozen v2 Chip Hardwires Gemini Architecture for 6-10x TPU Efficiency, Set for 2028

Google is developing Frozen v2, a dedicated AI chip that hardwires the Gemini model architecture into silicon for 6-10x energy efficiency per token over current TPUs. It is a new product line, planned for 2028, with weight update flexibility but a frozen architecture. This validates the industry shift from general-purpose GPUs to dedicated ASICs for AI inference.

NVIDIA Other 2026-07-20

NVIDIA Invests $2B in CoreWeave, Debuts 'Compute Central Bank' Model

NVIDIA invests $2 billion in CoreWeave and launches the AI Compute Partner Program, featuring credit enhancement, revenue sharing, and GPU buyback. This transforms NVIDIA from a hardware vendor into a 'compute central bank', tightening control over the AI cloud leasing ecosystem and squeezing intermediaries.

Meta Other 2026-07-20

Meta Launches Muse Spark 1.1 API at 25% Competitor Price, Ends Open-Source Era

Meta releases Muse Spark 1.1, a multimodal reasoning model with 1M token context window, and launches its first paid API at 25% of competitors' price. This ends the Llama open-source era, signaling a strategic shift to proprietary API monetization and aggressive market share capture.

Google Cloud Other 2026-07-20

Google DeepMind AlphaEvolve GA: AI Self-Evolution for Data Center and Algorithm Optimization

On July 19, 2026, Google DeepMind announced the GA of AlphaEvolve, a Gemini-based multi-agent evolution system for algorithmic discovery, mathematical discovery, and data center efficiency optimization, already used in Borg and Orca, aiming to reduce Capex in massive AI compute investments.

Other Other 2026-07-20

Moonshot AI Launches Kimi K3: 2.8T Parameter Open-Source MoE Model at $3/$15

Moonshot AI unveils Kimi K3, a 2.8T parameter open-source MoE model with 896 experts (16 active), native vision, and 1M context. API pricing at $3/$15 input/output per million tokens undercuts rivals. Open weights release July 27. GPU capacity exhausted within 2 days. Arena score 1679 tops Fable 5.

NVIDIA Other 2026-07-20

NVIDIA Agent Toolkit Shifts AI Agent Control from Cloud to Local DGX Station

NVIDIA launches Agent Toolkit for DGX Station, comprising NemoClaw, Nemotron 3 Ultra, Omniverse Libraries, and OpenShell. It enables local AI agent deployment in 30 minutes, locking developers into NVIDIA's hardware-software stack and shifting control from cloud services to on-premises hardware.

NVIDIA Other 2026-07-20

测试来源:36氪

...

Huawei Other 2026-07-19

Huawei Ascend 950 Supernode: Self-Developed HCCS Interconnect for Sovereign AI Compute Ecosystem

Huawei unveiled the Ascend 950 Supernode at WAIC, integrating 32 self-developed Ascend 950 AI processors with HCCS high-speed interconnect, achieving 2.5x compute density and supporting trillion-parameter model training, offering a sovereign AI compute alternative free from overseas supply chains.

NVIDIA Other 2026-07-19

NVIDIA Vera Rubin Platform and Dynamo 1.0 Disaggregate Inference, Shift Focus to Intelligence per Dollar

NVIDIA unveils Vera Rubin platform with a 7-chip stack (Vera CPU, Rubin GPU, NVLink 6, etc.) and Dynamo 1.0 inference disaggregation. A single NVL72 rack packs 72 GPUs/36 CPUs with 1.6 PB/s bandwidth, achieving up to 7x inference performance. The new 'intelligence per dollar' metric signals a shift from training to inference cost competition.