Filter

×
Active Filters Clear All
Keyword: 开源 ×
86 Total Reports
2/5 Page
NVIDIA Other 2026-07-21

NVIDIA Open-Sources Cosmos 3 Edge 4B World Model for Real-Time Robot Control at 15Hz on Jetson Thor

NVIDIA open-sources Cosmos 3 Edge, a 4B parameter world action model for edge robotics. It achieves 15Hz real-time inference with 32 actions per inference on the Jetson Thor module. This extends NVIDIA's physical AI stack from training to real-time deployment, enabling end-to-end robot control at the edge.

NVIDIA Other 2026-07-20

NVIDIA Expands Agent Toolkit with Omniverse Libraries for Physical AI Simulation

At SIGGRAPH 2026, NVIDIA announced an expansion to its Agent Toolkit, adding Omniverse libraries that enable AI agents to build and simulate 3D worlds. The company also open-sourced Cosmos 3 Edge, a 4B-parameter world action model, completing its physical AI ecosystem from training to edge deployment.

AMD Other 2026-07-20

Microsoft Azure Deploys AMD Helios Rack with MI455X GPUs, Breaking NVIDIA's Cloud AI Monopoly

Microsoft Azure officially adopts AMD Helios rack-scale AI infrastructure, featuring 72 MI455X GPUs (432GB HBM4, 19.6TB/s), Venice EPYC CPUs, and Pensando DPUs. Three new instances (ND MI455X v7, HDv2, HXv2) are launched, marking Azure's shift from exclusive NVIDIA dependency to a multi-vendor AI strategy.

Meta Other 2026-07-20

Meta Launches Muse Spark 1.1 API at 25% Competitor Price, Ends Open-Source Era

Meta releases Muse Spark 1.1, a multimodal reasoning model with 1M token context window, and launches its first paid API at 25% of competitors' price. This ends the Llama open-source era, signaling a strategic shift to proprietary API monetization and aggressive market share capture.

Other Other 2026-07-20

Moonshot AI Launches Kimi K3: 2.8T Parameter Open-Source MoE Model at $3/$15

Moonshot AI unveils Kimi K3, a 2.8T parameter open-source MoE model with 896 experts (16 active), native vision, and 1M context. API pricing at $3/$15 input/output per million tokens undercuts rivals. Open weights release July 27. GPU capacity exhausted within 2 days. Arena score 1679 tops Fable 5.

NVIDIA Other 2026-07-20

NVIDIA Agent Toolkit Shifts AI Agent Control from Cloud to Local DGX Station

NVIDIA launches Agent Toolkit for DGX Station, comprising NemoClaw, Nemotron 3 Ultra, Omniverse Libraries, and OpenShell. It enables local AI agent deployment in 30 minutes, locking developers into NVIDIA's hardware-software stack and shifting control from cloud services to on-premises hardware.

Other Other 2026-07-19

Alibaba Launches 2.4T Parameter Qwen3.8-Max MoE Model with 0.2x Pricing

Alibaba released Qwen3.8-Max-Preview, a 2.4 trillion parameter MoE multimodal model with 1M context window. It launched Qoder platform and Token Plan with aggressive discounts up to 0.2x, significantly reducing inference cost. The company claims it is second only to Anthropic Fable 5, marking China's AI entry into dual-track of parameter arms race and open-source competition.

Microsoft Other 2026-07-16

Microsoft Replaces OpenAI/Anthropic with In-House MAI Models to Cut Costs and Reduce Dependency

Microsoft has started replacing OpenAI and Anthropic AI calls in Excel and Outlook with its in-house MAI models, handling tens of thousands of prompts weekly. The move aims to cut costs and reduce dependency on Anthropic, signaling a strategic shift toward internal AI models and impacting the AI vendor ecosystem.

Microsoft Other 2026-07-15

Microsoft Releases Go SDK for Agent Framework, Challenging Google in Go Ecosystem

In July 2026, Microsoft released the Go SDK for Agent Framework in public preview, supporting MCP and multi-agent coordination. This positions Microsoft alongside Google as the only major cloud vendors offering native Go Agent SDKs, while OpenAI and Anthropic lag with Python-only support, risking developer ecosystem erosion.

Apple Other 2026-07-10

PrismML's 1-bit Compression: 27B Qwen Model Runs Fully on iPhone 17 Pro in 4GB

PrismML compressed a 27B-parameter dense LLM (Qwen 3.6) to 4GB, running fully on iPhone 17 Pro. Using native 1-bit quantization (weights as {-1, +1}), it achieves >92% compression, 8x faster inference, and 75-80% energy reduction. This challenges Apple's sparse architecture, potentially shifting edge AI from cloud-reliant to device-native.

OpenAI Other 2026-07-09

OpenAI Reopens with GPT-oss Models: Apache 2.0 License Hides Cloud Offload Control

OpenAI launches GPT-oss-120b and GPT-oss-20b under Apache 2.0 license, capable of running on a single 80GB GPU. However, a built-in cloud offload mechanism routes complex queries to proprietary models, masking a strategic control point shift behind the open-source facade.

Microsoft Other 2026-07-07

AI Giants Bet $10B on Forward Deployed Engineers: Control Shifts from Models to Engineering

Microsoft, OpenAI, Anthropic, and AWS collectively announced nearly $10B investment in Forward Deployed Engineer (FDE) model. Model interchangeability is now assumed; scarce resource moves from model parameters to engineering capability of embedding AI into business processes. This signals a fundamental paradigm shift in enterprise AI deployment.

Google Cloud Other 2026-07-06

Google Cloud Launches Blackwell GPU Confidential VM & Open-Source Prompt Encryption SDK, Redefining AI Security

Google Cloud upgrades its confidential computing portfolio with Blackwell GPU-based confidential VMs (Confidential G4 VMs preview), open-source Prompt Encryption SDK, and enhanced Confidential Space featuring Intel Trust Authority and Hopper GPU support, addressing TEE vulnerability CVE-2026-33697 to bolster AI inference and cross-organization training security.

OpenAI Other 2026-07-05

OpenAI Winds Down Fine-Tuning API: A Strategic Shift in AI Customization Landscape

OpenAI plans to phase out its fine-tuning API by 2027, stopping new task creation but allowing inference on existing models. This forces startups relying on fine-tuning for differentiation to migrate to open-source models or RAG, reshaping the AI customization ecosystem.

Research Other 2026-06-30

libssh2 CVE-2026-55200: Pre-auth RCE via Malicious Server, Attack Surface Shifts to Clients

A critical heap out-of-bounds write vulnerability (CVE-2026-55200, CVSS 9.2) in libssh2 allows a malicious SSH server to achieve pre-auth RCE on connecting clients. The flaw affects curl, Git, PHP, and many other projects statically linking the library, expanding the attack surface from servers to virtually any client application, including CI/CD, backup, and embedded systems.

NVIDIA Other 2026-06-29

Jia Yangqing exits NVIDIA as DGX Lepton shutdown reveals software layer failure

Jia Yangqing leaves NVIDIA after DGX Lepton underperforms and open-source commitments are broken. NVIDIA acquired Lepton AI for ~$700M, rebranded as DGX Cloud Lepton, but service ceased mid-2025. The event signals NVIDIA's failed software layer expansion, shifting control back to hyperscalers.

Qualcomm Other 2026-06-26

Qualcomm Acquires Modular for $3.9B, Open-Sources Mojo to Break CUDA Lock-In

Qualcomm acquires Modular for $3.9B in stock and open-sources Mojo, a Python-compatible systems language. Mojo targets CUDA dependency, aiming to provide a high-performance alternative for AI developers. This move strengthens Qualcomm's AI inference chip software stack and edge AI competitiveness.

Google Other 2026-06-23

FSFE Accuses Google of Silently Reinstalling AI Components on Android, DMA Compliance Under Fire

The Free Software Foundation Europe (FSFE) has filed a complaint with the European Commission, alleging Google silently reinstalls AI models on Android after user uninstallation, violating the Digital Markets Act (DMA). FSFE demands users be able to fully remove preloaded AI components and be protected from silent reinstalls. The dispute also targets Google's upcoming developer verification program, which could restrict access to alternative app stores like F-Droid.

Google Other 2026-06-19

Google Deprecates Open-Source Gemini CLI, Forces Migration to Closed-Source Antigravity

On June 18, 2026, Google deprecated the open-source Gemini CLI (Apache 2.0, 6000+ community PRs) for free users, mandating migration to the closed-source, Go-rewritten Antigravity CLI. Enterprise users retain Gemini CLI access, while a new AI Ultra tier ($100/month) offers 5x Antigravity quotas. Antigravity 2.0 replaces traditional IDE with Agent, signaling a strategic shift from open to proprietary developer tooling.

Microsoft Other 2026-06-18

Microsoft Shifts Copilot Cowork to Usage-Based Pricing, Eyes DeepSeek for Cost-Efficiency

Microsoft transitions Copilot Cowork to usage-based billing (Copilot Credits) and considers integrating fine-tuned DeepSeek V4 or open-source models as low-cost alternatives, hosted on Azure. This move addresses high costs from intensive usage and signals a multi-model strategy.