Filter

×
Active Filters Clear All
Keyword: 0% ×
1254 Total Reports
9/63 Page
Samsung Electronics Other 2026-07-10

Samsung GAIA AI PC Chip Samples with Memory-Centric NPU, Targeting 50 TOPS

Samsung launches GAIA AI PC processor with 4nm process and memory-centric NPU, integrating LPDDR5X controller with NPU for near-memory computing, achieving 40% energy efficiency improvement and 50 TOPS. Certified for Microsoft Copilot+ PC, Lenovo to adopt in Q4 2026.

Amazon Other 2026-07-10

AWS Sells Trainium 3 Externally, Challenging NVIDIA's AI Training Chip Dominance

AWS begins external sales of its Trainium 3 AI training chip, fabricated on TSMC 3nm process, delivering 2.52 PFLOPS per chip. Early customers include Anthropic and Uber. This move directly challenges NVIDIA's dominance and marks AWS's strategic shift from cloud provider to chip vendor.

AMD Other 2026-07-10

AMD's Experimental Topological Ghost Protocol Boosts MI300X Inference 10x

AMD introduces experimental Topological Ghost Protocol (TGP) on MI300X GPUs, achieving 431 tokens/sec with 100% success in high-concurrency inference, 10x improvement over standard vLLM. TGP uses KV-cache recycling and segmented state management, still experimental but potentially redefining AI inference benchmarks.

AMD Other 2026-07-10

Towards Feature Complete Triton Support in JAX-Triton — ROCm Blogs

...

OpenAI Other 2026-07-09

OpenAI GPT-5.6发布测试

...

OpenAI Other 2026-07-09

OpenAI Reopens with GPT-oss Models: Apache 2.0 License Hides Cloud Offload Control

OpenAI launches GPT-oss-120b and GPT-oss-20b under Apache 2.0 license, capable of running on a single 80GB GPU. However, a built-in cloud offload mechanism routes complex queries to proprietary models, masking a strategic control point shift behind the open-source facade.

Google Other 2026-07-09

Google Gemini 3.5 Pro Rebuilds from Scratch: 2M Token Context Window Reshapes AI Frontier

Google DeepMind targets July 17 for Gemini 3.5 Pro, a full architectural rewrite of its pretraining stack to overcome deficits in math reasoning, SVG generation, and image quality. Specs include a 2M token context window, Deep Think reasoning layer, and multi-step autonomous workflows, though unconfirmed by Google.

NVIDIA Other 2026-07-09

SambaNova完成11亿美元融资估值110亿美元:推理芯片新格局确立

...

Anthropic Other 2026-07-09

GhostApproval Vuln Exposes Systemic AI Coding Tool Flaw: Symlink Bypass in Human Review

Wiz Research discloses GhostApproval vulnerability affecting six major AI coding tools (Claude Code, Codex, Cursor, Amazon Q, Antigravity). Attackers use symlinks to bypass human review, achieving persistent remote access. The flaw reveals fundamental UI-level security gaps in Human-in-the-Loop mechanisms as agent permissions expand, requiring a redesign of confirmation workflows.

CrowdStrike Other 2026-07-08

CrowdStrike Capitalizes on 5x AIDR Growth to Enter Identity Security, Seizing AI Runtime Control Plane

CrowdStrike reports 5x growth in its AIDR product, expanding into identity security. AIDR monitors AI app data flows, detects prompt injection and model jailbreaks, and launches Shadow AI Discovery for Endpoint to auto-discover AI apps and LLM runtimes on endpoints. This signals a control plane shift from traditional endpoint detection to converged AI workload and identity security.

TSMC Other 2026-07-08

TSMC Ramps PIC Capacity to 25K Wafers, CPO Silicon Photonics Poised to Disrupt AI Interconnects

TSMC plans to expand its PIC capacity to 25,000 wafers per month by 2028, with its COUPE platform becoming critical for reducing latency and power in AI systems. Initial capacity is allocated to NVIDIA, Broadcom, and AMD, marking CPO's transition from lab to mass production and accelerating the shift from electrical to optical AI interconnects.

NVIDIA Other 2026-07-08

NVIDIA Rigel Core: Single-Threaded CPU as the New Control Plane for Agentic AI

NVIDIA unveils Rosa CPU architecture with custom Rigel core (Arm v9.2), targeting single-threaded performance for Agentic AI workloads, paired with Feynman GPU (1.6nm, 50 PFLOPS) in 2028. This shifts CPU design from core-count scaling to serial-latency optimization, directly challenging AMD EPYC and Intel Xeon dominance.

NVIDIA Other 2026-07-07

NVIDIA Vera CPU获Perplexity/OpenAI/Anthropic/Oracle采用 AI Agent性能验证1.5-1.9x加速

...

NVIDIA Other 2026-07-07

NVIDIA Vera CPU: Max Single-Threaded Performance at Scale for Agentic AI

NVIDIA launches Vera CPU, a max single-threaded CPU at scale for agentic AI. With Olympus cores delivering 1.8x sustained per-core performance over x86, 1.2TB/s LPDDR5X bandwidth, and 3.4TB/s core-to-core bandwidth, Vera integrates into NVIDIA's unified AI factory architecture, aiming to lock users into its ecosystem.

NVIDIA Other 2026-07-07

AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters

...

Anthropic Other 2026-07-07

Anthropic企业AI采用首超OpenAI 300亿年化收入运行率确认

...

Cisco Other 2026-07-07

Cisco Locks AI Data Center Security Control Plane with Silicon One and Hypershield

Cisco launches next-gen security for AI data centers, deeply integrating Splunk SIEM with its Silicon One 51.2Tbps chip and Hypershield architecture to push security policies to the network edge. This move aims to shift the security control plane from standalone appliances to its proprietary ASIC and management platform, creating hardware lock-in.

MediaTek Other 2026-07-07

MediaTek and Alibaba Cloud Deploy Tongyi Qianwen LLM on Dimensity Chips

MediaTek partners with Alibaba Cloud to deploy a small version of the Tongyi Qianwen LLM on Dimensity 9300/8300 mobile platforms, enabling offline multi-turn conversations. This move aims to capture edge AI inference control via NPU optimization and SDK integration, directly challenging Qualcomm.

Cloudflare Other 2026-07-07

Cloudflare Ultimatum: Mandatory AI Crawler Separation to Control Web Data Flow

Cloudflare mandates that AI companies must separate search crawlers from AI training/agent crawlers by Sep 15, 2026, or face global blocking. It also launches Monetization Gateway and Pay Per Use, using digital fingerprinting to charge for content citation. This will reshape AI data acquisition and may worsen data scarcity.

Amazon Other 2026-07-07

AWS Boosts Trainium3 ASIC Shipments, Accelerating Custom AI Chip Ecosystem Against NVIDIA

Amazon AWS has notified its supply chain to increase Q3 2026 shipments of Trainium3-based ASIC servers by 20-30%. This reflects growing confidence in its custom AI chips and a strategic push to reduce reliance on NVIDIA GPUs. AWS also partnered with OpenAI to develop a Stateful Runtime Environment on Bedrock.