Filter

×
Active Filters Clear All
Keyword: OpenAI ×
289 Total Reports
4/15 Page
Microsoft Other 2026-07-16

Microsoft Replaces OpenAI/Anthropic with In-House MAI Models to Cut Costs and Reduce Dependency

Microsoft has started replacing OpenAI and Anthropic AI calls in Excel and Outlook with its in-house MAI models, handling tens of thousands of prompts weekly. The move aims to cut costs and reduce dependency on Anthropic, signaling a strategic shift toward internal AI models and impacting the AI vendor ecosystem.

Intel Other 2026-07-15

Intel 18A Yield Hits 85%, Secures Orders from NVIDIA, OpenAI, Reshaping Foundry Landscape

Intel reports 18A process yield improvement to 85%, from 65% last quarter, nearing TSMC N2's 90%. Secured foundry deals with NVIDIA, AMD, OpenAI, etc. EMIB advanced packaging yield reaches 98%, used in NVIDIA Feynman, Google TPU. This marks a strategic inflection in AI chip manufacturing.

Apple Other 2026-07-15

Apple in Talks with PrismML to Compress Qwen 27B Model 15x for On-Device AI

Apple is negotiating with AI startup PrismML to deploy a compressed version of Alibaba's Qwen 27B parameter model on iPhone. PrismML's compression technology reduces memory usage by 15x, enabling 27B models to run locally with 10GB VRAM, shifting Apple's AI strategy from cloud-dependent to on-device inference.

Microsoft Other 2026-07-15

Microsoft Releases Go SDK for Agent Framework, Challenging Google in Go Ecosystem

In July 2026, Microsoft released the Go SDK for Agent Framework in public preview, supporting MCP and multi-agent coordination. This positions Microsoft alongside Google as the only major cloud vendors offering native Go Agent SDKs, while OpenAI and Anthropic lag with Python-only support, risking developer ecosystem erosion.

Other Other 2026-07-14

SANS Identifies Distributed Scanning of MCP Servers and AI Assistant Configs

SANS Internet Storm Center reports systematic scanning of MCP servers, AI assistant configs, and local LLM endpoints. 49 IPs targeted MCP handshakes, exploiting CVEs in MCP SDKs, signaling AI infrastructure as a new attack vector.

Meta Other 2026-07-13

Meta Iris Chip to Mass Produce in September: 6-Month Cadence Threatens NVIDIA GPU Hegemony

Reuters confirms Meta's Iris AI chip mass production in September, targeting 2.5GW by end-2026 and 14GW by 2027. Meta's 6-month MTIA generation cadence directly challenges NVIDIA's annual GPU cycle, signaling a hyperscaler shift from GPU dependency to custom ASIC sovereignty.

TSMC Other 2026-07-13

TSMC Hikes Sub-7nm Prices 8-12%, Extends Lead Times to 26 Weeks, Triggering AI Chip Cost Inflation

TSMC raises sub-7nm wafer prices by 8-12% and extends lead times to 26 weeks, effective July 2026. New v2.1 directive mandates EDA tool validation for PDK access. This directly inflates AI chip TCO, delays new product launches, and solidifies TSMC's control over the AI supply chain.

Microsoft Other 2026-07-12

Microsoft Takes Over OpenAI's Arctic Data Center, Seizing AI Compute Control

Microsoft leases a data center in Norway's Arctic Circle from Nscale, deploying 30,000 NVIDIA Vera Rubin GPUs, filling the gap left by OpenAI's retreat. OpenAI slashes its 2030 infrastructure budget from $140B to $60B. Microsoft surpasses OpenAI in AI compute capacity and gains geographical redundancy.

Anthropic Other 2026-07-12

Anthropic Locks 3.5GW TPU Compute with Broadcom, Signaling Shift to Custom AI ASICs

Broadcom's Q2 FY2026 filing reveals a 3.5GW TPU compute deal with Anthropic starting 2027. This marks a strategic shift from general-purpose GPUs to custom ASICs for AI workloads, with OpenAI and Meta making similar multi-GW commitments, signaling a fundamental change in AI infrastructure.

Apple Other 2026-07-10

PrismML's 1-bit Compression: 27B Qwen Model Runs Fully on iPhone 17 Pro in 4GB

PrismML compressed a 27B-parameter dense LLM (Qwen 3.6) to 4GB, running fully on iPhone 17 Pro. Using native 1-bit quantization (weights as {-1, +1}), it achieves >92% compression, 8x faster inference, and 75-80% energy reduction. This challenges Apple's sparse architecture, potentially shifting edge AI from cloud-reliant to device-native.

OpenAI Other 2026-07-09

OpenAI GPT-5.6发布测试

...

OpenAI Other 2026-07-09

OpenAI Reopens with GPT-oss Models: Apache 2.0 License Hides Cloud Offload Control

OpenAI launches GPT-oss-120b and GPT-oss-20b under Apache 2.0 license, capable of running on a single 80GB GPU. However, a built-in cloud offload mechanism routes complex queries to proprietary models, masking a strategic control point shift behind the open-source facade.

Google Other 2026-07-09

Google Gemini 3.5 Pro Rebuilds from Scratch: 2M Token Context Window Reshapes AI Frontier

Google DeepMind targets July 17 for Gemini 3.5 Pro, a full architectural rewrite of its pretraining stack to overcome deficits in math reasoning, SVG generation, and image quality. Specs include a 2M token context window, Deep Think reasoning layer, and multi-step autonomous workflows, though unconfirmed by Google.

Anthropic Other 2026-07-09

GhostApproval Vuln Exposes Systemic AI Coding Tool Flaw: Symlink Bypass in Human Review

Wiz Research discloses GhostApproval vulnerability affecting six major AI coding tools (Claude Code, Codex, Cursor, Amazon Q, Antigravity). Attackers use symlinks to bypass human review, achieving persistent remote access. The flaw reveals fundamental UI-level security gaps in Human-in-the-Loop mechanisms as agent permissions expand, requiring a redesign of confirmation workflows.

NVIDIA Other 2026-07-07

NVIDIA Vera CPU获Perplexity/OpenAI/Anthropic/Oracle采用 AI Agent性能验证1.5-1.9x加速

...

Anthropic Other 2026-07-07

Anthropic企业AI采用首超OpenAI 300亿年化收入运行率确认

...

Cloudflare Other 2026-07-07

Cloudflare Ultimatum: Mandatory AI Crawler Separation to Control Web Data Flow

Cloudflare mandates that AI companies must separate search crawlers from AI training/agent crawlers by Sep 15, 2026, or face global blocking. It also launches Monetization Gateway and Pay Per Use, using digital fingerprinting to charge for content citation. This will reshape AI data acquisition and may worsen data scarcity.

Amazon Other 2026-07-07

AWS Boosts Trainium3 ASIC Shipments, Accelerating Custom AI Chip Ecosystem Against NVIDIA

Amazon AWS has notified its supply chain to increase Q3 2026 shipments of Trainium3-based ASIC servers by 20-30%. This reflects growing confidence in its custom AI chips and a strategic push to reduce reliance on NVIDIA GPUs. AWS also partnered with OpenAI to develop a Stateful Runtime Environment on Bedrock.

OpenAI Other 2026-07-07

OpenAI Accepts US Gov Pre-Release AI Model Review, Regulatory Framework Reshapes Deployment Cadence

OpenAI commits to a voluntary US government framework requiring 30-day pre-release access for safety evaluation of frontier AI models. This shift from pure market-driven to regulated deployment will affect release cadence for models like GPT-5. Anthropic also signals participation.

Microsoft Other 2026-07-07

AI Giants Bet $10B on Forward Deployed Engineers: Control Shifts from Models to Engineering

Microsoft, OpenAI, Anthropic, and AWS collectively announced nearly $10B investment in Forward Deployed Engineer (FDE) model. Model interchangeability is now assumed; scarce resource moves from model parameters to engineering capability of embedding AI into business processes. This signals a fundamental paradigm shift in enterprise AI deployment.