Filter

×
Active Filters Clear All
Keyword: 0% ×
1254 Total Reports
10/63 Page
Meta Other 2026-07-07

Meta Cuts 1,395 Reality Labs Jobs, Pivots to AI Cloud to Challenge AWS and Azure

Meta plans to lay off 1,395 employees in July 2026, primarily from Reality Labs, while raising capex to $125-145B to focus on AI infrastructure. It is building a cloud business to sell AI compute externally, signaling a strategic pivot from AR/VR to AI cloud services.

OpenAI Other 2026-07-07

OpenAI Accepts US Gov Pre-Release AI Model Review, Regulatory Framework Reshapes Deployment Cadence

OpenAI commits to a voluntary US government framework requiring 30-day pre-release access for safety evaluation of frontier AI models. This shift from pure market-driven to regulated deployment will affect release cadence for models like GPT-5. Anthropic also signals participation.

Microsoft Other 2026-07-07

AI Giants Bet $10B on Forward Deployed Engineers: Control Shifts from Models to Engineering

Microsoft, OpenAI, Anthropic, and AWS collectively announced nearly $10B investment in Forward Deployed Engineer (FDE) model. Model interchangeability is now assumed; scarce resource moves from model parameters to engineering capability of embedding AI into business processes. This signals a fundamental paradigm shift in enterprise AI deployment.

NVIDIA Other 2026-07-07

NVIDIA Denies Kyber NVL144 Delay, But 78-Layer PCB Bottleneck Exposes AI Hardware Physics Limit

NVIDIA officially denies reports of Kyber NVL144 rack delay to 2028, but SemiAnalysis revelations about a 78-layer ultra-high-density PCB midplane bottleneck and Rubin Ultra cancellation expose hard physical limits in signal integrity and manufacturing, opening a strategic window for AMD and Google.

Amazon Other 2026-07-06

AWS boosts Trainium 3 shipments, accelerating ASIC substitution for NVIDIA GPUs

Supply chain sources indicate Amazon AWS has instructed vendors to increase Trainium 3 shipments for Q3 2026 by 20-30%. This signals strong confidence in its custom ASIC strategy to reduce dependence on NVIDIA GPUs, leveraging superior cost and power efficiency for cloud AI training.

NVIDIA Other 2026-07-06

NVIDIA Kyber NVL144 Delayed to 2028: Midplane PCB Manufacturing Becomes AI Scaling Bottleneck

SemiAnalysis reveals NVIDIA's Kyber NVL144 delayed beyond 12 months to 2028 due to 78-layer Orthogonal Backplane manufacturing challenges. The interim NVL72x2 solution is cancelled due to operational burdens, and the 4-die Rubin Ultra is also scrapped, leaving a product gap in NVIDIA's scaling roadmap.

Huawei Other 2026-07-06

Huawei Unveils Tao's Law V2: Kirin 2026 Boosts AI Inference 40% on Same Node

Huawei's He Tingbo releases Tao's Law V2, detailing Kirin 2026 metrics: 238 MTr/mm² transistor density (+55%), 41% power reduction at iso-performance, and 40% SRAM frequency increase. Without EUV lithography, co-optimization of architecture, circuit, and process delivers equivalent performance gains, proving system-level optimization as a viable alternative to Moore's Law scaling.

OpenAI Other 2026-07-06

OpenAI Launches GPT-5.6 Series, Regulatory Compliance Becomes Prerequisite for Frontier Models

OpenAI releases GPT-5.6 series with Sol achieving 96.7% SOTA on Terminal-Bench 2.1 via Ultra mode with sub-agent parallelism. Terra matches GPT-5.5 at half price, Luna for low-cost high-concurrency. Initial access limited to 20 trusted partners, subject to US government safety review.

Anthropic Other 2026-07-06

Anthropic Starts Custom AI Chip Development, Talks Samsung 2nm, Aims for Compute Independence

Anthropic has initiated its own AI chip development and is in talks with Samsung for 2nm foundry services. The move aims to reduce reliance on NVIDIA GPUs, optimize inference costs, and strengthen its technology moat ahead of a potential IPO. It joins OpenAI, Google, and others in the custom ASIC race, signaling a shift from software to hardware competition.

AMD Other 2026-07-06

AMD Unveils Zen 6/7 CPU and MI400/500 GPU Roadmap, Targets NVIDIA Rubin with HBM4 and 2nm

AMD unveiled its Zen 6/7 CPU and MI400/500 GPU roadmap at its 2026 Financial Analyst Day, featuring TSMC 2nm process and HBM4 memory. The MI400 series boasts 432GB memory, 19.6TB/s bandwidth, and 40 PFLOPs FP4 performance, directly targeting NVIDIA's Vera Rubin architecture with an annual cadence to disrupt the AI hardware monopoly.

Amazon Other 2026-07-06

AWS Trainium 3 Shipments Surge 20-30%, Shifting AI Compute Control from NVIDIA to Custom Silicon

Supply chain sources indicate AWS has raised Q3 Trainium 3 server shipments by 20-30%, driven by Anthropic. Trainium 2 is sold out, Trainium 3 nearly fully booked, with customers already queuing for Trainium 4 and development of Trainium 5 underway. This signals AWS's aggressive push to own the AI compute stack via custom silicon.

Anthropic Other 2026-07-06

Anthropic's $15B Australia Bet: AI Infra Shifts to Energy Arbitrage

Anthropic plans to invest $15B to secure 1.4GW of data center capacity in Australia, aiming to activate 1GW by next year. This move bypasses US grid bottlenecks from local opposition and litigation, building a hybrid model of self-build, partnerships, and cloud leasing. It signals a shift in AI infra deployment toward energy and regulatory arbitrage.

Google Cloud Other 2026-07-06

Google Cloud Launches Blackwell GPU Confidential VM & Open-Source Prompt Encryption SDK, Redefining AI Security

Google Cloud upgrades its confidential computing portfolio with Blackwell GPU-based confidential VMs (Confidential G4 VMs preview), open-source Prompt Encryption SDK, and enhanced Confidential Space featuring Intel Trust Authority and Hopper GPU support, addressing TEE vulnerability CVE-2026-33697 to bolster AI inference and cross-organization training security.

Anthropic Other 2026-07-05

Anthropic Launches Custom AI Chip: Vertical Integration to Control Inference Cost and Supply

Anthropic launched Claude Sonnet 5 and revealed a custom AI chip initiative, using Samsung foundry. This move aims to reduce dependency on NVIDIA, control long-term inference costs, and marks Anthropic's shift from a pure software company to a vertically integrated infrastructure firm.

OpenAI Other 2026-07-05

OpenAI Winds Down Fine-Tuning API: A Strategic Shift in AI Customization Landscape

OpenAI plans to phase out its fine-tuning API by 2027, stopping new task creation but allowing inference on existing models. This forces startups relying on fine-tuning for differentiation to migrate to open-source models or RAG, reshaping the AI customization ecosystem.

Cloudflare Other 2026-07-05

Cloudflare Default Blocks AI Crawlers: Infrastructure Layer Becomes Data Gatekeeper

Cloudflare announces default blocking of hybrid AI crawlers (e.g., Googlebot) for all sites starting Sept 15, allowing only pure search index crawlers unless manually overridden. This shifts AI data access control from websites/search engines to the CDN infrastructure layer, paired with a 'Pay Per Use' model to redefine content value exchange.

OpenAI Other 2026-07-05

OpenAI Ends Azure Exclusivity: Model Delivery Control Shifts from Microsoft to Multi-Cloud

OpenAI and Microsoft restructured their partnership in April 2026, ending exclusive Azure licensing and capacity commitments. OpenAI can now serve customers on any cloud; Microsoft retains right of first refusal and revenue share only on its platform. Driven by GPT-5.1's ~3 exaflops inference demand and FTC antitrust scrutiny.

Intel Other 2026-07-04

Critical Relay Attack Found in Attestation TLS Protocol: Both Intel TDX and AMD SEV-SNP Affected

A critical architecture flaw in the attestation TLS protocol, enabling relay attacks, has been discovered affecting both Intel TDX and AMD SEV-SNP platforms. With a CVSS score of 7.5, it surpasses recent high-profile confidential computing vulnerabilities. No official patch is currently available.

NVIDIA Other 2026-07-04

NVIDIA Vera Rubin AI Platform Slated for July 2026 Shipments, Iterative Compute Upgrade

NVIDIA confirms its next-gen AI compute platform, Vera Rubin, will start shipping in July 2026 to major cloud providers like Microsoft and Google. The platform uses an advanced process node to boost AI training and inference performance, representing an iterative upgrade over Hopper and Blackwell without a fundamental architectural shift.

NVIDIA Other 2026-07-04

英伟达RTX 5080公版显卡将在BW2026限量发售,售价8299元

...