Reports
AI-generated structured vendor updates
Anthropic企业AI采用首超OpenAI 300亿年化收入运行率确认
...
Cisco Locks AI Data Center Security Control Plane with Silicon One and Hypershield
Cisco launches next-gen security for AI data centers, deeply integrating Splunk SIEM with its Silicon One 51.2Tbps chip and Hypershield architecture to push security policies to the network edge. This move aims to shift the security control plane from standalone appliances to its proprietary ASIC and management platform, creating hardware lock-in.
MediaTek and Alibaba Cloud Deploy Tongyi Qianwen LLM on Dimensity Chips
MediaTek partners with Alibaba Cloud to deploy a small version of the Tongyi Qianwen LLM on Dimensity 9300/8300 mobile platforms, enabling offline multi-turn conversations. This move aims to capture edge AI inference control via NPU optimization and SDK integration, directly challenging Qualcomm.
Cloudflare Ultimatum: Mandatory AI Crawler Separation to Control Web Data Flow
Cloudflare mandates that AI companies must separate search crawlers from AI training/agent crawlers by Sep 15, 2026, or face global blocking. It also launches Monetization Gateway and Pay Per Use, using digital fingerprinting to charge for content citation. This will reshape AI data acquisition and may worsen data scarcity.
AWS Boosts Trainium3 ASIC Shipments, Accelerating Custom AI Chip Ecosystem Against NVIDIA
Amazon AWS has notified its supply chain to increase Q3 2026 shipments of Trainium3-based ASIC servers by 20-30%. This reflects growing confidence in its custom AI chips and a strategic push to reduce reliance on NVIDIA GPUs. AWS also partnered with OpenAI to develop a Stateful Runtime Environment on Bedrock.
Meta Cuts 1,395 Reality Labs Jobs, Pivots to AI Cloud to Challenge AWS and Azure
Meta plans to lay off 1,395 employees in July 2026, primarily from Reality Labs, while raising capex to $125-145B to focus on AI infrastructure. It is building a cloud business to sell AI compute externally, signaling a strategic pivot from AR/VR to AI cloud services.
OpenAI Accepts US Gov Pre-Release AI Model Review, Regulatory Framework Reshapes Deployment Cadence
OpenAI commits to a voluntary US government framework requiring 30-day pre-release access for safety evaluation of frontier AI models. This shift from pure market-driven to regulated deployment will affect release cadence for models like GPT-5. Anthropic also signals participation.
AI Giants Bet $10B on Forward Deployed Engineers: Control Shifts from Models to Engineering
Microsoft, OpenAI, Anthropic, and AWS collectively announced nearly $10B investment in Forward Deployed Engineer (FDE) model. Model interchangeability is now assumed; scarce resource moves from model parameters to engineering capability of embedding AI into business processes. This signals a fundamental paradigm shift in enterprise AI deployment.
NVIDIA Denies Kyber NVL144 Delay, But 78-Layer PCB Bottleneck Exposes AI Hardware Physics Limit
NVIDIA officially denies reports of Kyber NVL144 rack delay to 2028, but SemiAnalysis revelations about a 78-layer ultra-high-density PCB midplane bottleneck and Rubin Ultra cancellation expose hard physical limits in signal integrity and manufacturing, opening a strategic window for AMD and Google.
AWS boosts Trainium 3 shipments, accelerating ASIC substitution for NVIDIA GPUs
Supply chain sources indicate Amazon AWS has instructed vendors to increase Trainium 3 shipments for Q3 2026 by 20-30%. This signals strong confidence in its custom ASIC strategy to reduce dependence on NVIDIA GPUs, leveraging superior cost and power efficiency for cloud AI training.
NVIDIA Kyber NVL144 Delayed to 2028: Midplane PCB Manufacturing Becomes AI Scaling Bottleneck
SemiAnalysis reveals NVIDIA's Kyber NVL144 delayed beyond 12 months to 2028 due to 78-layer Orthogonal Backplane manufacturing challenges. The interim NVL72x2 solution is cancelled due to operational burdens, and the 4-die Rubin Ultra is also scrapped, leaving a product gap in NVIDIA's scaling roadmap.
Huawei Unveils Tao's Law V2: Kirin 2026 Boosts AI Inference 40% on Same Node
Huawei's He Tingbo releases Tao's Law V2, detailing Kirin 2026 metrics: 238 MTr/mm² transistor density (+55%), 41% power reduction at iso-performance, and 40% SRAM frequency increase. Without EUV lithography, co-optimization of architecture, circuit, and process delivers equivalent performance gains, proving system-level optimization as a viable alternative to Moore's Law scaling.
OpenAI Launches GPT-5.6 Series, Regulatory Compliance Becomes Prerequisite for Frontier Models
OpenAI releases GPT-5.6 series with Sol achieving 96.7% SOTA on Terminal-Bench 2.1 via Ultra mode with sub-agent parallelism. Terra matches GPT-5.5 at half price, Luna for low-cost high-concurrency. Initial access limited to 20 trusted partners, subject to US government safety review.
Anthropic Starts Custom AI Chip Development, Talks Samsung 2nm, Aims for Compute Independence
Anthropic has initiated its own AI chip development and is in talks with Samsung for 2nm foundry services. The move aims to reduce reliance on NVIDIA GPUs, optimize inference costs, and strengthen its technology moat ahead of a potential IPO. It joins OpenAI, Google, and others in the custom ASIC race, signaling a shift from software to hardware competition.
AMD Unveils Zen 6/7 CPU and MI400/500 GPU Roadmap, Targets NVIDIA Rubin with HBM4 and 2nm
AMD unveiled its Zen 6/7 CPU and MI400/500 GPU roadmap at its 2026 Financial Analyst Day, featuring TSMC 2nm process and HBM4 memory. The MI400 series boasts 432GB memory, 19.6TB/s bandwidth, and 40 PFLOPs FP4 performance, directly targeting NVIDIA's Vera Rubin architecture with an annual cadence to disrupt the AI hardware monopoly.
AWS Trainium 3 Shipments Surge 20-30%, Shifting AI Compute Control from NVIDIA to Custom Silicon
Supply chain sources indicate AWS has raised Q3 Trainium 3 server shipments by 20-30%, driven by Anthropic. Trainium 2 is sold out, Trainium 3 nearly fully booked, with customers already queuing for Trainium 4 and development of Trainium 5 underway. This signals AWS's aggressive push to own the AI compute stack via custom silicon.
Anthropic's $15B Australia Bet: AI Infra Shifts to Energy Arbitrage
Anthropic plans to invest $15B to secure 1.4GW of data center capacity in Australia, aiming to activate 1GW by next year. This move bypasses US grid bottlenecks from local opposition and litigation, building a hybrid model of self-build, partnerships, and cloud leasing. It signals a shift in AI infra deployment toward energy and regulatory arbitrage.
Google Cloud Launches Blackwell GPU Confidential VM & Open-Source Prompt Encryption SDK, Redefining AI Security
Google Cloud upgrades its confidential computing portfolio with Blackwell GPU-based confidential VMs (Confidential G4 VMs preview), open-source Prompt Encryption SDK, and enhanced Confidential Space featuring Intel Trust Authority and Hopper GPU support, addressing TEE vulnerability CVE-2026-33697 to bolster AI inference and cross-organization training security.
Anthropic Launches Custom AI Chip: Vertical Integration to Control Inference Cost and Supply
Anthropic launched Claude Sonnet 5 and revealed a custom AI chip initiative, using Samsung foundry. This move aims to reduce dependency on NVIDIA, control long-term inference costs, and marks Anthropic's shift from a pure software company to a vertically integrated infrastructure firm.
OpenAI Winds Down Fine-Tuning API: A Strategic Shift in AI Customization Landscape
OpenAI plans to phase out its fine-tuning API by 2027, stopping new task creation but allowing inference on existing models. This forces startups relying on fine-tuning for differentiation to migrate to open-source models or RAG, reshaping the AI customization ecosystem.