Reports
AI-generated structured vendor updates
测试来源:36氪
...
Huawei Ascend 950 Supernode: Self-Developed HCCS Interconnect for Sovereign AI Compute Ecosystem
Huawei unveiled the Ascend 950 Supernode at WAIC, integrating 32 self-developed Ascend 950 AI processors with HCCS high-speed interconnect, achieving 2.5x compute density and supporting trillion-parameter model training, offering a sovereign AI compute alternative free from overseas supply chains.
Cisco Deploys AgenticOps Autonomous Network, AI Agents Take Over Operations Control
Cisco deploys AgenticOps autonomous network architecture globally, using an AI-native engine for self-healing. Over 90,000 employees use AI agents, reducing MTTR from hours to minutes, targeting 40% opex reduction. A third-party agent development framework is also launched.
Azure Arc Extends Control Plane: SQL Server Migration to Azure VM Now Unified
Microsoft has made SQL Server migration to Azure VMs generally available through Azure Arc, offering a unified guided workflow from assessment to cutover. The migration uses backup-restore and log shipping with Azure Blob as staging, but has key limitations like region binding and no migration of agent jobs or SSIS packages, aiming to strengthen Azure's unified control plane and user lock-in.
ARM Launches AGI CPU, Achieves 3x Performance Per Watt, Tops Supercomputing Rankings
At ISC 2026, ARM announced the Armv9-based LineShine supercomputer as the first to exceed 2 exaflops, topping the TOP500. It also launched the AGI CPU with 136 Neoverse V3 cores for gigawatt-scale AI datacenters, with Neoverse systems achieving 3x performance per watt over x86 and leading the Green500.
NVIDIA Vera Rubin Platform and Dynamo 1.0 Disaggregate Inference, Shift Focus to Intelligence per Dollar
NVIDIA unveils Vera Rubin platform with a 7-chip stack (Vera CPU, Rubin GPU, NVLink 6, etc.) and Dynamo 1.0 inference disaggregation. A single NVL72 rack packs 72 GPUs/36 CPUs with 1.6 PB/s bandwidth, achieving up to 7x inference performance. The new 'intelligence per dollar' metric signals a shift from training to inference cost competition.
AMD and HPE Launch Helios Open AI Infrastructure to Rival NVIDIA Ecosystem
AMD and HPE expand partnership to launch Helios, an open-stack AI infrastructure platform integrating EPYC CPUs, Instinct MI455X GPUs, Pensando networking, and ROCm software. Each rack delivers up to 2.9 exaFLOPS FP4, built on OCP principles with Juniper switches, targeting simplified deployment and energy efficiency.
Meta to lease AI compute to Anthropic, signaling infrastructure monetization push
Meta is in talks to lease AI compute capacity to Anthropic, aiming to monetize its massive infrastructure investment. This marks Meta's shift from internal consumer to external provider, potentially reshaping the AI compute market and intensifying competition with cloud providers.
EU Forces Google to Open Android AI Access and Share Search Data
The EU mandates Google to open 11 system-level Android functions to third-party AI assistants by 2027-2028, and share search click/query data from 2027, under the Digital Markets Act. Non-compliance risks fines up to $40 billion annually. This will reshape the AI assistant and search markets.
Alibaba Launches 2.4T Parameter Qwen3.8-Max MoE Model with 0.2x Pricing
Alibaba released Qwen3.8-Max-Preview, a 2.4 trillion parameter MoE multimodal model with 1M context window. It launched Qoder platform and Token Plan with aggressive discounts up to 0.2x, significantly reducing inference cost. The company claims it is second only to Anthropic Fable 5, marking China's AI entry into dual-track of parameter arms race and open-source competition.
PPIO Launches Agentic Cloud, Intelligent Model Gateway Becomes New Control Point
PPIO unveiled Agentic Cloud and Intelligent Model Gateway at WAIC 2026, targeting AI agent workloads with semantic routing and cost-aware scheduling. With over 1.2 trillion daily tokens and sub-200ms sandbox cold start, it signals the emergence of dedicated agent infrastructure.
Huawei unveils Atlas 950 SuperPoD: 1024-card memory-coherent AI supernode
Huawei unveiled the Atlas 950 SuperPoD at WAIC 2026, powered by the 950DT chip, supporting 1024 interconnected cards with 256TB unified memory addressing. Designed for trillion-parameter model training and Agentic AI inference, it marks a shift from chip-level stacking to system-level unified architecture.
FortiBleed Credential Leak Exposes 75K FortiGate Firewalls, Management Plane Security Gaps
The FortiBleed campaign leaked configs from ~75,000 internet-facing FortiGate firewalls, with admin credentials hashed using weak SHA-256 (pre-PBKDF2) easily cracked offline. This highlights risks of exposed management interfaces and inadequate password policies.
Palo Alto Networks Launches AI Gateway as Centralized Control Plane for Enterprise AI
Palo Alto Networks announces general availability of AI Gateway, integrating Portkey technology, positioned as the enterprise AI control plane. It unifies LLM, MCP, and A2A gateway execution, processing over 68 trillion tokens with sub-millisecond latency and 99.999% availability.
EU Forces Google to Open Android to Third-Party AI Assistants, Share Search Data from 2027
The European Commission mandates Google to grant third-party AI assistants (e.g., ChatGPT) system-level access on Android by Android 18 (2027), including wake word, Home button, context reading, and device AI compute. Additionally, Google must share search data with rivals from 2027, risking fines up to 10% of global revenue (~$40B).
Huawei Atlas 950 SuperPoD & 灵衢2.0: A Systemic Pivot in China's AI Compute from Chip to Cluster
At WAIC 2026, Huawei publicly demonstrated the Atlas 950 SuperPoD, a 1024-ascend NPU card cluster, and unveiled the 灵衢2.0 high-speed interconnect protocol. This signals a strategic shift in China's AI infrastructure from single-chip to system-level leadership, creating a closed-loop ecosystem that directly challenges NVIDIA's NVL series dominance.
TSMC Pledges $100B More for 6 US Fabs, Localizing 3nm for AI Chip Supply Chain
TSMC announces an additional $100B investment in Arizona, bringing total US commitment to $265B, with plans for 6 fabs focused on 3nm and beyond. This move localizes advanced process for AI chip demand from NVIDIA, Apple, AMD, reshaping global semiconductor supply chain. Q2 net profit surged 77% YoY, FY capex raised to $60-64B.
Huawei Ascend 950 SuperPoD: 1024 NPUs with 256TB Unified Memory Redefines AI Compute
Huawei unveiled the Ascend 950 SuperPoD at WAIC 2026, featuring 1024 NPUs per cabinet with 256TB unified memory and 3μs latency. This system-level innovation compensates for process limitations, scaling to 500,000 NPUs for trillion-parameter model training, shifting the compute race from single-chip to system efficiency.
Microsoft Azure Cuts 200-400 Jobs in China Amid Cloud Growth, Reshapes Geopolitical Compliance
Microsoft Azure is cutting 200-400 jobs in China even as its cloud business grows 40% YoY, offering some employees relocation to Canada. This signals a strategic shift to move compliance and operational control out of China, reshaping enterprise multi-cloud and data sovereignty decisions.
Cloudflare Launches AI Payment Gateway, Revives HTTP 402 for Bot Monetization
Cloudflare introduces an AI content payment gateway leveraging HTTP 402 and stablecoin settlements to enforce payments at the edge, enabling websites to charge AI crawlers for access. This system aims to monetize bot traffic and end free scraping.