Reports
AI-generated structured vendor updates
Microsoft Launches MAI Models, Slashes GPU Costs 89%, Reducing OpenAI Dependency
Microsoft unveiled MAI-Image-2.5-Pro and MAI-Voice-2-Flash on Azure Foundry, achieving 96.8% text rendering accuracy at 8K and reducing GPU costs by 84-89% vs GPT. Integrated across Bing, PowerPoint, and Dynamics 365, it marks a strategic shift from OpenAI dependency. Also, NVIDIA Jetson heads to the moon for edge AI.
AI Kill Switch Act: Mandatory Shutdown Powers Reshape AI Safety Infrastructure
US lawmakers propose the AI Kill Switch Act, granting DHS emergency shutdown powers over AI systems with training costs over $100M and annual revenue over $500M. Non-compliance fines reach $2M/day, with $20M/day for violating shutdown orders. Concurrently, a researcher claims a universal jailbreak affecting GPT-5.6, Claude Opus 5, and Fable.
AMD Helios Enters Production: 12-Stack HBM4 Outmuscles NVIDIA, UALoE Opens AI Network
AMD's second-generation Helios rack-scale AI server enters full production, featuring 72 MI455X GPUs with Samsung's exclusive 12-stack HBM4 (31TB per rack). Compared to NVIDIA's 8-stack design, it offers 50% more memory and 30% lower token cost. Microsoft Azure commits to large-scale deployment, solidifying hyperscaler dual-vendor strategy.
Anthropic Launches Claude Opus 5 at Half Price, Deep AWS Integration Shifts Control
Anthropic releases Claude Opus 5 with pricing unchanged from Opus 4.8 but performance approaching Fable 5, effectively halving cost. AWS announces Claude Platform GA, deeply integrating Anthropic API into AWS IAM/billing/management, shifting control from standalone API to cloud platform.
Anthropic partners with SpaceX for 220K GPUs, shifting AI compute ecosystem beyond hyperscalers
Anthropic signs a deal with SpaceX for 300MW capacity and 220,000 NVIDIA GPUs at Colossus 1 data center, with exploration of orbital AI compute. This diversifies Anthropic's compute sources beyond hyperscalers and doubles Claude Code usage limits.
Meta Transforms into AI Compute Landlord with $10B Anthropic Lease Deal
Meta is in early talks with Anthropic for a $10 billion compute lease deal, marking its strategic shift from internal AI infrastructure user to commercial landlord. This move monetizes Meta's massive capex and creates an inter-cloud leasing model, fundamentally reshaping the AI compute supply chain.
NVIDIA and SK Group Lock HBM4 Supply and Launch Sovereign AI Factory Model with $500B+ Deal
NVIDIA and SK Group announced a $500B+ AI partnership including a 2GW AI factory using Vera Rubin and HBM4, long-term HBM4 supply lock, and a $1B NVIDIA investment in Naver (with $9B from Brookfield). Samsung and Broadcom signed a $200B deal. This signals a new era of sovereign AI infrastructure and supply chain deep-locking.
AMD and Cerebras Unveil Disaggregated AI Inference with Wafer-Scale Engine
AMD and Cerebras launch a disaggregated AI inference solution combining the Helios Rackscale system (6th-gen EPYC Venice CPUs + up to 72 Instinct MI455X GPUs) with the Cerebras WSE-3 (4 trillion transistors) via Infinity Fabric, targeting ultra-low latency and high throughput for AI inference, challenging traditional GPU clusters.
AMD launches world's first 2nm GPU MI455X and Zen 6 EPYC, targets NVIDIA and Intel
At Advancing AI 2026, AMD announced 46% data center CPU market share and launched the world's first 2nm GPU, Instinct MI455X, with CDNA architecture and HBM4 memory. The 6th-gen EPYC Venice (Zen 6) was also unveiled, targeting AI workloads.
NVIDIA Leads 25 Companies in Open Letter Against Restricting Open-Weight AI and Distillation
NVIDIA CEO Jensen Huang posted his first tweet on X, attaching an open letter signed by 25 companies including Microsoft, Meta, and IBM, urging Congress not to restrict open-weight AI models and distillation. The letter marks the formal split of the AI industry into open-weight and closed-source camps, with OpenAI, Anthropic, and Google notably absent, reshaping industry alliances and influencing global AI governance.
NVIDIA Prepay $1.5B to Amkor for US 2nm/3nm/HBM4 Advanced Packaging
NVIDIA prepays $1.5B to Amkor to expand advanced packaging capacity in Arizona, covering 2nm/3nm/HBM4. This onshores AI chip packaging, complementing Wistron's system integration, to reduce reliance on Taiwan and secure supply for Vera Rubin and Blackwell Ultra.
Microsoft and Databricks Expand Partnership: Full Azure Migration with Cobalt 200 ARM, Locking AI Agent Control Plane
Databricks will fully migrate to Azure, using Microsoft Cobalt 200 ARM chips for data and AI workloads, with deep integration of Genie and Unity AI Gateway into Microsoft products, locking in long-term partnership. Databricks raises $18.8B valuation.
Intel Accelerates 14A to 2027H2 Risk Production, 18A Yield Exceeds Target by 25%, Capex Raised to $20B
Intel announced 14A process acceleration to risk production in 2027H2 with HVM in 2028, 18A yield exceeding target by 25% supporting Panther Lake volume, and raising 2026 capex to $20B, signaling an aggressive foundry push. Advanced packaging EMIB-T becomes a profit pillar.
NVIDIA and Wistron Launch US-Based GB300 Production, Shifting AI Manufacturing Landscape
NVIDIA partnered with Wistron to open a $700M factory in Texas, achieving L6 integration of the GB300 Grace Blackwell Ultra system. The first US-made unit contains 1.5M parts, weighs 2 tons, and costs $4M. This milestone accelerates US AI capex localization from 5% to 30%, reshaping the global AI supply chain.
Tiered AI Chip Market Emerges as US Allows H200 Exports to China with 25% Levy
The US Commerce Department approved NVIDIA H200 exports to China with a 25% sales tax, while maintaining a ban on Blackwell. This formalizes a tiered AI chip market, making H200 the best available imported chip for China, but the performance gap and added tax burden increase deployment costs and complexity for Chinese AI infrastructure.
Cisco Proposes Logically Air-Gapped Model with eBPF, Shifting Security to Kernel
Cisco introduces a logically air-gapped governance model using eBPF and Cilium to create a software-defined cryptographic perimeter at the kernel level. Integrating Cisco Secure Workload with Isovalent, it aims to provide data residency and regulatory compliance for containerized, virtualized, and bare-metal environments without sacrificing cloud agility.
华为WAIC 2026首次展示全系列计算模组
...
NVIDIA SIGGRAPH 2026展示Vera Rubin与AI生产工具
...
AMD发布Helios机架级AI平台与MI455X GPU
...
OpenAI Launches Presence Agent Platform, Shifts Control Plane to Lock Enterprise AI Deployment
OpenAI launches Presence, an enterprise agent deployment platform integrating model inference, permissions, policies, evaluation, and escalation tools, shifting the control plane from models to the platform. ChatGPT Health is now fully available to US users 18+, integrating Apple Health, accelerating consumer AI agent adoption.