Reports
AI-generated structured vendor updates
AMD Reports Inference at 60% of AI Workloads, Launches Embedded AI Chip, Secures Meta 6GW Deal
At AMD Advancing AI 2026, CEO Lisa Su reported inference now accounts for 60% of AI workloads. AMD launched Ryzen AI Embedded X100 with CPU/GPU/NPU for physical AI, and announced a 6GW multi-generational Instinct GPU agreement with Meta, plus powering the US Sovereign AI Factory with MI355X, EPYC, and Pensando.
AMD Unveils 6th Gen EPYC Venice, MI400 GPUs, Helios Rack for AI Inference
AMD launches 6th Gen EPYC Venice (2nm) and MI400 series GPUs, with MI455X claiming 34x token throughput improvement. The Helios rack solution integrates 72 MI455X GPUs with 18 EPYC CPUs via Pensando networking and ROCm software, offering 30% more inference tokens per dollar than competitors. Adopted by OpenAI, Meta, and others.
AMD发布第六代EPYC Venice处理器与Helios机架级AI解决方案
...
AMD Launches Helios Rack-Scale AI Platform with MI400 GPUs, Targeting Inference TCO
At Advancing AI 2026, AMD unveiled the Helios rackscale platform integrating 72 MI455X GPUs and 18 EPYC Venice CPUs per rack, delivering 2.9 exaflops FP4 inference and 31TB HBM4 memory. The MI430X offers 288 TFLOPS FP64 for HPC. AMD claims up to 30% more inference tokens per dollar vs. competitors.
AMD and Cerebras Unveil Disaggregated AI Inference with Wafer-Scale Engine
AMD and Cerebras launch a disaggregated AI inference solution combining the Helios Rackscale system (6th-gen EPYC Venice CPUs + up to 72 Instinct MI455X GPUs) with the Cerebras WSE-3 (4 trillion transistors) via Infinity Fabric, targeting ultra-low latency and high throughput for AI inference, challenging traditional GPU clusters.
AMD launches world's first 2nm GPU MI455X and Zen 6 EPYC, targets NVIDIA and Intel
At Advancing AI 2026, AMD announced 46% data center CPU market share and launched the world's first 2nm GPU, Instinct MI455X, with CDNA architecture and HBM4 memory. The 6th-gen EPYC Venice (Zen 6) was also unveiled, targeting AI workloads.
AMD发布Helios机架级AI平台与MI455X GPU
...
AMD Helios Rack Challenges NVIDIA NVLink with Open UALoE Interconnect
At Advancing AI 2026, AMD launched the Helios rack with 72 MI455X GPUs, 18 Venice EPYC CPUs, and Pensando networking, claiming 30% higher inference token/$ vs NVIDIA NVL72. It introduced UALoE open interconnect to break NVLink lock-in, partnering with Cerebras, Cisco, and major AI firms.
AMD Helios Full-Stack AI Server Deployed on Azure, UALoE Open Standard Challenges NVLink
AMD and Microsoft Azure announce large-scale deployment of Helios full-stack AI servers, featuring 72 MI455X GPUs, 18 EPYC Venice CPUs, and Pensando DPUs, with 31TB HBM4 and 1.7PB/s memory bandwidth. Two new Azure VM series target Agentic AI and semiconductor design, marking Microsoft's multi-vendor strategy and challenging NVIDIA's dominance.
AMD Invests $5B in Anthropic, Secures 2GW MI450 Deployment, Reshaping AI Compute Ecosystem
AMD and Anthropic announce a strategic partnership: Anthropic will deploy up to 2GW of AMD Instinct MI450 GPUs, with AMD investing up to $5B in Anthropic. They will collaborate on ROCm optimization and Claude workload tuning, marking AMD's transition from chip vendor to AI ecosystem investor and accelerating multi-sourcing in AI compute.
AMD Unveils Zen 6 Venice, MI455X, and Helios Rack-Level Design to Challenge NVIDIA
At Advancing AI 2026, AMD launched Zen 6 EPYC Venice (2nm, up to 256 cores) and MI455X (CDNA5, 432GB HBM4, 40 PFLOPS FP4), along with Helios rack reference design (2.9 exaFLOPS FP4 per rack), claiming a 1000x AI performance roadmap, with major commitments from Meta, OpenAI, and others.
Microsoft Azure Deploys AMD Helios Rack with MI455X GPUs, Breaking NVIDIA's Cloud AI Monopoly
Microsoft Azure officially adopts AMD Helios rack-scale AI infrastructure, featuring 72 MI455X GPUs (432GB HBM4, 19.6TB/s), Venice EPYC CPUs, and Pensando DPUs. Three new instances (ND MI455X v7, HDv2, HXv2) are launched, marking Azure's shift from exclusive NVIDIA dependency to a multi-vendor AI strategy.
AMD Confirms Zen 6 EPYC Venice: First 2nm Server CPU Launching July 2026
AMD confirms Zen 6 EPYC Venice launch at Advancing AI 2026 (July 22-23). As the first 2nm server CPU, it features triple-core hybrid architecture, up to 192 cores, ~29% single-thread and ~22% multi-thread gains, targeting AI inference and tight CPU-GPU synergy via Infinity Fabric.