Reports
AI-generated structured vendor updates
AMD Announces Breakthrough MLPerf Inference 6.0 Results, Showcasing Multinode Scaling and Multimodal Capabilities
AMD's MLPerf Inference 6.0 submission, powered by Instinct MI355X GPUs, surpassed 1 million tokens per second for the first time on models like Llama 2 70B and GPT-OSS-120B. The results highlight efficient multinode scaling, rapid enablement of new workloads (e.g., text-to-video model Wan-2.2-t2v), and reproducible performance across a broad partner ecosystem.
Cisco Discloses Memory Poisoning Attack Method in AI Coding Assistants
Cisco's security team discovered and validated a persistent memory poisoning attack method targeting AI coding assistants like Claude Code, demonstrating how tampering with MEMORY.md system files can persistently manipulate AI behavior. This vulnerability prompted Anthropic to remove user memory files' system prompt privileges in v2.1.50.
Cisco Achieves Financial Network Compliance Automation via Services as Code (SaC)
Cisco disclosed its use of Services as Code (SaC) to help Intesa Sanpaolo achieve DORA compliance, automating the management of 8,000 switch configurations through Infrastructure as Code, reducing implementation time by 70%. This case demonstrates the feasibility of large-scale automated network device configuration management.
Intel Demonstrates AI Performance with Xeon 6 and Arc Pro GPUs in MLPerf Inference
Intel showcased the performance of its Xeon 6 CPUs and Arc Pro B-Series GPUs in the MLPerf Inference v6.0 benchmarks, particularly in handling large language models (LLMs). The results indicate that a system with four Arc Pro B70 GPUs can process 120B parameter models, delivering up to 1.8x higher inference performance in multi-GPU setups.
Google Launches Gemini API Docs MCP & Agent Skills for AI Coding Agents
Google introduces Gemini API Docs MCP protocol and Agent Skills toolkit, enabling real-time access to updated API documentation and injecting best-practice patterns to resolve outdated code generation. Combined usage achieves 96.3% pass rate with 63% fewer tokens per correct answer.
Google Launches Gemini API Docs MCP and Agent Skills to Enhance Coding Agent Performance
Google introduced two new tools, Gemini API Docs MCP and Agent Skills, to address the issue of coding agents generating outdated code due to training data cutoff dates. MCP connects to current Gemini API documentation via the Model Context Protocol, ensuring access to the latest APIs and code, while Agent Skills provides best-practice guidance and resource links. Combined use achieves a 96.3% pass rate with 63% fewer tokens per correct answer.
Cisco Launches Open-Source AI Agent Security Solution DefenseClaw
Cisco released open-source security solution DefenseClaw with four protection engines for OpenClaw AI Agent, covering prompt inspection, tool detection, installation scanning and code review. The solution demonstrates defense against 11.9% identified threats including malicious skills and unsafe MCP servers through hands-on labs.
Cisco Proposes Unified AI Fabric Architecture for Training/Inference Traffic
Cisco introduces unified AI fabric architecture using N9000 switches to intelligently route both training and inference traffic, addressing resource inefficiencies in dual-fabric setups. The solution features silicon-level low latency, real-time telemetry and automated policy tuning, targeting neocloud providers' platform transformation.
Meta Elevates Product Privacy Review to AI-Driven Company-Wide Risk Review
Meta announced it is expanding its product Privacy Review program into a broader, AI-centric company-wide Risk Review. The program leverages AI to automate compliance workflows, identify risks earlier in product development, and enable continuous monitoring, aiming to make manual processes the fallback.
Meta Elevates AI-Powered Risk Review to Cross-Company Program
Meta transforms its product Privacy Review into an AI-centric cross-company Risk Review program, automating documentation pre-filling, proactive development-phase scanning, and continuous monitoring for earlier risk identification. The initiative combines AI scalability with human expertise to establish an automation-first compliance culture.
Cisco Open Sources DefenseClaw for AI Agent Security Governance
Cisco launched open-source DefenseClaw, providing three-layer security architecture for AI agents like OpenClaw: supply chain scanning, runtime inspection, and system boundary control. The solution integrates NVIDIA's OpenShell sandbox for end-to-end automated governance.
Cisco Deploys Enterprise-Grade Networking and Security Architecture in Humanitarian Response Scenario
Cisco's Crisis Response team deployed an industrial-grade wireless network solution for the first time at the Musenyi refugee camp in Burundi. The solution integrates enterprise technologies like Cisco Identity Services Engine, Secure Connect, and Meraki cloud management to establish reliable and secure connectivity in harsh environments with limited infrastructure. This demonstrates Cisco's capability to adapt and validate its mature enterprise networking and zero-trust security architecture for extreme edge scenarios.
NVIDIA Forms Nemotron Coalition to Advance Open Frontier Models
NVIDIA announced the Nemotron Coalition at GTC, a collaboration with model builders and AI labs like Mistral AI to advance open, frontier-level foundation models. The initiative aims to foster the open model ecosystem by sharing expertise, data, and compute, emphasizing a future where AI is powered by a system of both open and proprietary models.
NVIDIA Demonstrates AI Factories as Flexible Grid Assets for Peak Demand Management
NVIDIA, in collaboration with EPRI, National Grid, and Emerald AI, demonstrated how AI factories powered by Blackwell GPU clusters can dynamically adjust power consumption in response to grid signals. This allows them to act as 'shock absorbers' during peak demand while maintaining performance for high-priority AI workloads.
ARM Launches AGI CPU for Agentic AI Infrastructure Era
ARM introduces the Arm AGI CPU, its first silicon product, designed for agentic AI infrastructure on Neoverse. Optimized for massively parallel workloads, it supports 272 cores per blade in a 1OU design, delivering 8160 cores per rack and over 2x performance vs. x86 systems.
ARM Launches AGI CPU Silicon for AI Infrastructure Market
ARM introduced its first production AGI CPU silicon in March 2026, marking a strategic shift from IP licensing to full silicon solutions provider. Designed for next-gen AI infrastructure, this move may reshape the data center processor ecosystem.
Arm Neoverse Reshapes Control Layer in AI Infrastructure
ARM introduces Neoverse infrastructure CPU cores optimized for cloud, AI, and HPC workloads, adopted by NVIDIA, AWS, Microsoft, and Google for their AI platforms, delivering performance gains and energy efficiency. This architecture enables high-density AI workload deployment in cloud and edge environments with enhanced multi-tenant security.
NVIDIA Launches OpenShell, Establishing Runtime Sandbox for Secure Autonomous AI Agents
NVIDIA introduces OpenShell, an open-source project designed as a secure-by-design runtime for autonomous AI agents. It employs a "browser tab" model, isolating agent operations from policy enforcement at the system level to prevent policy overrides and data leaks. NVIDIA is collaborating with key security vendors to establish a unified policy layer for enterprise AI agents.
NVIDIA CEO Outlines Accelerated Computing Paradigm, Signaling AI Infrastructure Evolution
In an interview, NVIDIA CEO Jensen Huang systematically elaborated on accelerated computing as a fundamental shift in computer architecture. He emphasized the data center's transition from general-purpose CPUs to specialized acceleration platforms led by GPUs, and believes the future computing stack will be re-architected around accelerated computing.
NVIDIA Extends RTX AI Capabilities to Local Agentic AI, Accelerating Gemma 4 Inference
At GTC 2026, NVIDIA announced it is extending its RTX platform capabilities to the domain of local Agentic AI, aiming to accelerate the inference performance of open models like Gemma 4 on end-user devices. This move seeks to leverage local, real-time context to enhance the value of AI agents, driving innovation beyond the cloud.