Filter

×
Active Filters Clear All
Keyword: Vera ×
318 Total Reports
2/16 Page
NVIDIA Other 2026-07-22

NVIDIA Reveals Vera Rubin GPU and Vera CPU: 3360B Transistors, 88-Core Olympus, 10x Agentic AI Efficiency

NVIDIA fully discloses Vera Rubin GPU and Vera CPU specifications. The GPU features 3360B transistors, HBM4 288GB, and 10x agentic AI efficiency over Blackwell. The CPU has 88 custom Olympus cores, delivering 2.2x faster agentic AI performance than Intel Sapphire Rapids. This solidifies NVIDIA's full-stack strategy against x86 incumbents.

Microsoft Other 2026-07-22

Microsoft and Mistral Partner to Build Sovereign AI Infrastructure for Regulated European Industries

Microsoft and Mistral expand their partnership with a multi-billion dollar deal. Mistral gains thousands of NVIDIA Vera Rubin GPUs and integrates its Medium 3.5 and OCR 4 models into Microsoft Foundry and Copilot Studio, offering cloud, connected, and offline deployment modes for European regulated industries under EU AI Act.

NVIDIA Other 2026-07-22

NVIDIA and Wistron Open US Factory for GB300 and Vera Rubin AI Superchips

Wistron opens its first US manufacturing facility in Fort Worth, producing NVIDIA GB300 Grace Blackwell Ultra and Vera Rubin superchips. The $700M plant aims for tens of thousands of boards monthly, marking NVIDIA's strategic shift to domestic AI hardware production.

NVIDIA Other 2026-07-21

NVIDIA Vera Rubin Platform Specs Revealed: 10x Tokens per Watt, Monolithic CPU+GPU Design

NVIDIA unveiled Vera Rubin platform specs with a monolithic design pairing 2 Rubin GPUs with 1 Vera CPU, flagship NVL72 integrating 36 CPUs and 72 GPUs. Claims 10x tokens per watt and 3x memory bandwidth over Grace Blackwell. Vera CPU sold standalone. First customers: Microsoft, OpenAI, Oracle. Mass production H2 2026. Performance claims await independent verification.

NVIDIA Other 2026-07-21

NVIDIA Spectrum-6 102.4Tbps Switch Goes Commercial, Cisco Adoption Confirms Bandwidth Inflection

NVIDIA announces Spectrum-6 102.4Tbps Ethernet switch for AI factories, doubling bandwidth with CPO and liquid cooling. Cisco confirms adoption in N9100 series, while Broadcom launches Tomahawk 6, signaling a terabit Ethernet race for AI infrastructure.

NVIDIA Other 2026-07-21

NVIDIA Open-Sources Cosmos 3 Edge 4B World Model for Real-Time Robot Control at 15Hz on Jetson Thor

NVIDIA open-sources Cosmos 3 Edge, a 4B parameter world action model for edge robotics. It achieves 15Hz real-time inference with 32 actions per inference on the Jetson Thor module. This extends NVIDIA's physical AI stack from training to real-time deployment, enabling end-to-end robot control at the edge.

NVIDIA Other 2026-07-20

NVIDIA Expands Agent Toolkit with Omniverse Libraries for Physical AI Simulation

At SIGGRAPH 2026, NVIDIA announced an expansion to its Agent Toolkit, adding Omniverse libraries that enable AI agents to build and simulate 3D worlds. The company also open-sourced Cosmos 3 Edge, a 4B-parameter world action model, completing its physical AI ecosystem from training to edge deployment.

NVIDIA Other 2026-07-20

NVIDIA Vera Rubin at BMS: Mission Control and BioNeMo Shift the AI Factory Control Plane

BMS deploys NVIDIA DGX SuperPOD with Vera CPU and Rubin GPU, managed by Mission Control and leveraging BioNeMo Agent Toolkit for drug discovery. This signals NVIDIA's shift from hardware vendor to AI factory control plane provider, potentially locking enterprises into its ecosystem.

Google Other 2026-07-20

Google's Frozen v2 Chip Hardwires Gemini Architecture for 6-10x TPU Efficiency, Set for 2028

Google is developing Frozen v2, a dedicated AI chip that hardwires the Gemini model architecture into silicon for 6-10x energy efficiency per token over current TPUs. It is a new product line, planned for 2028, with weight update flexibility but a frozen architecture. This validates the industry shift from general-purpose GPUs to dedicated ASICs for AI inference.

AMD Other 2026-07-20

Microsoft Azure Deploys AMD Helios Rack with MI455X GPUs, Breaking NVIDIA's Cloud AI Monopoly

Microsoft Azure officially adopts AMD Helios rack-scale AI infrastructure, featuring 72 MI455X GPUs (432GB HBM4, 19.6TB/s), Venice EPYC CPUs, and Pensando DPUs. Three new instances (ND MI455X v7, HDv2, HXv2) are launched, marking Azure's shift from exclusive NVIDIA dependency to a multi-vendor AI strategy.

Intel Other 2026-07-20

Intel Foundry 18A Yields Jump to 85%+; EMIB Packaging Hits 98%, Challenging TSMC N2

Intel Foundry 18A yields surged from 65% to 85%+ in a single quarter, approaching TSMC N2's 90%. EMIB advanced packaging yields reached 90-98%, turning a former bottleneck into a selling point. NVIDIA, AMD, Apple signed on but mostly as secondary suppliers.

Hewlett Packard Enterprise Other 2026-07-20

HPE Expands Private Cloud AI with NVIDIA Vera Rubin, Enabling Agent-Native AI Factory

HPE expands its Private Cloud AI line with NVIDIA Vera Rubin NVL72 and HGX Rubin NVL8, introducing Compute XD700, Cray GX240 blade with Vera CPU, and Quantum-X800 InfiniBand. New software includes Agent Toolkit and NemoClaw for agent-native AI, with Alletra Storage MP X10000 in Q4 2026.

NVIDIA Other 2026-07-19

NVIDIA Vera Rubin Platform and Dynamo 1.0 Disaggregate Inference, Shift Focus to Intelligence per Dollar

NVIDIA unveils Vera Rubin platform with a 7-chip stack (Vera CPU, Rubin GPU, NVLink 6, etc.) and Dynamo 1.0 inference disaggregation. A single NVL72 rack packs 72 GPUs/36 CPUs with 1.6 PB/s bandwidth, achieving up to 7x inference performance. The new 'intelligence per dollar' metric signals a shift from training to inference cost competition.

TSMC Other 2026-07-17

TSMC Pledges $100B More for 6 US Fabs, Localizing 3nm for AI Chip Supply Chain

TSMC announces an additional $100B investment in Arizona, bringing total US commitment to $265B, with plans for 6 fabs focused on 3nm and beyond. This move localizes advanced process for AI chip demand from NVIDIA, Apple, AMD, reshaping global semiconductor supply chain. Q2 net profit surged 77% YoY, FY capex raised to $60-64B.

NVIDIA Other 2026-07-16

Huang Denies Vera Rubin Delay; NVIDIA Defends AI Compute Throne

Jensen Huang officially denies rumors of a delay for the Vera Rubin platform, stating it is already in production and on track for mass deployment. This move aims to quell market anxiety over NVIDIA's product roadmap and solidify its leadership in AI training and inference chips.

NVIDIA Other 2026-07-16

NVIDIA Debuts T3000/T2000 Modules and Cosmos 3 Edge, Builds Sovereign AI Ecosystem in Japan

NVIDIA unveils T3000/T2000 compute modules (Thor architecture) and Cosmos 3 Edge world model, signs Japan Noetra alliance for 13,750 Vera CPUs + 27,500 Rubin GPUs (140MW). Sovereign AI revenue triples to $30B+ in FY2026, accelerating the physical AI ecosystem.

NVIDIA Other 2026-07-16

NVIDIA CUDA 13.3 Introduces clmad for Hardware-Accelerated Carryless Multiplication on GPUs

NVIDIA CUDA 13.3 adds the clmad hardware instruction for carryless multiply-accumulate on Ampere+ GPUs. GHASH throughput reaches 6.3 TB/s on B200, up to 18.8x faster than bitsliced. Sum-check protocol accelerates 3-13x. The instruction also benefits CRC, Reed-Solomon, and post-quantum cryptography.

Google Other 2026-07-15

Google Deeply Integrates Gemini Enterprise Telemetry with BigQuery for AI Governance

Google Cloud enables streaming Gemini Enterprise app telemetry (prompts, responses, activity logs) into BigQuery for real-time analysis. Leveraging BigQuery's AI capabilities (Conversational Analytics, auto-schema), it automates auditing, compliance, and insights for large-scale AI deployments, driving data-driven AI observability.

AMD Other 2026-07-15

AMD Confirms Zen 6 EPYC Venice: First 2nm Server CPU Launching July 2026

AMD confirms Zen 6 EPYC Venice launch at Advancing AI 2026 (July 22-23). As the first 2nm server CPU, it features triple-core hybrid architecture, up to 192 cores, ~29% single-thread and ~22% multi-thread gains, targeting AI inference and tight CPU-GPU synergy via Infinity Fabric.

NVIDIA Other 2026-07-14

NVIDIA's HVDC Power Shift Reshapes AI Data Center Energy Efficiency and Supply Chain

NVIDIA is driving a shift from AC to HVDC power systems for AI data centers, aiming to reduce conversion losses and improve efficiency. This move will reshape the entire supply chain for servers, power equipment, and cooling, but faces challenges in safety and standardization. It signals a generational change in AI infrastructure power delivery.