NVIDIA Vera CPU Launched for Agentic AI Workloads

NVIDIA has introduced the Vera CPU, a processor built specifically for agentic AI and reinforcement learning workloads. According to NVIDIA, it delivers results with 50% faster performance and twice the efficiency compared to traditional rack-scale CPUs.
Technical Specifications
The Vera CPU features 88 custom NVIDIA-designed Olympus cores, each capable of running two tasks using NVIDIA Spatial Multithreading. It includes a high-bandwidth memory subsystem built on LPDDR5X memory and uses the second-generation NVIDIA Scalable Coherency Fabric for faster agentic responses under high utilization conditions.
System Configurations
- New Vera CPU rack integrates 256 liquid-cooled Vera CPUs
- Sustains more than 22,500 concurrent CPU environments running independently at full performance
- Built using NVIDIA MGX modular reference architecture
- Part of NVIDIA Vera Rubin NVL72 platform with NVIDIA GPUs connected via NVIDIA NVLink-C2C interconnect
- Provides 1.8 TB/s of coherent bandwidth (7x PCIe Gen 6 bandwidth)
- Also serves as host CPU for NVIDIA HGX Rubin NVL8 systems
- Systems integrate NVIDIA ConnectX SuperNIC cards and NVIDIA BlueField-4 DPUs
Adoption and Partners
Customers collaborating with NVIDIA to deploy Vera CPU include Alibaba, ByteDance, Meta, Oracle Cloud Infrastructure, CoreWeave, Lambda, Nebius, and Nscale. Manufacturing partners include Dell Technologies, HPE, Lenovo, Supermicro, ASUS, Compal, Foxconn, GIGABYTE, Pegatron, Quanta Cloud Technology (QCT), Wistron, and Wiwynn.
Target Workloads
Vera systems are designed for reinforcement learning, agentic inference, data processing, orchestration, storage management, cloud applications, and high-performance computing. Systems partners provide both dual and single-socket CPU server configurations.
According to Jensen Huang, NVIDIA's CEO, "The CPU is no longer simply supporting the model; it's driving it. With breakthrough performance and energy efficiency, Vera unlocks AI systems that think faster and scale further."
📖 Read the full source: HN AI Agents
👀 See Also

Longitudinal study finds AI productivity gains at 10%, not 10x
A longitudinal study tracking 40 companies from November 2024 through February 2026 found AI usage increased by 65% on average, but pull request throughput only increased by 9.97%. The data suggests coding was never the primary bottleneck in software development.

OpenClaw's New Release: A Simple Name Change or a Major Upgrade?
OpenClaw, previously known as ClawDBot, has undergone a transformation. Read on to find out whether this change is merely cosmetic or introduces new features and improved stability.

Apple Builds New AI Architecture on Google Gemini Foundation Models
Apple announced a major overhaul of Apple Intelligence, built on foundation models co-developed with Google using Gemini technology. The new architecture includes an orchestrator, on-device and server-side models, and multimodal capabilities.

Claude Platform on AWS Now GA: Managed Agents, Code Execution, and Full API Parity via IAM
Claude Platform on AWS brings native Claude API features (Managed Agents, code execution, skills) to AWS customers with IAM auth, CloudTrail logging, and commitment retirement.