NVIDIA Vera: The First CPU Purpose-Built for AI Agents Is Shipping

NVIDIA's first custom CPU — 88 Olympus cores, 1.2TB/s memory bandwidth, up to 1.8x faster per-core agentic performance — ships at scale to AWS, Oracle, Anthropic, OpenAI, and SpaceXAI.

Thursday August 27, 2026 Source: NVIDIA
TL;DR — Quick Answer

NVIDIA's Vera CPU — its first custom processor — is shipping at scale, hand-delivered to AWS, Oracle Cloud Infrastructure, Anthropic, OpenAI, and SpaceXAI. With 88 custom Olympus cores, 1.2TB/s of memory bandwidth, and up to 1.8x faster per-core performance on agentic workloads, Vera is built on the insight that AI agents don't run on GPUs alone: every tool call, sandbox, and orchestration layer is CPU work.

Key Takeaways

NVIDIA Vera: The First CPU Purpose-Built for AI Agents Is Shipping — AI news article illustration

NVIDIA’s Vera CPU — its first custom processor — is shipping now, and the deliveries read like a tour of the AI industry’s biggest names.

A CPU for the Agent Era

The insight behind Vera: agentic AI puts pressure on CPUs in ways traditional designs never anticipated. An agent’s sandbox, its tool calls, the orchestration layers, long-context state management — that’s all CPU work, concurrent and real-time. Vera’s answer:

As NVIDIA VP Ian Buck put it: “AI agents don’t run on GPUs alone… Agentic AI is creating a new CPU moment in the AI factory — as models move from answering to acting, Vera is purpose-built to keep that work moving at scale.”

Hand-Delivered Across the Ecosystem

The delivery tour is its own story: Buck personally handed Vera systems to Oracle Cloud Infrastructure (first cloud deploying at hyperscale — hundreds of thousands of Vera CPUs planned), Anthropic (head of compute James Bradbury), OpenAI (compute infrastructure head Sachin Katti), SpaceXAI (evaluating Vera for RL workloads and agent simulation, with Musk quizzing the team on cores, memory layout, and cooling), and most recently AWS — alongside an AWS-NVIDIA expansion that adds 2 million more NVIDIA GPUs including Vera Rubin GPU infrastructure.

Vera also serves as the host processor in Vera Rubin NVL72, paired with Rubin GPUs via NVLink-C2C in a unified memory architecture.

Read the full story at blogs.nvidia.com/blog/vera-cpu-delivery.


Frequently Asked Questions

What is the NVIDIA Vera CPU?

Vera is NVIDIA's first custom CPU, purpose-built for the age of agentic AI. It has 88 custom NVIDIA-designed Olympus cores, 1.2TB/s of memory bandwidth, and delivers up to 1.8x faster per-core performance on agentic AI workloads.

Why do AI agents need a special CPU?

AI agents don't run on GPUs alone. Every agent sandbox, tool call, orchestration layer, and long-context retrieval operation is CPU work — concurrent, real-time tasks that traditional core-density-focused CPUs weren't built to prioritize.

Who is deploying Vera?

As of August 2026, Vera systems have been delivered to AWS, Oracle Cloud Infrastructure (first cloud to deploy at hyperscale, with hundreds of thousands planned), and directly to AI labs Anthropic, OpenAI, and SpaceXAI.

What is Vera Rubin NVL72?

Vera is also the host processor for the Vera Rubin NVL72 rack, where it pairs via second-generation NVLink-C2C with two Rubin GPUs, sharing a unified memory architecture at 2x the energy efficiency of traditional infrastructure.

This article is based on the official announcement from NVIDIA . Read the original for full technical details.

Related Articles

Back to all news