Nvidia used its Hot Chips 2026 presentation to flesh out the architecture of Vera, its first CPU built around a completely custom core design (dubbed Olympus). Unlike the previous Grace CPU, which used a stock Arm core, Vera represents Nvidia's deepest foray into processor design and is explicitly aimed at the rapidly evolving 'agentic AI' market — autonomous software agents that perform complex, multi-step tasks.
Vera ships as a single monolithic 88-core SKU, a contrast with the chiplet-based designs typical of x86 competitors. It employs what Nvidia calls 'spatial multithreading,' a novel approach to parallel processing that the company says is optimized for the unpredictable, branching workloads characteristic of AI agents. The chip is paired with a new memory subsystem using LPDDR5X (marketed as SOCAMM2) that delivers 1.2 TB/s of bandwidth.
Nvidia provided concrete performance comparisons against AMD's 96-core EPYC 9655P (Turin). In a headless browser test designed to simulate how an AI agent would fetch and process web information, Vera was 24% faster. In Linux kernel compilation — a proxy for the software-building tasks agents often perform — Vera achieved a 22% improvement when compiling natively for AArch64 and 14% faster when cross-compiling for x86.
However, Nvidia acknowledged that benchmarking agentic AI remains a challenge, calling it 'the most complex computing workload in history.' Performance can vary dramatically depending on how an agent's workflow is optimized. For instance, agents can strip out GUI rendering, font loading, and media decoding to run a browsing workflow up to 4.5x faster than a human-driven process.
Notably, AMD executive Joe Macri was quoted as being 'very happy to see Nvidia's Vera performance results,' adding, 'I actually thought we were beating them by smaller numbers.' The comment suggests that while Nvidia holds a lead in these specific benchmarks, AMD's Venice architecture (due to compete with Vera) may be closer than Nvidia's marketing suggests.
Nvidia has previously disclosed that Vera's Olympus core features a 10-wide decode, a large branch prediction unit with neural prediction, and a massive L1 cache — design choices aimed at minimizing latency for the erratic instruction streams typical of AI agents. The company also confirmed that Vera will sample to customers in the first half of 2026, with volume shipments expected later that year.