Earlier quoted context omitted.
A 9700X is twice the performance of a 9900K and M5 Max is almost 3X the performance. The megahertz myth is a myth.
I replied to the sibling comment: I was making simplifying assumptions for two specific use cases and naively treated physical cores and clock rate as my variables.
Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
81–90 of 109 posts
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#82Earlier quoted context omitted.
Features like hardware FP8 support definitely make it apples-to-oranges.
But doesn't the Apple M series NPU support FP8, and as it's a monolithic die (except for the GPU in the M5 Pro and Max) it could be argued it has hardware FP8 support, no?
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#83Earlier quoted context omitted.
Wow impressive. What's the story with this?
It's a tech demonstrator for a company that turns models into custom silicon for fast inference. In this case llama3.1-8b https://taalas.com/products/
I’m guessing it’s some form of ASIC because I can’t imagine crafting the logic of Llama on silicon is a very quick or easy job. Not that doing it on an ASIC is a piece of cake either.
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#84Agentic AI CPU? No. It’s a CPU designed for an AI cluster. Their last CPU Grace was the same thing and no one called it agentic. Vera now just has more performance/more bandwidth. It’s cool, I’d like to have one of these clusters, but this is not new. It’s marketed as agentic AI because that’s fashionable in 2026.
They significantly lowered latency compared to EPYC/Xeon, which is critical for streaming agents (e.g. text/audio/video agents).
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#85Earlier quoted context omitted.
Your 9900k at 5ghz does work slower than a Ryzen 9800X3D at 5ghz. A lot slower (1700 single core geekbench vs 3300, and just about any benchmark will tell the same story). Clock speed alone doesn't mean anything.
From the newegg listing: >8 Cores and 16 processing threads, based on AMD "Zen 5" architecture which is the same thread geometry as my 9900K. My main concerns at the time were: 1. More cores for running large workloads on k8s since I had just upgraded to 128G RAM 2. More thread level parallelism for my C++ code Naively I thought that, ceteris paribus and assuming good L1 cache utilization, having more physical cores…
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#86Earlier quoted context omitted.
From the newegg listing: >8 Cores and 16 processing threads, based on AMD "Zen 5" architecture which is the same thread geometry as my 9900K. My main concerns at the time were: 1. More cores for running large workloads on k8s since I had just upgraded to 128G RAM 2. More thread level parallelism for my C++ code Naively I thought that, ceteris paribus and assuming good L1 cache utilization, having more physical cores…
The 9800X3D has wider everything . Decoder, execution ports, vectors, cache, memory bandwidth...
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#87Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#88Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#89It is a 88-core ARM v9 chip, for somewhat more detailed spec.
Hmm, the 128-core Ampere Altra CPU is already available, and in a case from System76. I wonder what else differentiates it. If they're going to build CPUs I wish they had used Risc-V instead. They are using it somewhat already.
The Nvidia CPUs are designed for a very specific use case. They are designed for high performance with less concern about cost control.
The newer AmpereOne CPUs use DDR5 with the AmpereOne M supporting even higher memory bandwidth. Even then, I doubt the AmpereOne CPUs will match the performance of the Nvidia Rubin CPUs. But the Ampere processors are available for general use. I am guessing that Nvidia is only going to sell the complete rack system and only to high-volume customers.
Re: Nvidia Launches Vera CPU, Purpose-Built for Agentic AI
#90Given the price of these systems the ridiculously expensive network cards isn't such a huge huge deal, but I can't help but wonder at the absurdly amazing bandwidth hanging off Vera, the amazing brags about "7x more bandwidth than pcie gen 6" (amazing), but then having to go to pcie to network to chat with anyone else. It might be 800Gbe but it's still so many hops, pcie is weighty. I keep expecting we see fabric gai…
Most of the big AI/HPC clusters these systems are aimed at aren’t running regular PCIe Ethernet between nodes, they’re usually wired up with InfiniBand fabrics (HDR/NDR now, XDR soon)