NVIDIA Vera CPU

Seagate Lyve

NVIDIA Vera CPU

Built for Scalable Accelerated Systems

The NVIDIA® Vera CPU is engineered to support agentic AI reasoning by managing data flow, memory access, and workflows across accelerated computing environments. Designed to integrate seamlessly with NVIDIA GPUs, Vera enhances AI system performance while also operating independently for analytics, cloud services, system orchestration, storage, and high-performance computing (HPC) workloads. Featuring high-performance, energy-efficient cores, extensive low-power memory bandwidth, and predictable latency, Vera helps maximize GPU utilization while delivering up to 2× the performance of the previous generation with exceptional energy efficiency.

NVIDIA Vera Rubin NVL72

The Vera Rubin platform opens the next frontier of agentic AI with seven chips to scale the world’s AI factories - the NVIDIA Vera CPU, NVIDIA Rubin GPU, NVIDIA NVLink™ 6 Switch, NVIDIA ConnectX®-9 SuperNIC, NVIDIA BlueField®-4 DPU and NVIDIA Spectrum™-6 Ethernet switch, and the NVIDIA Groq 3 LPU. Designed to operate together as one incredible AI supercomputer, the chips power every phase of AI — from massive-scale pretraining, post-training and test-time scaling to real-time agentic inference.

Power Efficiency vs B300Perf vs B300
NVFP4 Inference3.6 EFLOPS3.6X5X
FP8 Training1.2 EFLOPS2.5X3.5X
LPDDR5X Capacity54TB2.5X
HBM Capacity20.7 TB1X
HBM4 Bandwidth1.6 PB/s1.7X2.5X
Scale-Up Bandwidth260 TB/s1.4X2X
NVL72 System Power187 kw 1.4X

NVIDIA Rubin GPU

The NVIDIA Rubin GPU is NVIDIA’s next‑generation data‑center GPU architecture, designed to dramatically scale AI training and inference while improving power efficiency. Rubin introduces the 3rd‑generation Transformer Engine with native NVFP4 support and hardware‑assisted adaptive sparsity, enabling higher throughput without sacrificing model accuracy.
Power Efficiency vs B300Perf vs B300
NVFP4 Inference50 PFLOPS2.9X5X
NVFP4 Training35 PFLOPS2.0X3.5X
HBM4 Bandwidth22TB/s2X2.8X
NVLink Bandwidth per GPU3.6 TB/s1.2x2X
NVL8 System Power24 kW1.7x
Rubin features 6th‑generation NVLink technology with 3.6 TB/s of scale‑up bandwidth per GPU, enabling fast all‑to‑all communication across large GPU clusters and rack‑scale systems such as NVL8 and NVL72. Together, these advances make Rubin a foundational building block for next‑generation AI factories—delivering higher performance, better efficiency, and smaller hardware footprints for training and inference at scale.

Quick Quote.

What can we quote you on today?

We monitor our inbox’s like a hawk. Send us your pricing needs and we’ll get back to you ASAP.

What to expect

Expect an email from a Hypertec Solutions Partner representative after the form is submitted to schedule a discovery call or with the information requested.

 

Quick Quote (USA-EN)

Name(Required)