NVIDIA Rubin is an upcoming GPU microarchitecture (the generation after Blackwell), named after astrophysicist Vera Rubin. It was announced by Jensen Huang at Computex in Taipei in 2024 and is expected in the second half of 2026 (Q3).
The chip is to be manufactured on TSMC's 3 nm process and support HBM4 memory. The architecture delivers 50 sparse petaflops in FP4 — a significant increase over Blackwell's 20 petaflops. Rubin is paired with the Vera CPU (the Vera Rubin NVL144 system).

AI Accelerator · serves as: AI Inference, High-level compute.
Which group NVIDIA Rubin belongs to and how it is built
Compute Modules is a subcategory of hardware components that provide processing power for robotic systems. It encompasses onboard computers, single-board computers (SBCs), AI accelerators, embedded processors, GPU/NPU compute modules, and other units responsible for processing sensor data and executing control logic. These modules form the foundation of modern autonomous, humanoid, and perception-capable robots.
An AI Accelerator is a specialized hardware component designed for efficient execution of artificial intelligence computations, particularly neural network inference, computer vision processing, and sensor data analysis. In robotics, AI accelerators are used to run perception models, object recognition, image segmentation, planning, and other tasks that require high computational throughput under constrained power budgets. They may take the form of dedicated NPU, TPU, VPU, or GPU chips, or specialized embedded modules.