NVIDIA CEO Jensen Huang Leaves Everyone SPEECHLESS (CES Supercut)

Quick Overview

NVIDIA CEO Jensen Huang revealed the Vera Rubin platform, which is in full production, featuring the Vera CPU and Reuben GPU, delivering significantly higher performance than Blackwell through extreme co-design strategies, including new MVF FP4 tensor cores and revolutionary liquid-cooled MGX chassis architecture.

Key Points: NVIDIA is now in full production of the Vera Rubin system, driven by the necessity to advance computation annually to meet the 10x annual increase in AI model size and computation demand. The Vera CPU delivers two times the performance per watt of the world's most advanced CPUs in a power-constrained world, boasting 88 physical cores with 176 threads using spatial multi-threading. The Reuben GPU offers 5x Blackwell's floating-point performance while only having 1.6 times the transistor count of Blackwell, achieved through extreme co-design across the entire stack. A key innovation is the MVF FP4 tensor core, an entire processing unit that adaptively adjusts precision to maximize throughput without sacrificing necessary accuracy in transformer models. The MGX chassis assembly time reduced from two hours to five minutes, and the system is 80-100% liquid-cooled using water at 45°C, enabling significant data center power savings. NVIDIA introduced Spectrum X, AI Ethernet, and Bluefield 4 Data Processing Units (DPUs) to handle intense east-west and north-south traffic, with Bluefield 4 providing 150 terabytes of context memory backing store per rack. The system achieves a 10x throughput improvement over Hopper in factory throughput (revenue potential) and is projected to require only 1/4th the number of systems compared to Blackwell to train a 10 trillion parameter model in one month.

Context: NVIDIA CEO Jensen Huang presented a supercut of announcements, likely at CES, focusing on the escalating race for AI advancement, which demands exponential increases in computational power annually. The context is driven by the fact that AI models are increasing by a factor of 10 yearly, necessitating aggressive innovation beyond the slowing pace of Moore's Law, leading NVIDIA to implement 'extreme co-design' across their entire hardware and software stack.

Raw markdown version of this recap