On Sunday, Nvidia CEO Jensen Huang reached past Blackwell and revealed the corporate’s next-generation AI-accelerating GPU platform throughout his keynote at Computex 2024 in Taiwan. Huang additionally detailed plans for an annual tick-tock-style improve cycle of its AI acceleration platforms, mentioning an upcoming Blackwell Ultra chip slated for 2025 and a subsequent platform known as “Rubin” set for 2026.
Nvidia’s knowledge heart GPUs at present energy a big majority of cloud-based AI fashions, equivalent to ChatGPT, in each improvement (coaching) and deployment (inference) phases, and traders are maintaining a detailed watch on the corporate, with expectations to maintain that run going.
During the keynote, Huang appeared considerably hesitant to make the Rubin announcement, maybe cautious of invoking the so-called Osborne impact, whereby an organization’s untimely announcement of the subsequent iteration of a tech product eats into the present iteration’s gross sales. “This is the very first time that this subsequent click on as been made,” Huang mentioned, holding up his presentation distant simply earlier than the Rubin announcement. “And I’m undecided but whether or not I’m going to remorse this or not.”
Nvidia Keynote at Computex 2023.
The Rubin AI platform, anticipated in 2026, will use HBM4 (a brand new type of high-bandwidth reminiscence) and NVLink 6 Switch, working at 3,600GBps. Following that launch, Nvidia will launch a tick-tock iteration known as “Rubin Ultra.” While Huang didn’t present intensive specs for the upcoming merchandise, he promised value and vitality financial savings associated to the brand new chipsets.
During the keynote, Huang additionally launched a brand new ARM-based CPU known as “Vera,” which can be featured on a brand new accelerator board known as “Vera Rubin,” alongside one of the Rubin GPUs.
Much like Nvidia’s Grace Hopper structure, which mixes a “Grace” CPU and a “Hopper” GPU to pay tribute to the pioneering pc scientist of the identical title, Vera Rubin refers to Vera Florence Cooper Rubin (1928–2016), an American astronomer who made discoveries in the sphere of deep area astronomy. She is finest recognized for her pioneering work on galaxy rotation charges, which supplied robust proof for the existence of darkish matter.
A calculated danger
Nvidia’s reveal of Rubin shouldn’t be a shock in the sense that the majority large tech corporations are repeatedly engaged on follow-up merchandise properly in advance of launch, but it surely’s notable as a result of it comes simply three months after the corporate revealed Blackwell, which is barely out of the gate and not but extensively transport.
At the second, the corporate appears to be snug leapfrogging itself with new bulletins and catching up later; Nvidia simply introduced that its GH200 Grace Hopper “Superchip,” unveiled one 12 months in the past at Computex 2023, is now in full manufacturing.
With Nvidia inventory rising and the corporate possessing an estimated 70–95 p.c of the info heart GPU market share, the Rubin reveal is a calculated danger that appears to come back from a spot of confidence. That confidence may change into misplaced if a so-called “AI bubble” pops or if Nvidia misjudges the capabilities of its rivals. The announcement may stem from stress to proceed Nvidia’s astronomical progress in market cap with nonstop guarantees of bettering know-how.
Accordingly, Huang has been wanting to showcase the corporate’s plans to proceed pushing silicon fabrication tech to its limits and extensively broadcast that Nvidia plans to maintain releasing new AI chips at a gentle cadence.
“Our firm has a one-year rhythm. Our primary philosophy may be very easy: construct your complete knowledge heart scale, disaggregate and promote to you elements on a one-year rhythm, and we push the whole lot to know-how limits,” Huang mentioned throughout Sunday’s Computex keynote.
Despite Nvidia’s latest market efficiency, the corporate’s run could not proceed indefinitely. With ample cash pouring into the info heart AI area, Nvidia is not alone in growing accelerator chips. Competitors like AMD (with the Instinct collection) and Intel (with Gaudi 3) additionally need to win a slice of the info heart GPU market away from Nvidia’s present command of the AI-accelerator area. And OpenAI’s Sam Altman is attempting to encourage diversified manufacturing of GPU {hardware} that can energy the corporate’s subsequent technology of AI fashions in the years ahead.



