Nvidia's AI Super Legion is here! Hwang In-hoon says: Demand for AI computing power will skyrocket 100 times

Nvidia founder and CEO Huang Renxun told thousands of viewers at the company's annual AI developer conference on Tuesday that artificial intelligence is at a “critical inflection point.”
At GTC 2025, known as the “Super Bowl of the AI World,” Hwang In-hoon's keynote address focused on Nvidia's latest breakthroughs in the AI field and shared his predictions for the industry's development in the next few years.
This trend towards AI agents and inference AI means “a significant increase in the amount of computation required to train and infer these models.”
Hwang In-hoon's remarks are meant to emphasize that the AI industry still needs a large number of Nvidia GPUs.
However, in January of this year, AI startup DeepSeek revealed that they only used 2,000 slower Nvidia H800 chips to train high-performance basic AI models, while companies such as OpenAI usually require tens of thousands or more GPUs.
This news once raised concerns in the market, causing Nvidia's stock price to plummet, and the market value evaporated by nearly 600 billion US dollars within a day. Wall Street investors feared that GPU demand was overestimated.
However, Hwang In-hoon believes that the future development of AI agents and inference AI will bring greater demand. He predicted,In the future, the world's 1 billion knowledge workers will have 10 billion AI agents working together。
Strong growth in market demand is already being traced.
Hwang In-hoon revealed that in the year demand for Nvidia Hopper GPUs was at its peak, the company delivered 1.3 million chips to the four major cloud computing companies AWS, Microsoft, Google, and Oracle. In its first year of launch, the latest Blackwell-architecture GPUs have already shipped 3.6 million units.
Also, during the presentation, Hwang In-hoon showed an AI model duel — Meta's LLAMA open source model and DeepSeek's R1 inference model. The user asked the two models a wedding seating arrangement question: at the 7-seat table, make sure that the bride and groom's parents are not next to each other, and that other restrictions are met.
Llama quickly gave an answer, generating 439 tokens (each token is about 0.75 words), but the answers were wrong. Although R1 answered correctly, the calculation time was longer and 8,559 tokens were generated. Since users are charged per token, the calculation cost is also higher.
Hwang In-hoon said,Although optimization technology can improve AI computing efficiency and thereby reduce the consumption of computing resources, overall demand will continue to grow。
Blackwell Ultra: Scheduled to launch in the second half of 2025; Vera Rubin (RubinAI chip, named after famous astronomer Vera Rubin): Expected to be released by the end of 2026; Rubin Ultra: Expected to debut in 2027.
He said that in the past ten years, AI has evolved from initial perception and computer vision to generative AI, and now it is moving towards intelligent AI (Agentic AI) with inference capabilities.This means that AI can not only understand and generate content, but also has the ability to make autonomous reasoning and intelligent decisions。
The release of GTC 2025 marks Nvidia's entry into a new stage in the field of AI computing, and also indicates that the future development of AI will enter a higher dimension.
DGX Spark: Previously unveiled as “Project Digits,” it is a handheld supercomputer that can provide up to 1000 trillion operations (TOPS), and is specially designed for AI fine-tuning and inference. DGX Station: A desktop-level supercomputer with datacenter-level computing performance, suitable for high-intensity AI computing tasks.



