Sony Pixel Power calrec Sony

We Created a Processor for the Generative AI Era,' NVIDIA CEO Says

18/03/2024

Generative AI promises to revolutionize every industry it touches - all that's been needed is the technology to meet the challenge.

NVIDIA founder and CEO Jensen Huang on Monday introduced that technology - the company's new Blackwell computing platform - as he outlined the major advances that increased computing power can deliver for everything from software to services, robotics to medical technology and more.

Accelerated computing has reached the tipping point - general purpose computing has run out of steam, Huang told more than 11,000 GTC attendees gathered in-person - and many tens of thousands more online - for his keynote address at Silicon Valley's cavernous SAP Center arena.

We need another way of doing computing - so that we can continue to scale so that we can continue to drive down the cost of computing, so that we can continue to consume more and more computing while being sustainable. Accelerated computing is a dramatic speedup over general-purpose computing, in every single industry.

Huang spoke in front of massive images on a 40-foot tall, 8K screen the size of a tennis court to a crowd packed with CEOs and developers, AI enthusiasts and entrepreneurs, who walked together 20 minutes to the arena from the San Jose Convention Center on a dazzling spring day.

Delivering a massive upgrade to the world's AI infrastructure, Huang introduced the NVIDIA Blackwell platform to unleash real-time generative AI on trillion-parameter large language models.

Huang presented NVIDIA NIM - a reference to NVIDIA inference microservices - a new way of packaging and delivering software that connects developers with hundreds of millions of GPUs to deploy custom AI of all kinds.

And bringing AI into the physical world, Huang introduced Omniverse Cloud APIs to deliver advanced simulation capabilities.

Huang punctuated these major announcements with powerful demos, partnerships with some of the world's largest enterprises and more than a score of announcements detailing his vision.

GTC - which in 15 years has grown from the confines of a local hotel ballroom to the world's most important AI conference - is returning to a physical event for the first time in five years.

This year's has over 900 sessions - including a panel discussion on transformers moderated by Huang with the eight pioneers who first developed the technology, more than 300 exhibits and 20-plus technical workshops.

It's an event that's at the intersection of AI and just about everything. In a stunning opening act to the keynote, Refik Anadol, the world's leading AI artist, showed a massive real-time AI data sculpture with wave-like swirls in greens, blues, yellows and reds, crashing, twisting and unraveling across the screen.

As he kicked off his talk, Huang explained that the rise of multi-modal AI - able to process diverse data types handled by different models - gives AI greater adaptability and power. By increasing their parameters, these models can handle more complex analyses.

But this also means a significant rise in the need for computing power. And as these collaborative, multi-modal systems become more intricate - with as many as a trillion parameters - the demand for advanced computing infrastructure intensifies.

We need even larger models, Huang said. We're going to train it with multimodality data, not just text on the internet, we're going to train it on texts and images, graphs and charts, and just as we learned watching TV, there's going to be a whole bunch of watching video.

The Next Generation of Accelerated Computing In short, Huang said we need bigger GPUs. The Blackwell platform is built to meet this challenge. Huang pulled a Blackwell chip out of his pocket and held it up side-by-side with a Hopper chip, which it dwarfed.

Named for David Harold Blackwell - a University of California, Berkeley mathematician specializing in game theory and statistics, and the first Black scholar inducted into the National Academy of Sciences - the new architecture succeeds the NVIDIA Hopper architecture, launched two years ago.

Blackwell delivers 2.5x its predecessor's performance in FP8 for training, per chip, and 5x with FP4 for inference. It features a fifth-generation NVLink interconnect that's twice as fast as Hopper and scales up to 576 GPUs.

And the NVIDIA GB200 Grace Blackwell Superchip connects two Blackwell NVIDIA B200 Tensor Core GPUs to the NVIDIA Grace CPU over a 900GB/s ultra-low-power NVLink chip-to-chip interconnect.

Huang held up a board with the system. This computer is the first of its kind where this much computing fits into this small of a space, Huang said. Since this is memory coherent, they feel like it's one big happy family working on one application together.

For the highest AI performance, GB200-powered systems can be connected with the NVIDIA Quantum-X800 InfiniBand and Spectrum-X800 Ethernet platforms, also announced today, which deliver advanced networking at speeds up to 800Gb/s.

The amount of energy we save, the amount of networking bandwidth we save, the amount of wasted time we save, will be tremendous, Huang said. The future is generative which is why this is a brand new industry. The way we compute is fundamentally different. We created a processor for the generative AI era.

To scale up Blackwell, NVIDIA built a new chip called NVLink Switch. Each can connect four NVLink interconnects at 1.8 terabytes per second and eliminate traffic by doing in-network reduction.

NVIDIA Switch and GB200 are key components of what Huang described as one giant GPU, the NVIDIA GB200 NVL72, a multi-node, liquid-cooled, rack-scale system that harnesses Blackwell to offer supercharged compute for trillion-parameter models, with 720 petaflops of AI training performance and 1.4 exaflops of AI inference perform
LINK: https://blogs.nvidia.com/blog/2024-gtc-keynote/...
See more stories from nvidia

More from Nvidia

25/04/2024

AI Drives Future of Transportation at Asia's Largest Automotive Show

The latest trends and technologies in the automotive industry are in the spotlight at the Beijing International Automotive Exhibition, aka Auto China, which ope...

25/04/2024

Into the Omniverse: Unlocking the Future of Manufacturing With OpenUSD on Siemens Teamcenter X

Editor's note: This post is part of Into the Omniverse, a series focused on ...

25/04/2024

Blast From the Past: Stream StarCraft' and Diablo' on GeForce NOW

Support for Battle.net on GeForce NOW expands this GFN Thursday, as titles from the iconic StarCraft and Diablo series come to the cloud. StarCraft Remastered,...

24/04/2024

Rays Up: Decoding AI-Powered DLSS 3.5 Ray Reconstruction

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

24/04/2024

Forecasting the Future: AI2's Christopher Bretherton Discusses Using Machine Learning for Climate Modeling

Can machine learning help predict extreme weather events and climate change? Chr...

24/04/2024

NVIDIA to Acquire GPU Orchestration Software Provider Run:ai

To help customers make more efficient use of their AI computing resources, NVIDIA today announced it has entered into a definitive agreement to acquire Run:ai, ...

24/04/2024

How Virtual Factories Are Making Industrial Digitalization a Reality

To address the shift to electric vehicles, increased semiconductor demand, manufacturing onshoring, and ambitions for greater sustainability, manufacturers are ...

23/04/2024

Small and Mighty: NVIDIA Accelerates Microsoft's Open Phi-3 Mini Language Models

NVIDIA announced today its acceleration of Microsoft's new Phi-3 Mini open l...

22/04/2024

Climate Tech Startups Integrate NVIDIA AI for Sustainability Applications

Whether they're monitoring miniscule insects or delivering insights from satellites in space, NVIDIA-accelerated startups are making every day Earth Day. S...

18/04/2024

Wide Open: NVIDIA Accelerates Inference on Meta Llama 3

NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM). The open mod...

18/04/2024

Up to No Good: No Rest for the Wicked' Early Access Launches on GeForce NOW

It's time to get a little wicked. Members can now stream No Rest for the Wicked from the cloud. It leads six new games joining the GeForce NOW library of m...

18/04/2024

NVIDIA Honors Partners of the Year in Europe, Middle East, Africa

NVIDIA today recognized 18 partners in Europe, the Middle East and Africa for their achievements and commitment to driving AI adoption. The recipients were hon...

17/04/2024

Seeing Beyond: Living Optics CEO Robin Wang on Democratizing Hyperspectral Imaging

Step into the realm of the unseen with Robin Wang, CEO of Living Optics. The sta...

17/04/2024

Moving Pictures: Transform Images Into 3D Scenes With NVIDIA Instant NeRF

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

16/04/2024

New NVIDIA RTX A400 and A1000 GPUs Enhance AI-Powered Design and Productivity Workflows

AI integration across design and productivity applications is becoming the new s...

16/04/2024

To Cut a Long Story Short: Video Editors Benefit From DaVinci Resolve's New AI Features Powered by RTX

Editor's note: This post is part of our In the NVIDIA Studio series, which c...

15/04/2024

AI Is Tech's Greatest Contribution to Social Elevation,' NVIDIA CEO Tells Oregon State Students

AI promises to bring the full benefits of the digital revolution to billions acr...

10/04/2024

The Building Blocks of AI: Decoding the Role and Significance of Foundation Models

Editor's note: This post is part of the AI Decoded series, which demystifies...

10/04/2024

Combating Corruption With Data: Cleanlab and Berkeley Research Group on Using AI-Powered Investigative Analytics

Talk about scrubbing data. Curtis Northcutt, cofounder and CEO of Cleanlab, and ...

09/04/2024

NVIDIA Joins $110 Million Partnership to Help Universities Teach AI Skills

The Biden Administration has announced a new $110 million AI partnership between Japan and the United States that includes an initiative to fund research throug...

09/04/2024

Broadcasting Breakthroughs: NVIDIA Holoscan for Media, Available Now, Transforms Live Media With Easy AI Integration

Whether delivering live sports programming, streaming services, network broadcas...

09/04/2024

Start Up Your Engines: NVIDIA and Google Cloud Collaborate to Accelerate AI Development

NVIDIA and Google Cloud have announced a new collaboration to help startups arou...

04/04/2024

NVIDIA Ranked by Fortune at No. 3 on 100 Best Companies to Work For' List

NVIDIA jumped to No. 3 on the latest list of America's 100 Best Companies to Work For by Fortune magazine and Great Place to Work. It's the company'...

04/04/2024

The Elder Scrolls Online' Joins GeForce NOW for Game's 10th Anniversary

Rain or shine, a new month means new games. GeForce NOW kicks off April with nearly 20 new games, seven of which are available to play this week. GFN Thursday ...

03/04/2024

A New Lens: Dotlumen CEO Cornel Amariei on Assistive Technology for the Visually Impaired

Dotlumen is illuminating a new technology to help people with visual impairments...

03/04/2024

Coming Up ACEs: Decoding the AI Technology That's Enhancing Games With Realistic Digital Humans

Editor's note: This post is part of the AI Decoded series, which demystifies...

28/03/2024

Greater Scope: Doctors Get Inside Look at Gut Health With AI-Powered Endoscopy

From humble beginnings as a university spinoff to an acquisition by the leading global medtech company in its field, Odin Vision has been on an accelerated jour...

28/03/2024

Get Cozy With Palia' on GeForce NOW

Ease into spring with the warm, cozy vibes of Palia, coming to the cloud this GFN Thursday. It's part of six new titles joining the GeForce NOW library of ...

27/03/2024

Software Developers Launch OpenUSD and Generative AI-Powered Product Configurators Built on NVIDIA Omniverse

From designing dream cars to customizing clothing, 3D product configurators are ...

27/03/2024

NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf

It's official: NVIDIA delivered the world's fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks...

27/03/2024

Viome's Guru Banavar Discusses AI for Personalized Health

In the latest episode of NVIDIA's AI Podcast, Viome Chief Technology Officer Guru Banavar spoke with host Noah Kravitz about how AI and RNA sequencing are r...

27/03/2024

Unlocking Peak Generations: TensorRT Accelerates AI on RTX PCs and Workstations

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

26/03/2024

Boom in AI-Enabled Medical Devices Transforms Healthcare

The future of healthcare is software-defined and AI-enabled. Around 700 FDA-cleared, AI-enabled medical devices are now on the market - more than 10x the number...

26/03/2024

Model Innovators: How Digital Twins Are Making Industries More Efficient

A manufacturing plant near Hsinchu, Taiwan's Silicon Valley, is among facilities worldwide boosting energy efficiency with AI-enabled digital twins. A virt...

26/03/2024

Into the Omniverse: Groundbreaking OpenUSD Advancements Put NVIDIA GTC Spotlight on Developers

Editor's note: This post is part of Into the Omniverse, a series focused on ...

25/03/2024

NVIDIA Blackwell and Automotive Industry Innovators Dazzle at NVIDIA GTC

Generative AI, in the data center and in the car, is making vehicle experiences safer and more enjoyable. The latest advancements in automotive technology were...

21/03/2024

AI's New Frontier: From Daydreams to Digital Deeds

Imagine a world where you can whisper your digital wishes into your device, and poof, it happens. That world may be coming sooner than you think. But if you...

21/03/2024

You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper

Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by...

21/03/2024

Instant Latte: NVIDIA Gen AI Research Brews 3D Shapes in Under a Second

NVIDIA researchers have pumped a double shot of acceleration into their latest text-to-3D generative AI model, dubbed LATTE3D. Like a virtual 3D printer, LATTE...

21/03/2024

Here Be Dragons: Dragon's Dogma 2' Comes to GeForce NOW

Arise for a new adventure with Dragon's Dogma 2, leading two new titles joining the GeForce NOW library this week. Set Forth, Arisen Fulfill a forgotten de...

20/03/2024

AI Decoded From GTC: The Latest Developer Tools and Apps Accelerating AI on PC and Workstation

Editor's note: This post is part of the AI Decoded series, which demystifies...

19/03/2024

NVIDIA Celebrates Americas Partners Driving AI-Powered Transformation

NVIDIA recognized 14 partners in the Americas for their achievements in transforming businesses with AI, this week at GTC. The winners of the NVIDIA Partner Ne...

19/03/2024

Climate Pioneers: 3 Startups Harnessing NVIDIA's AI and Earth-2 Platforms

To help mitigate climate change - one of humanity's greatest challenges - researchers are turning to AI and sustainable computing to accelerate and operatio...

19/03/2024

Secure by Design: NVIDIA AIOps Partner Ecosystem Blends AI for Businesses

In today's complex business environments, IT teams face a constant flow of challenges, from simple issues like employee account lockouts to critical securit...

19/03/2024

Generation Sensation: New Generative AI and RTX Tools Boost Content Creation

Editor's note: This post is part of our In the NVIDIA Studio series, which celebrates featured artists, offers creative tips and tricks, and demonstrates ho...

19/03/2024

NVIDIA, Huang Win Top Honors in Innovation, Engineering

NVIDIA today was named the world's most innovative company by Fast Company magazine. The accolade comes on the heels of company founder and CEO Jensen Huan...

18/03/2024

NVIDIA Edify Unlocks 3D Generative AI, New Image Controls for Visual Content Providers

NVIDIA Edify, a multimodal architecture for visual generative AI, is entering a ...

18/03/2024

From Atoms to Supercomputers: NVIDIA, Partners Scale Quantum Computing

The latest advances in quantum computing include investigating molecules, deploying giant supercomputers and building the quantum workforce with a new academic ...

18/03/2024

New NVIDIA Storage Partner Validation Program Streamlines Enterprise AI Deployments

A sharp increase in generative AI deployments is driving business innovation for...

18/03/2024

NVIDIA Unveils Digital Blueprint for Building Next-Gen Data Centers

Designing, simulating and bringing up modern data centers is incredibly complex, involving multiple considerations like performance, energy efficiency and scala...