Sony Pixel Power calrec Sony

Why GPUs Are Great for AI

04/12/2023

GPUs have been called the rare Earth metals - even the gold - of artificial intelligence, because they're foundational for today's generative AI era.

Three technical reasons, and many stories, explain why that's so. Each reason has multiple facets well worth exploring, but at a high level:

GPUs employ parallel processing.

GPU systems scale up to supercomputing heights.

The GPU software stack for AI is broad and deep.

The net result is GPUs perform technical calculations faster and with greater energy efficiency than CPUs. That means they deliver leading performance for AI training and inference as well as gains across a wide array of applications that use accelerated computing.

In its recent report on AI, Stanford's Human-Centered AI group provided some context. GPU performance has increased roughly 7,000 times since 2003 and price per performance is 5,600 times greater, it reported.

A 2023 report captured the steep rise in GPU performance and price/performance. The report also cited analysis from Epoch, an independent research group that measures and forecasts AI advances.

GPUs are the dominant computing platform for accelerating machine learning workloads, and most (if not all) of the biggest models over the last five years have been trained on GPUs [they have] thereby centrally contributed to the recent progress in AI, Epoch said on its site.

A 2020 study assessing AI technology for the U.S. government drew similar conclusions.

We expect [leading-edge] AI chips are one to three orders of magnitude more cost-effective than leading-node CPUs when counting production and operating costs, it said.

NVIDIA GPUs have increased performance on AI inference 1,000x in the last ten years, said Bill Dally, the company's chief scientist in a keynote at Hot Chips, an annual gathering of semiconductor and systems engineers.

ChatGPT Spread the News ChatGPT provided a powerful example of how GPUs are great for AI. The large language model (LLM), trained and run on thousands of NVIDIA GPUs, runs generative AI services used by more than 100 million people.

Since its 2018 launch, MLPerf, the industry-standard benchmark for AI, has provided numbers that detail the leading performance of NVIDIA GPUs on both AI training and inference.

For example, NVIDIA Grace Hopper Superchips swept the latest round of inference tests. NVIDIA TensorRT-LLM, inference software released since that test, delivers up to an 8x boost in performance and more than a 5x reduction in energy use and total cost of ownership. Indeed, NVIDIA GPUs have won every round of MLPerf training and inference tests since the benchmark was released in 2019.

In February, NVIDIA GPUs delivered leading results for inference, serving up thousands of inferences per second on the most demanding models in the STAC-ML Markets benchmark, a key technology performance gauge for the financial services industry.

A RedHat software engineering team put it succinctly in a blog: GPUs have become the foundation of artificial intelligence.

AI Under the Hood A brief look under the hood shows why GPUs and AI make a powerful pairing.

An AI model, also called a neural network, is essentially a mathematical lasagna, made from layer upon layer of linear algebra equations. Each equation represents the likelihood that one piece of data is related to another.

For their part, GPUs pack thousands of cores, tiny calculators working in parallel to slice through the math that makes up an AI model. This, at a high level, is how AI computing works.

Highly Tuned Tensor Cores Over time, NVIDIA's engineers have tuned GPU cores to the evolving needs of AI models. The latest GPUs include Tensor Cores that are 60x more powerful than the first-generation designs for processing the matrix math neural networks use.

In addition, NVIDIA Hopper Tensor Core GPUs include a Transformer Engine that can automatically adjust to the optimal precision needed to process transformer models, the class of neural networks that spawned generative AI.

Along the way, each GPU generation has packed more memory and optimized techniques to store an entire AI model in a single GPU or set of GPUs.

Models Grow, Systems Expand The complexity of AI models is expanding a whopping 10x a year.

The current state-of-the-art LLM, GPT4, packs more than a trillion parameters, a metric of its mathematical density. That's up from less than 100 million parameters for a popular LLM in 2018.

In a recent talk at Hot Chips, NVIDIA Chief Scientist Bill Dally described how single-GPU performance on AI inference expanded 1,000x in the last decade. GPU systems have kept pace by ganging up on the challenge. They scale up to supercomputers, thanks to their fast NVLink interconnects and NVIDIA Quantum InfiniBand networks.

For example, the DGX GH200, a large-memory AI supercomputer, combines up to 256 NVIDIA GH200 Grace Hopper Superchips into a single data-center-sized GPU with 144 terabytes of shared memory.

Each GH200 superchip is a single server with 72 Arm Neoverse CPU cores and four petaflops of AI performance. A new four-way Grace Hopper systems configuration puts in a single compute node a whopping 288 Arm cores and 16 petaflops of AI performance with up to 2.3 terabytes of high-speed memory.

And NVIDIA H200 Tensor Core GPUs announced in November pack up to 288 gigabytes of the latest HBM3e memory technology.

Software Covers the Waterfront An expanding ocean of GPU software has evolved since 2007 to enable every facet of AI, from deep-tech features to high-level applications.

The NVIDIA AI platform includes hundreds of software libraries and apps. The CUDA programming language and the cuDNN-X library for deep learning provide a base on top of which developers have created software like NVIDIA NeMo, a framework to let users build, customize and run inference on their o
LINK: https://blogs.nvidia.com/blog/why-gpus-are-great-for-ai/...
See more stories from nvidia

North America Stories

20/06/2026

What's Next for Apogee? Start Here.

What exactly is Apogee Control V3? Control V3 is a new mixer application that controls Apogee interfaces. The new hit feature is that V3 finally allows for...

19/06/2026

NBC Sports U.S. Open Coverage Fires Up 92 Cameras, Bunker cams

Split compound eases operational challenges at Shinnecock Hills Golf Club...

19/06/2026

ESPN's Men's College World Series Production Adds Onsite Studio, POVORA CapCams, Expanded Drone Coverage for Finale in Omaha

North Carolina, Oklahoma meet in the best-of-three Finals as ESPN leans into spe...

19/06/2026

Eurovision secures top four position as content distributor rankings hold steady in Poland

Data from May shows seasonal outdoor trends triggers lower viewing Warsaw, Pola...

19/06/2026

Bitfocus Buttons wins another top industry award

Buttons is best control system in the rAVe Best of Infocomm Awards 2026...

19/06/2026

Mavis Studio Makes iPad Production More Powerful

Mavis Studio Makes iPad Production More Powerful Brie Clayton June 19, 2026 0 Comments InfoComm update brings new NDI Preview, PTZ control, USB audio ...

19/06/2026

Immersive Studio Metaverse Stage Tackles Post with Blackmagic Design

Immersive Studio Metaverse Stage Tackles Post with Blackmagic Design Brie Clayton June 19, 2026 0 Comments New narrative projects rely on DaVinci Reso...

19/06/2026

How to Run the Original 1993 After Effects

How to Run the Original 1993 After Effects Graham Quince June 19, 2026 0 Comments How to the original After Effects v1 in an emulator, and you don'...

19/06/2026

IBC Show to Increase Focus on Networking, Startups

Share Copy link Facebook X Linkedin Bluesky Email...

19/06/2026

Irdeto Taps Axel Gallant as CEO

Share Copy link Facebook X Linkedin Bluesky Email...

19/06/2026

SMPTE Makes Its Standards Freely Accessible - Opening St...

SMPTE , the home of media professionals, technologists and engineers, has announced that its entire Standards catalog is now freely available to the global medi...

19/06/2026

nsign and BrightSign partner to expand deployment options...

nsign, the digital signage SaaS platform built around its core principle of Simplify Complexity, has announced a partnership with BrightSign , expanding the dep...

19/06/2026

Visual Productions Unveils RdmRelay2 Four-channel Relay...

Visual Productions announces the availability of its new RdmRelay2 at InfoComm 2026 (ACT Entertainment, Booth N6813). A networked, four-channel DMX relay, it is...

19/06/2026

Adobe Unveils Major Expansion of Creative Agent Across Firefly and Creative Cloud Apps Including Photoshop and Premiere

Adobe Unveils Major Expansion of Creative Agent Across Firefly and Creative Clou...

19/06/2026

Immersive Studio Metaverse Stage Innovates Storytelling with URSA Cine Immersive

Immersive Studio Metaverse Stage Innovates Storytelling with URSA Cine Immersive Brie Clayton June 18, 2026 0 Comments Two new narrative short films c...

19/06/2026

June 18, 2026

Lab studies explain how new cancer drug works as it enters patient testing Immunologists at Scripps Research show how a new, experimental drug revives immune ce...

18/06/2026

Ratings Roundup: Knicks-Spurs NBA Finals Is Most Watched Since Jordan Bulls Era; FIFA World Cup Opens Big

Ratings Roundup is a rundown of recent rating news and is derived from press rel...

18/06/2026

InfoComm 2026: PTZOptics Debuts Healthcare Visual Reasoning Integration with LayerJot

PTZOptics has unveiled new Visual Reasoning demonstrations at InfoComm 2026 (Boo...

18/06/2026

IBC 2026: Announces New Future Tech Ignite Initiative, Expanded Exhibitor Participation

IBC2026 will take place at the RAI Amsterdam from September 11-14, bringing toge...

18/06/2026

InfoComm 2026 Holds Inaugural Media Day with Product Announcements from 19 Exhibitors

InfoComm 2026 held its first-ever Media Day on June 17, providing journalists an...

18/06/2026

Info Comm 2026: FOR-A America Announces TAA-Compliant LED Display Package

FOR-A America has announced a Trade Agreements Act (TAA)-compliant LED display solution combining Alfalite's Litepix LED displays and Brompton Technology...

18/06/2026

SVG GameDay, Ep. 20: Cleveland Browns Kyle Millen - Run of Shows & Rock n Roll

In-venue and creative video staffers at the professional and collegiate level have one major thing in common: the intensity and attention to detail ramps up dur...

18/06/2026

International Federation of American Football, TMRW Sports Partner on Global Growth of Flag Football

The International Federation of American Football (IFAF) and TMRW Sports have an...

18/06/2026

AJA Debuts Io Xpand Thunderbolt 5 Expansion Chassis for KONA and Corvid PCIe I/O Cards

AJA Video Systems has unveiled Io Xpand, a Thunderbolt 5-enabled PCIe expansion ...

18/06/2026

ESPN Marks 30th Anniversary of WNBA's Inaugural Game With Liberty-Sparks Broadcast

ESPN has announced its coverage plans for the 30th anniversary of the WNBA's...

18/06/2026

FOX Sports' Big Noon Kickoff Heads to London for Union Jack Classic From Wembley Stadium

FOX Sports' Big Noon Kickoff will broadcast live from Wembley Stadium in Lon...

18/06/2026

InfoComm 2026: Show Opens With Microsoft Keynote, Media Day, and New Industry Initiatives

InfoComm 2026 opened on Wednesday at the Las Vegas Convention Center, bringing t...

18/06/2026

SVG New Sponsor Spotlight: Akta's Matt Smith on Building AI Into the Foundation of the Video Workflow

As media companies look to deliver more live, VOD, and snackable sports content ...

18/06/2026

2026 Sundance Film Festival: Local Lens

Top L-R: Take Me Home, The Lake Bottom L-R: TheyDream, Union County Free Summer Screening Series Announced Screenings for the Local Utah Community at...

18/06/2026

The Gauge: Poland | May 2026

The average daily TV screen time in May was 3 hours and 36 minutes, marking a clear decrease of 15 minutes compared to April. This trend proved to be much stron...

18/06/2026

Gracenote and PubMatic Bring Curated Live Sports and Content-Level Deals to Programmatic CTV

Integration embeds Gracenote content intelligence, including contextual segments...

18/06/2026

AJA Debuts Io Xpand Thunderbolt 5 Expansion Chassis for...

AJA Video Systems unveiled Io Xpand, a high-performance Thunderbolt 5-enabled PCIe expansion chassis for AJA KONA and Corvid video and audio I/O cards. As dema...

18/06/2026

Providius Announces Providius Direct for Faster AV-over-I...

New NVRT-driven workflow enables on-demand traffic mirroring and guided troubleshooting for AV teams, integrators, and support organizations ahead of InfoComm 2...

18/06/2026

Mavis Studio Brings New Muscle to Mobile Production on iP...

At InfoComm 2026, Mavis announced a major update to Mavis Studio, its live production app for the iPad, with new features designed to make professional AV produ...

18/06/2026

Harmonic Completes Divestiture of Video Business to Media...

Transaction Positions Harmonic as a Pure-Play Broadband Company Harmonic Inc. (NASDAQ: HLIT), the worldwide leader in virtualized broadband solutions, today a...

18/06/2026

ACT Entertainment and NETGEAR Partner to Power the Futur...

ACT Entertainment and NETGEAR have announced a strategic partnership that establishes verified interoperability between NETGEAR Switches and key technologies fr...

18/06/2026

IBC2026 creates new pathways from innovation to real-worl...

IBC2026 is set to bring the global media, entertainment and technology community together at the RAI Amsterdam from 11 14 September, enabling industry players f...

18/06/2026

PLIANT TECHNOLOGIES DEBUTS CREWCOM FLEX THE WORLDS FIRST...

Pliant Technologies is proud to introduce CrewCom Flex , the world's first frameless matrix intercom system and the only complete matrix IP-based intercom s...

18/06/2026

XR Sports Alliance Strengthens its Capabilities with New...

The XR Sports Alliance (XRSA) has announced that a new cohort of members has joined the strategic initiative: ActionStreamer, Antigravity, Creative Artists Agen...

18/06/2026

Hollywood Filmmakers Launch AI Production Platform Cascade

Share Copy link Facebook X Linkedin Bluesky Email...

18/06/2026

ATVA Blasts Deltavision Media for Demanding 'Egregious' Retrans Fees

Share Copy link Facebook X Linkedin Bluesky Email...

18/06/2026

MediaKind Completes Merger with Harmonic's Video Business

Share Copy link Facebook X Linkedin Bluesky Email...

18/06/2026

AJA Unveils Io Xpand Thunderbolt 5 Expansion Chassis At InfoComm 2026

Share Copy link Facebook X Linkedin Bluesky Email...

18/06/2026

Linda May Han Oh Receives Guggenheim Fellowship

Linda May Han Oh Receives Guggenheim Fellowship The Berklee professor, bassist, and composer will use the fellowship to debut Dreams of Knowing, an interdisci...

18/06/2026

How FERC's Large-Load Interconnection Actions Help Address Grid Stress, Improve Affordability

In a consequential grid infrastructure decision, the Federal Energy Regulatory C...

18/06/2026

Sync and Stream: GeForce NOW Connects to Members' Game Libraries Across Devices

Play favorite titles from popular game libraries, keep progress synced and jump ...

18/06/2026

At Cannes Lions, NVIDIA Partners Reshape Advertising and Marketing With AI

The digital era gave the advertising and marketing industry speed; the AI era is giving it autonomous operations. For companies building next-generation techn...

17/06/2026

EVS Achieves EcoVadis Gold Medal, Ranking in Top 5% of Companies Globally

EVS has announced it has received the EcoVadis Gold Medal for sustainability performance, ranking among the top 5% of companies globally in the Technology/Mid-S...

17/06/2026

Chyron Releases Weather 2.4 with Updated DataFlow Module

Chyron has released Weather 2.4, an update to its weather suite for broadcasters and meteorologists. The release focuses on enhancements to the DataFlow module,...

17/06/2026

VSiN Launches Best Bets TV, a 24/7 FAST Channel

VSiN, The Sports Betting Network, has announced the launch of Best Bets TV, a free ad-supported streaming TV (FAST) channel. The 24/7 channel is currently avail...