Sony Pixel Power calrec Sony

TOPS of the Class: Decoding AI Performance on RTX AI PCs and Workstations

12/06/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for RTX PC users.

The era of the AI PC is here, and it's powered by NVIDIA RTX and GeForce RTX technologies. With it comes a new way to evaluate performance for AI-accelerated tasks, and a new language that can be daunting to decipher when choosing between the desktops and laptops available.

While PC gamers understand frames per second (FPS) and similar stats, measuring AI performance requires new metrics.

Coming Out on TOPS The first baseline is TOPS, or trillions of operations per second. Trillions is the important word here - the processing numbers behind generative AI tasks are absolutely massive. Think of TOPS as a raw performance metric, similar to an engine's horsepower rating. More is better.

Compare, for example, the recently announced Copilot+ PC lineup by Microsoft, which includes neural processing units (NPUs) able to perform upwards of 40 TOPS. Performing 40 TOPS is sufficient for some light AI-assisted tasks, like asking a local chatbot where yesterday's notes are.

But many generative AI tasks are more demanding. NVIDIA RTX and GeForce RTX GPUs deliver unprecedented performance across all generative tasks - the GeForce RTX 4090 GPU offers more than 1,300 TOPS. This is the kind of horsepower needed to handle AI-assisted digital content creation, AI super resolution in PC gaming, generating images from text or video, querying local large language models (LLMs) and more.

Insert Tokens to Play TOPS is only the beginning of the story. LLM performance is measured in the number of tokens generated by the model.

Tokens are the output of the LLM. A token can be a word in a sentence, or even a smaller fragment like punctuation or whitespace. Performance for AI-accelerated tasks can be measured in tokens per second.

Another important factor is batch size, or the number of inputs processed simultaneously in a single inference pass. As an LLM will sit at the core of many modern AI systems, the ability to handle multiple inputs (e.g. from a single application or across multiple applications) will be a key differentiator. While larger batch sizes improve performance for concurrent inputs, they also require more memory, especially when combined with larger models.

The more you batch, the more (time) you save. RTX GPUs are exceptionally well-suited for LLMs due to their large amounts of dedicated video random access memory (VRAM), Tensor Cores and TensorRT-LLM software.

GeForce RTX GPUs offer up to 24GB of high-speed VRAM, and NVIDIA RTX GPUs up to 48GB, which can handle larger models and enable higher batch sizes. RTX GPUs also take advantage of Tensor Cores - dedicated AI accelerators that dramatically speed up the computationally intensive operations required for deep learning and generative AI models. That maximum performance is easily accessed when an application uses the NVIDIA TensorRT software development kit (SDK), which unlocks the highest-performance generative AI on the more than 100 million Windows PCs and workstations powered by RTX GPUs.

The combination of memory, dedicated AI accelerators and optimized software gives RTX GPUs massive throughput gains, especially as batch sizes increase.

Text-to-Image, Faster Than Ever Measuring image generation speed is another way to evaluate performance. One of the most straightforward ways uses Stable Diffusion, a popular image-based AI model that allows users to easily convert text descriptions into complex visual representations.

With Stable Diffusion, users can quickly create and refine images from text prompts to achieve their desired output. When using an RTX GPU, these results can be generated faster than processing the AI model on a CPU or NPU.

That performance is even higher when using the TensorRT extension for the popular Automatic1111 interface. RTX users can generate images from prompts up to 2x faster with the SDXL Base checkpoint - significantly streamlining Stable Diffusion workflows.

ComfyUI, another popular Stable Diffusion user interface, added TensorRT acceleration last week. RTX users can now generate images from prompts up to 60% faster, and can even convert these images to videos using Stable Video Diffuson up to 70% faster with TensorRT.

TensorRT acceleration can be put to the test in the new UL Procyon AI Image Generation benchmark, which delivers speedups of 50% on a GeForce RTX 4080 SUPER GPU compared with the fastest non-TensorRT implementation.

TensorRT acceleration will soon be released for Stable Diffusion 3 - Stability AI's new, highly anticipated text-to-image model - boosting performance by 50%. Plus, the new TensorRT-Model Optimizer enables accelerating performance even further. This results in a 70% speedup compared with the non-TensorRT implementation, along with a 50% reduction in memory consumption.

Of course, seeing is believing - the true test is in the real-world use case of iterating on an original prompt. Users can refine image generation by tweaking prompts significantly faster on RTX GPUs, taking seconds per iteration compared with minutes on a Macbook Pro M3 Max. Plus, users get both speed and security with everything remaining private when running locally on an RTX-powered PC or workstation.

The Results Are in and Open Sourced But don't just take our word for it. The team of AI researchers and engineers behind the open-source Jan.ai recently integrated TensorRT-LLM into its local chatbot app, then tested these optimizations for themselves.

Source: Jan.ai The researchers tested its implementation of TensorRT-LLM against the open-source llama.cpp inference engine across a variety of GPUs and CPUs used by the community. They found that TensorRT is 30-70% faster than llam
LINK: https://blogs.nvidia.com/blog/ai-decoded-tops/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

04/08/2026

Dalet Announces Commercial Availability of Dalia, Bringing Media-Aware Agentic AI to Enterprise Productions

Dalet, a leading technology and service provider for media-rich organizations, t...

04/07/2026

Detective Conan: Fallen Angel of the Highway Opens in Dolby Cinemas Across Japan, Presented in Dolby Atmos and Dolby ...

April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...

04/06/2026

Sony's New PTZ Cameras Deliver 4K 60p; New STARVIS Sensor Meets Low-Light Demands

Sony Electronics is introducing the SRG-AS10, a 4K 60p-compatible PTZ auto-frami...

04/06/2026

SVG Students To Watch: Alex Albert, Texas A&M University

This recent grad from Spring, TX, led creative-video output for the Aggies' men's basketball team last season and has been producing video and creating ...

04/06/2026

USGA Brings AI Recaps, 3D Range Tracking, and Predictive Shot Tracing to U.S. Womens Open

For the first time at a women's golf major, every player in the field will r...

04/06/2026

Panasonic PT-RQ45 Projectors Power Lille Video Mapping Festival Opera House Installation

Three Panasonic PT-RQ45 40,000-lumen 3-Chip DLP projectors made their first live...

04/06/2026

Bitmovin and Akamai Support NRJ Groups Deployment of Akamai Adaptive Media Player 2

Bitmovin and Akamai have announced a collaboration with NRJ Group, a French mult...

04/06/2026

Telestream to Exhibit at InfoComm 2026 with Live Production and Media Workflow Demonstrations

Telestream will exhibit at InfoComm 2026 (Booth N7952), demonstrating media work...

04/06/2026

Sony Announces RIALTO 65 Image Sensor Block for VENICE 2, Targeting 2027 Release

Sony has announced the development of RIALTO 65, a 65mm format image sensor block for the VENICE 2 digital cinema camera, targeting release in the first half of...

04/06/2026

KOKUSAI DENKI Electric America to Exhibit 4K and Remote Production Solutions at InfoComm 2026

KOKUSAI DENKI Electric America will exhibit at InfoComm 2026 (Booth N8025, June ...

04/06/2026

Bell Media to Carry All 104 FIFA World Cup 2026 Matches Across TSN, RDS, and Streaming Platforms

Bell Media's TSN and RDS are the exclusive Canadian broadcasters of FIFA Wor...

04/06/2026

MASV Case Study: How MASV Reduced Miami HEAT's Road Game Video Transfer Times by 85%

The Challenge: Receiving Heavy Media Files From Road Games Quickly and ReliablyT...

04/06/2026

MASV Outlines Seven-Step Sports Analytics Workflow, Highlights File Transfer as Key Bottleneck

MASV, a managed file transfer platform used in broadcast and live sports product...

04/06/2026

NESN Appoints Fahad Haider as Vice President of Operations and Engineering

NESN has announced the appointment of Fahad Haider as Vice President of Operations and Engineering. Haider returns to NESN, where he previously served as Vice P...

04/06/2026

Sports Broadcaster, Executive, and Author David J. Halberstam Dies

David J. Halberstam, who spent almost 50 years in sports as a broadcaster and an executive, died June 2 after a years-long battle with brain cancer. Over his l...

04/06/2026

Grass Valleys Ben Dolinky on Offering Teachable Technology to College Students Across the Country

Although collegiate production programs are tasked with delivering high-quality ...

04/06/2026

Prime Video Caps First NBA Season With Sports Emmy Win, Carries In-House Production Into WNBA Campaign

California studio, two production trucks, global distribution system are combine...

04/06/2026

TikTok and Sundance Collab Launch Micro-Series Storytelling Program

New global program empowers and supports storytellers through scriptwriting course and access to industry experts TikTok and Sundance Institute today announce...

04/06/2026

Celemony announce Tonalic ARA support for Cubase & Nuendo

Steinberg DAWs now boast in-depth Tonalic integration Celemony's innovative virtual session musician plug-in has just received an update that brings ARA...

04/06/2026

GearExpo UK: Microphone Update

Get Hands-On With Over 20 Mic Brands GearExpo UK is fast approaching, and if you've been looking for a chance to check out some new mics, then you'r...

04/06/2026

Positive Grid launch Reactor amp range

Combos feature new Amplifier Intelligence engine Positive Grid's latest release sees the company introduce two new combo amplifiers that promise to offe...

04/06/2026

Is Your Job Making You Work this June?

Is Your Job Making You Work this June? 4 June, 2026 Media releases SBS Launches the World Cup Watchers' Rights Association to Stand Up For Australians&...

04/06/2026

Statement regarding unauthorised use of SBS logos on third party social content

Statement regarding unauthorised use of SBS logos on third party social content 4 June, 2026 Media releases SBS has become aware of social media posts in c...

04/06/2026

TiVo: TV Viewing Hits Post-Pandemic Peak

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

Bitmovin and Akamai Support NRJ Group to Deploy Next Gene...

Bitmovin, a leading provider of video streaming infrastructure, and Akamai, the cybersecurity and cloud computing company that powers and protects business onli...

04/06/2026

American Underground Opens at American Tobacco Campus, Completing a 16-Year Full-Circle Story

American Underground (AU), the Startup Hub of the South and a community of mor...

04/06/2026

Nielsen: Thunder Rolls as NBA's Most-Watched Team

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

AI Drives Lenovo's 2026 FIFA World Cup Broadcast Plans

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

ATSC Conference Looks Beyond Traditional TV for 3.0 Success

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

ATSC Awards Highest Technical Honor to Julia Kenyon

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

Lumine Group to Acquire Synamedias Video Network Business

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

BFOA Launches 5th Annual Giving Day

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

Hemisphere Media Group Brings WAPA+ Fast Channel to Prime Video

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

APTS to Hold June 4 'Protect My Public Media Day'

Share Copy link Facebook X Linkedin Bluesky Email...

04/06/2026

NVIDIA and Microsoft Reinvent Windows PCs for the Age of Personal AI

NVIDIA and Microsoft Reinvent Windows PCs for the Age of Personal AI Brie Clayton June 3, 2026 0 Comments RTX Spark - a 1-Petaflop Superchip, the Full...

04/06/2026

Inside the Peaky Blinders: The Immortal Man Grade

Inside the Peaky Blinders: The Immortal Man Grade Brie Clayton June 3, 2026 0 Comments Simone Grattarola discusses shaping the look in DaVinci Resolve...

04/06/2026

Cine Gear Expo Announces 2026 Awards of Excellence Recipients

Cine Gear Expo Announces 2026 Awards of Excellence Recipients Brie Clayton June 3, 2026 0 Comments Ed Lachman ASC, Caleb Deschanel ASC, and M. David M...

04/06/2026

St. Vincents Live with Orchestra Tour to Feature Berklee Alum Ruby Plume

St. Vincents Live with Orchestra Tour to Feature Berklee Alum Ruby Plume Berklee alumni St. Vincent and Ruby Plume will appear on the same bill across seven d...

04/06/2026

'All the Truth in My Lies' Coming to Netflix on August 28

Back to All News All the Truth in My Lies Coming to Netflix on August 28 Entertainment 04 June 2026 GlobalSpain Link copied to clipboard Download the imag...

04/06/2026

Larry Tanz, VP of Content, EMEA, Delivers a Keynote Speech at the Enders TMT Leaders Live 2026 Conference

Back to All News Larry Tanz, VP of Content, EMEA, Delivers a Keynote Speech at ...

04/06/2026

Turn Your Living Room into a Stadium With the New FIFA World Cup: Launch Edition' Game, Exclusively on Netflix Games June 11

Back to All News Turn Your Living Room into a Stadium With the New FIFA World ...

04/06/2026

Keeping conversations real on LinkedIn

Keeping conversations real on LinkedIn Published on Jun 4, 2026 Categories: Product News LinkedIn Corporate Communications Share LinkedIn Facebook ...

04/06/2026

Culture and family at the centre of new RT travel series

BackStory follows four Irish young people as they travel back to their parents' homelands Modern Irish identity is enriched by cultures and influences from...

04/06/2026

Forecast: Fun Ahead - 18 Games Join in June to Stream on GeForce NOW

June's forecast with GeForce NOW: 100% chance of gaming. GeForce NOW is lining up new adventures for the month, from big-name blockbusters to quirky indies...

03/06/2026

SES Launches Multi-Orbit Satellite Inflight Connectivity on Viva Airlines

SES and Viva, Mexico's ultra low-cost airline, have launched multi-orbit satellite inflight connectivity on Viva's Airbus aircraft. A total of 60 A320s ...

03/06/2026

CFP, ESPN, and TNT Sports Announce 2026-27 College Football Playoff Broadcast Schedule

The College Football Playoff, ESPN, and TNT Sports have announced kick times and...

03/06/2026

RED Digital Cinema to Host Panels and Demos at Cine Gear Expo 2026

RED Digital Cinema will exhibit at Cine Gear Expo 2026 (Booth 33, June 5-6, Universal Studios Lot), hosting three panels and hands-on product demonstrations. P...

03/06/2026

Roku Launches Soccer Zone for FIFA World Cup 2026

Roku has announced the Soccer Zone, a dedicated hub for FIFA World Cup 2026 content available across the United States, Canada, Mexico, Brazil, Colombia, Argent...