Sony Pixel Power calrec Sony

TOPS of the Class: Decoding AI Performance on RTX AI PCs and Workstations

12/06/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for RTX PC users.

The era of the AI PC is here, and it's powered by NVIDIA RTX and GeForce RTX technologies. With it comes a new way to evaluate performance for AI-accelerated tasks, and a new language that can be daunting to decipher when choosing between the desktops and laptops available.

While PC gamers understand frames per second (FPS) and similar stats, measuring AI performance requires new metrics.

Coming Out on TOPS The first baseline is TOPS, or trillions of operations per second. Trillions is the important word here - the processing numbers behind generative AI tasks are absolutely massive. Think of TOPS as a raw performance metric, similar to an engine's horsepower rating. More is better.

Compare, for example, the recently announced Copilot+ PC lineup by Microsoft, which includes neural processing units (NPUs) able to perform upwards of 40 TOPS. Performing 40 TOPS is sufficient for some light AI-assisted tasks, like asking a local chatbot where yesterday's notes are.

But many generative AI tasks are more demanding. NVIDIA RTX and GeForce RTX GPUs deliver unprecedented performance across all generative tasks - the GeForce RTX 4090 GPU offers more than 1,300 TOPS. This is the kind of horsepower needed to handle AI-assisted digital content creation, AI super resolution in PC gaming, generating images from text or video, querying local large language models (LLMs) and more.

Insert Tokens to Play TOPS is only the beginning of the story. LLM performance is measured in the number of tokens generated by the model.

Tokens are the output of the LLM. A token can be a word in a sentence, or even a smaller fragment like punctuation or whitespace. Performance for AI-accelerated tasks can be measured in tokens per second.

Another important factor is batch size, or the number of inputs processed simultaneously in a single inference pass. As an LLM will sit at the core of many modern AI systems, the ability to handle multiple inputs (e.g. from a single application or across multiple applications) will be a key differentiator. While larger batch sizes improve performance for concurrent inputs, they also require more memory, especially when combined with larger models.

The more you batch, the more (time) you save. RTX GPUs are exceptionally well-suited for LLMs due to their large amounts of dedicated video random access memory (VRAM), Tensor Cores and TensorRT-LLM software.

GeForce RTX GPUs offer up to 24GB of high-speed VRAM, and NVIDIA RTX GPUs up to 48GB, which can handle larger models and enable higher batch sizes. RTX GPUs also take advantage of Tensor Cores - dedicated AI accelerators that dramatically speed up the computationally intensive operations required for deep learning and generative AI models. That maximum performance is easily accessed when an application uses the NVIDIA TensorRT software development kit (SDK), which unlocks the highest-performance generative AI on the more than 100 million Windows PCs and workstations powered by RTX GPUs.

The combination of memory, dedicated AI accelerators and optimized software gives RTX GPUs massive throughput gains, especially as batch sizes increase.

Text-to-Image, Faster Than Ever Measuring image generation speed is another way to evaluate performance. One of the most straightforward ways uses Stable Diffusion, a popular image-based AI model that allows users to easily convert text descriptions into complex visual representations.

With Stable Diffusion, users can quickly create and refine images from text prompts to achieve their desired output. When using an RTX GPU, these results can be generated faster than processing the AI model on a CPU or NPU.

That performance is even higher when using the TensorRT extension for the popular Automatic1111 interface. RTX users can generate images from prompts up to 2x faster with the SDXL Base checkpoint - significantly streamlining Stable Diffusion workflows.

ComfyUI, another popular Stable Diffusion user interface, added TensorRT acceleration last week. RTX users can now generate images from prompts up to 60% faster, and can even convert these images to videos using Stable Video Diffuson up to 70% faster with TensorRT.

TensorRT acceleration can be put to the test in the new UL Procyon AI Image Generation benchmark, which delivers speedups of 50% on a GeForce RTX 4080 SUPER GPU compared with the fastest non-TensorRT implementation.

TensorRT acceleration will soon be released for Stable Diffusion 3 - Stability AI's new, highly anticipated text-to-image model - boosting performance by 50%. Plus, the new TensorRT-Model Optimizer enables accelerating performance even further. This results in a 70% speedup compared with the non-TensorRT implementation, along with a 50% reduction in memory consumption.

Of course, seeing is believing - the true test is in the real-world use case of iterating on an original prompt. Users can refine image generation by tweaking prompts significantly faster on RTX GPUs, taking seconds per iteration compared with minutes on a Macbook Pro M3 Max. Plus, users get both speed and security with everything remaining private when running locally on an RTX-powered PC or workstation.

The Results Are in and Open Sourced But don't just take our word for it. The team of AI researchers and engineers behind the open-source Jan.ai recently integrated TensorRT-LLM into its local chatbot app, then tested these optimizations for themselves.

Source: Jan.ai The researchers tested its implementation of TensorRT-LLM against the open-source llama.cpp inference engine across a variety of GPUs and CPUs used by the community. They found that TensorRT is 30-70% faster than llam
LINK: https://blogs.nvidia.com/blog/ai-decoded-tops/...
See more stories from nvidia

North America Stories

08/01/2026

Richard E. Wiley to Step Down as Media Institute's Chairman

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

08/01/2026

Cineverse Acquires Giant Worldwide

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

08/01/2026

Maxon Introduces Cinebench 2026

Maxon's new release of Cinebench features performance enhancements and adds support for the latest Nvidia and AMD GPUs as well as Apple Silicon. Maxon is t...

08/01/2026

Zixi Accelerates Global Growth with Appointment of Heathe...

Zixi, the industry leader in IP-based video transport and orchestration, today announced the appointment of Heather Mellish as Vice President, Global Sales. In...

08/01/2026

Pebble future-proofs playout at Canal Sur

Pebble, the leading automation, content management and integrated channel specialist, has provided a complete update of its installation at Canal Sur in Spain. ...

08/01/2026

Panasonics success in US market with Flagship Z95B OLED T...

iWedia, a global leader in software solutions for connected TV devices, proudly announces the success of its collaboration with Panasonic on the Z95B OLED TV, o...

08/01/2026

Secuoya Chile Invests in Ikegami UHK-X600 and UHL-X40 Cam...

Secuoya Chile, a leading provider of television content creation and supporting services, has invested in Ikegami UHK-X600 and UHL-X40 broadcast cameras as the ...

08/01/2026

Kiloview Highlights its Integrated AV-over-IP Ecosystem a...

Kiloview, a global leader in AV-over-IP solutions, will showcase its latest innovations at ISE 2026, highlighting the continued evolution of its complete, light...

08/01/2026

iWedia and Realtek Strengthen Collaboration to Shape the...

iWedia, a global leader in software solutions for connected TV devices, and Realtek, a leading global SoC design house, today announced the next phase of their ...

08/01/2026

CJP Broadcast delivers new pitch-side media facility for...

CJP Broadcast has completed a new pitch-side media installation for Cinderford RFC, creating a flexible production setup that supports match coverage, coaching ...

08/01/2026

PlayBox Neo CEO Defines Vision for 2026 after Record Year...

PlayBox Neo further drives momentum in Playout, Streaming, Media Management and Delivery "With a brand new year at PlayBox Neo already off to a flying start, I...

08/01/2026

Boston Conservatory at Berklee Presents Second Annual Commercial Dance BFA Concert

Boston Conservatory at Berklee Presents Second Annual Commercial Dance BFA Conce...

07/01/2026

SVG Summit 2025: All General Sessions Now Available to Watch on SVG PLAY

SVG Summit 2025: All General Sessions Now Available to Watch on SVG PLAYThe SVG Summit celebrated its 20th edition with a loaded agendaBy Brandon Costa, Directo...

07/01/2026

NBCU's Brings Rinkside Live' and Courtside Live' Features to Peacock for Legendary February'

NBCU's Brings Rinkside Live' and Courtside Live' Features to Peaco...

07/01/2026

For L3Harris, the Action Starts Six Seconds Before Artemis II Lifts Off

NASA's Space Launch System rocket carrying the Orion spacecraft launches on the Artemis I flight test, Wednesday, Nov. 16, 2022, from Launch Complex 39B at ...

07/01/2026

L3Harris ROVER and TNR Products Receive NSA Approval for International Sales

For the first time ever, new C' variants allow international users to interoperate with U.S. personnel via NSA-certified cryptography....

07/01/2026

ATSC 3.0 Home Gateways To Appear At CES 2026

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

SDVI Names Simon Eldridge Chief Operating Officer

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Globecast Promotes Chris Pulis to Group Chief Technology Officer

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

English Football League Secures Match Streams With Nagravision

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Kristy Santiago Named GM of Gray Stations in Missouri, Kentucky

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Proton to Feature New Zoom Capability for 4K Minicam at Hamburg Open

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Warner Bros. Discovery Rejects Latest Paramount Offer

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

NBCUniversal Sells Out Ad Inventory for Olympic Winter Games

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

New AES Technical Document Focuses on Dialogue Intelligibility

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

CTA: U.S. Consumer Tech Revenue to Hit $565 Billion in 2026

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

NBCUniversal Upgrades AI-Powered Guide for 2026 Winter Olympics

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Parks: Smart TVs Are Primary Streaming Device in U.S. Homes

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Clear-Com to Feature 4-Channel HelixNet Beltpack at ISE 2026

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

Cable Center to Sell Its Building to University of Denver

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

CES: Aktas AI-First Video Platform Adds New Capabilities

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

07/01/2026

The Trailer of 'Paparazzi King': The 5-Episode Docuseries Coming Only on Netflix January 9

Back to All News The Trailer of Paparazzi King: The 5-Episode Docuseries Coming...

07/01/2026

Netflix Supports Warner Bros. Discovery Board's Commitment to Merger Agreement

Back to All News Netflix Supports Warner Bros. Discovery Board's Commitment...

07/01/2026

What Next? Netflix Reveals Series, Films and Games Coming in 2026

Back to All News What Next? Netflix Reveals Series, Films and Games Coming in 2026 Entertainment 07 January 2026 Global Link copied to clipboard Download ...

07/01/2026

From Warehouse to Wallet: New State of AI in Retail and CPG Survey Uncovers How AI Is Rewiring Supply Chains and Customer Experiences

AI has transformed retail and consumer packaged goods (CPG) operations, enhancin...

06/01/2026

Peacock Bringing Dolby Vision and Dolby Atmos to Live Sports Content

Peacock Bringing Dolby Vision and Dolby Atmos to Live Sports ContentPeacock is the first streamer to embrace Dolbys full suite of advanced picture/sound innovat...

06/01/2026

Milano Cortina 2026: Listening to the Sounds of Powder and Ice With a Behind the Scenes Tour of OBS and NBC's Audio Set Ups

SVG Europe Audio: Listening to the sounds of powder and ice at Milano Cortina wi...

06/01/2026

Quintar Meta Spatial SDK Integration Promises Next-Level XR Experiences

Quintar Meta Spatial SDK Integration Promises Next-Level XR ExperiencesBy Ken Kerschbaumer, Editorial Director Tuesday, January 6, 2026 - 9:41 am Print This...

06/01/2026

Milano Cortina 2026: BBC Sport Previews Broadcast Ops, Studio Setup, Social Media Plans, and More

Milano Cortina 2026: BBC Sport Previews Broadcast Ops, Studio Setup, Social Medi...

06/01/2026

Dolby's Jason Power on Elevating Live Sports Through Immersive Audio and HDR Imagery

Dolby's Jason Power on Elevating Live Sports Through Immersive Audio and HDR...

06/01/2026

DAZN's Global CRO Walker Jacobs on the Streamer's Breakout Year in the U.S.

DAZN's Global CRO Walker Jacobs on the Streamer's Breakout Year in the U...

06/01/2026

California Dreamin': LA28's SVP of Media Jim Bell Previews the Olympic and Paralympic Games' Return to the U.S.

California Dreamin': LA28's SVP of Media Jim Bell Previews the Olympic a...

06/01/2026

One Month Out From Winter Olympics Opening Ceremony, NBC Sports in Final Prep for a Legendary February'

One Month Out From Winter Olympics Opening Ceremony, NBC Sports in Final Prep fo...

06/01/2026

Advanced HDR by Technicolor and Zinwell Integrated onto ATSC 3.0 Conversion Boxes

Advanced HDR by Technicolor and Zinwell Integrated onto ATSC 3.0 Conversion Boxe...

06/01/2026

Fox Sports President and COO Mark Silverman Shifts to Consulting Role

Fox Sports President and COO Mark Silverman Shifts to Consulting RoleBy Ken Kerschbaumer, Editorial Director Tuesday, January 6, 2026 - 2:37 pm Print This S...

06/01/2026

L3Harris CFO and Missile Solutions President Appears on CNBC

In a live broadcast, L3Harris CFO and Missile Solutions President Ken Bedingfield joined Morgan Brennan on CNBCs Closing Bell: Overtime. Bedingfield discussed...

06/01/2026

Index Exchange Launches Gracenote-Powered Show-Level Reporting with Spectrum Reach

First-of-its-kind SSP capability delivers program-level insight and brand suitab...

06/01/2026

iWedia, Skyworth Partner on Turnkey Solutions for NextGen TV

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

06/01/2026

Hub: Younger Viewers More Receptive to Ads on Streaming Services

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...