Sony Pixel Power calrec Sony

TOPS of the Class: Decoding AI Performance on RTX AI PCs and Workstations

12/06/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for RTX PC users.

The era of the AI PC is here, and it's powered by NVIDIA RTX and GeForce RTX technologies. With it comes a new way to evaluate performance for AI-accelerated tasks, and a new language that can be daunting to decipher when choosing between the desktops and laptops available.

While PC gamers understand frames per second (FPS) and similar stats, measuring AI performance requires new metrics.

Coming Out on TOPS The first baseline is TOPS, or trillions of operations per second. Trillions is the important word here - the processing numbers behind generative AI tasks are absolutely massive. Think of TOPS as a raw performance metric, similar to an engine's horsepower rating. More is better.

Compare, for example, the recently announced Copilot+ PC lineup by Microsoft, which includes neural processing units (NPUs) able to perform upwards of 40 TOPS. Performing 40 TOPS is sufficient for some light AI-assisted tasks, like asking a local chatbot where yesterday's notes are.

But many generative AI tasks are more demanding. NVIDIA RTX and GeForce RTX GPUs deliver unprecedented performance across all generative tasks - the GeForce RTX 4090 GPU offers more than 1,300 TOPS. This is the kind of horsepower needed to handle AI-assisted digital content creation, AI super resolution in PC gaming, generating images from text or video, querying local large language models (LLMs) and more.

Insert Tokens to Play TOPS is only the beginning of the story. LLM performance is measured in the number of tokens generated by the model.

Tokens are the output of the LLM. A token can be a word in a sentence, or even a smaller fragment like punctuation or whitespace. Performance for AI-accelerated tasks can be measured in tokens per second.

Another important factor is batch size, or the number of inputs processed simultaneously in a single inference pass. As an LLM will sit at the core of many modern AI systems, the ability to handle multiple inputs (e.g. from a single application or across multiple applications) will be a key differentiator. While larger batch sizes improve performance for concurrent inputs, they also require more memory, especially when combined with larger models.

The more you batch, the more (time) you save. RTX GPUs are exceptionally well-suited for LLMs due to their large amounts of dedicated video random access memory (VRAM), Tensor Cores and TensorRT-LLM software.

GeForce RTX GPUs offer up to 24GB of high-speed VRAM, and NVIDIA RTX GPUs up to 48GB, which can handle larger models and enable higher batch sizes. RTX GPUs also take advantage of Tensor Cores - dedicated AI accelerators that dramatically speed up the computationally intensive operations required for deep learning and generative AI models. That maximum performance is easily accessed when an application uses the NVIDIA TensorRT software development kit (SDK), which unlocks the highest-performance generative AI on the more than 100 million Windows PCs and workstations powered by RTX GPUs.

The combination of memory, dedicated AI accelerators and optimized software gives RTX GPUs massive throughput gains, especially as batch sizes increase.

Text-to-Image, Faster Than Ever Measuring image generation speed is another way to evaluate performance. One of the most straightforward ways uses Stable Diffusion, a popular image-based AI model that allows users to easily convert text descriptions into complex visual representations.

With Stable Diffusion, users can quickly create and refine images from text prompts to achieve their desired output. When using an RTX GPU, these results can be generated faster than processing the AI model on a CPU or NPU.

That performance is even higher when using the TensorRT extension for the popular Automatic1111 interface. RTX users can generate images from prompts up to 2x faster with the SDXL Base checkpoint - significantly streamlining Stable Diffusion workflows.

ComfyUI, another popular Stable Diffusion user interface, added TensorRT acceleration last week. RTX users can now generate images from prompts up to 60% faster, and can even convert these images to videos using Stable Video Diffuson up to 70% faster with TensorRT.

TensorRT acceleration can be put to the test in the new UL Procyon AI Image Generation benchmark, which delivers speedups of 50% on a GeForce RTX 4080 SUPER GPU compared with the fastest non-TensorRT implementation.

TensorRT acceleration will soon be released for Stable Diffusion 3 - Stability AI's new, highly anticipated text-to-image model - boosting performance by 50%. Plus, the new TensorRT-Model Optimizer enables accelerating performance even further. This results in a 70% speedup compared with the non-TensorRT implementation, along with a 50% reduction in memory consumption.

Of course, seeing is believing - the true test is in the real-world use case of iterating on an original prompt. Users can refine image generation by tweaking prompts significantly faster on RTX GPUs, taking seconds per iteration compared with minutes on a Macbook Pro M3 Max. Plus, users get both speed and security with everything remaining private when running locally on an RTX-powered PC or workstation.

The Results Are in and Open Sourced But don't just take our word for it. The team of AI researchers and engineers behind the open-source Jan.ai recently integrated TensorRT-LLM into its local chatbot app, then tested these optimizations for themselves.

Source: Jan.ai The researchers tested its implementation of TensorRT-LLM against the open-source llama.cpp inference engine across a variety of GPUs and CPUs used by the community. They found that TensorRT is 30-70% faster than llam
LINK: https://blogs.nvidia.com/blog/ai-decoded-tops/...
See more stories from nvidia

North America Stories

10/03/2026

Harvey Arnold, Bert Goldman to Be Honored at the 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

10/03/2026

Senators Urge FCC to Preserve Citizens Broadband Radio Service

Share Copy link Facebook X Linkedin Bluesky Email...

10/03/2026

SCTE TechExpo26 Issues Call for Content, Technical Papers

Share Copy link Facebook X Linkedin Bluesky Email...

10/03/2026

Zefr Receives MRC Accreditation

Share Copy link Facebook X Linkedin Bluesky Email...

10/03/2026

Study: Overloaded Sports Fans Fed Up with Fragmented Viewing Options

Share Copy link Facebook X Linkedin Bluesky Email...

09/03/2026

Foos Gone Wild, Combate Global Launch New Televised MMA Fight Series

Foos Gone Wild and Combate Global have teamed up to create a twist on combat sports competition, announcing the launch of a special amateur Mixed Martial Arts (...

09/03/2026

Harmonic Accelerates Streaming and Broadcast Transformations

At the 2026 NAB Show, Harmonic will introduce significant enhancements to its video appliances and SaaS solutions, highlighted by a next-generation media server...

09/03/2026

ESPN Delivers Most-Watched MLB Spring Training Game in 10 years with Team USA vs. San Francisco Giants

ESPN's March 3 spring training matchup between Team USA and the San Francisc...

09/03/2026

Most Valuable Promotions Launches Women's Boxing Platform, Signs Multi-Year Deal with ESPN

Most Valuable Promotions (MVP) announces the launch of MVPW, a new global platfo...

09/03/2026

Behind The Mic: CBS Sports and TNT Sports Share NCAA Division 1 Mens's Basketball Tournament Commentators

Behind The Mic provides a roundup of recent news regarding on-air talent, includ...

09/03/2026

SVG All-Stars: Jenna McKeon, Senior Director, Remote Technical Operations, CBS Sports

From Super Bowl compounds to Final Four setups, the Hofstra graduate helps coord...

09/03/2026

NBC's Paralympic Effort Embraces Cloud for Signal Transport

Stamford plays a key role, but a small team in Cortina and Milan powers local presence and mixed-zone coverage...

09/03/2026

Save the Date: SVG's New Cloud & Content Workflows Summit in NYC on July 28

The event brings together SVG's previous Cloud Production and Content Management Forums into a single, comprehensive day of programming...

09/03/2026

2,000-Year-Old Arena Hosts Paralympics Opening Ceremony, Olympics Closing Ceremony

Updated Mar 9, 2026 Live surround sound has been a part of the plan for Roman a...

09/03/2026

Judge Rules VOA's Kari Lake Has Acted Unlawfully'

Share Copy link Facebook X Linkedin Bluesky Email...

09/03/2026

Utah Scientific Adds Three Companies To Technology Partner Program

Share Copy link Facebook X Linkedin Bluesky Email...

09/03/2026

Broadpeak Showcases Premium Live Streaming Advanced Monet...

Broadpeak, a leader in streaming and monetization at scale, will showcase its latest innovations for broadcasters and streaming platforms at NAB Show 2026 (boot...

09/03/2026

'The Predator of Seville' premieres on Netflix on 27 March

Back to All News The Predator of Seville premieres on Netflix on 27 March Entertainment 09 March 2026 GlobalSpain Link copied to clipboard Download the im...

09/03/2026

Netflix Debuts the Trailer for 'Love is Blind: Sweden' Season 3

Back to All News Netflix Debuts the Trailer for Love is Blind: Sweden Season 3 Entertainment 09 March 2026 GlobalSweden Link copied to clipboard That wait...

09/03/2026

How AI Is Driving Revenue, Cutting Costs and Boosting Productivity for Every Industry in 2026

AI is everywhere and accelerating everything - becoming essential infrastructure...

09/03/2026

ABB Robotics Taps NVIDIA Omniverse to Deliver IndustrialGrade Physical AI at Scale

ABB Robotics and NVIDIA today announced a breakthrough partnership that brings i...

07/03/2026

NAB Show: Tedial to Showcase Solutions for Future of Media Operations

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

TNT Sports Acquires Exclusive U.S. English Language Broadcast Rights to FIBA Men's and Women's Tournaments

TNT Sports and the International Basketball Federation (FIBA) have reached a mul...

06/03/2026

OffBall to Partner with TOGETHXR Across Commercial Strategy and Operations

OffBall and TOGETHXR, two influential young media companies in sports, announce a strategic and operational partnership in a shared push to scale and create inn...

06/03/2026

InfoComm 2026 Names Shure as Exclusive Headline Partner, Showcasing Audio and Innovation Across Key Activations and Stages

InfoComm 2026, a destination for AV, IT, broadcast, and AI-driven systems, annou...

06/03/2026

LTN, MediaKind Partner to Deliver Integrated Reliable IP Transport and Edge Processing

LTN and MediaKind announce a strategic partnership to integrate MediaKind's ...

06/03/2026

X Games Brings First-Ever Summer Championship Event to New Orleans in July

X Games and the Greater New Orleans Sports Foundation (GNOSF) announce that New Orleans, Louisiana, will host the first-ever X Games Championship event - the fi...

06/03/2026

Chyron Releases New Edition of AXIS Maps

As part of a busy start of the year at Chyron, the AXIS team developed a set of improvements for AXIS Maps. The features released empower users with more flexib...

06/03/2026

BeckTV Launches BeckFlow at 2026 NAB Show While Highlighting its Latest Design and Integration Projects

At the 2026 NAB Show, BeckTV, a premier systems integrator for the broadcast ind...

06/03/2026

Net Insight Sets New Standard for Live Media Operations with Nimbra Live Intelligence

At NAB Show 2026, Net Insight introduces Nimbra Live Intelligence, which definin...

06/03/2026

SES Adds New MEO Capacity as Latest O3b mPOWER Satellites Enter Commercial Service

SES announces that it has added new Medium Earth Orbit (MEO) satellite capacity ...

06/03/2026

LucidLink Launches Connect to Extend Instant Access to Data Stores

LucidLink, the cloud-native file streaming platform for instant, secure access to large files, announce LucidLink Connect, a new solution that enables real-time...

06/03/2026

Case Study: Panasonic Projection Brings Paddington to Life in London's West End

Panasonic helped bring the world of Paddington: The Musical to life through imme...

06/03/2026

SVG GameDay, Ep. 6: Cincinnati Bengals' Alex Schweppe - Welcome to the Jungle

In-venue and creative video staffers at the professional and collegiate level ha...

06/03/2026

Ratings Roundup: Post Olympics High, NHL on ESPN Viewership Spikes Over 50%

Ratings Roundup is a rundown of recent rating news and is derived from press releases and reports around the industry. In this week's edition, NASCAR Cup Se...

06/03/2026

Riedel Communications Demos Product Innovations at 2026 NAB Show

At the 2026 NAB Show, Riedel Communications opens with a clear message to the North American market: production technology does not have to be complex. This ye...

06/03/2026

SVG Sit-Down: Midco Sports' Andy Price and Craig DeWit on How the Dakotas-Based RSN Is Redefining Regional Sports Media

Midco Sports isn't your typical regional sports network. Backed by nearly a ...

06/03/2026

Inside the Paralympics with OBS Producer Josephine Xiaofan

The stories around the Paralympics are really touching and the athletes, the atmosphere it is all amazing and we want to use new technologies to help us tell th...

06/03/2026

Apple TV Kicks Off F1 Era in U.S. With Driver Tracker, On-Board Cameras, Multiview, Sky Sports Feed

Apple TV subscribers will have access to as many as 30 additional live feeds acr...

06/03/2026

Nielsen: 46 Billion Minutes of Women's Sports were Consumed in 2025*

Nielsen Highlights Women's Sports Viewership Milestones Ahead of International Women's Day on March 8 New York March 5, 2026 According to Nielsen&#...

06/03/2026

MXL Moves Into Stable Production Release

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

LTN, MediaKind Partner on Industry Transition From C-Band to IP

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

Roku Unveils 'Roklue' Interactive Content Discovery Feature

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

FCC Releases Tentative Agenda for March Open Meeting

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

ESPN to Air Animated Version of Capitals vs. Rangers NHL Game

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

Tedial to Highlight AI-Fueled Media Lifecycle at 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

06/03/2026

esRadio Advances with DHD RX2 Audio Production Consoles

Spanish FM and online broadcaster esRadio has selected DHD RX2 audio production consoles for use at its studio headquarters in Madrid. Part of the Libertad Digi...

06/03/2026

LucidLink launches Connect to extend instant access to da...

LucidLink, the cloud-native file streaming platform for instant, secure access to large files, today announced LucidLink Connect, a new solution that enables re...

06/03/2026

BeckTV Launches BeckFlow at 2026 NAB Show While Highlight...

At the 2026 NAB Show, BeckTV, a premier systems integrator for the broadcast industry, is launching BeckFlow, the first web-based schematic documentation platfo...

06/03/2026

Net Insight sets a New Standard for Live Media Operations...

At NAB Show 2026, Net Insight introduces Nimbra Live Intelligence defining a new category and setting a new standard for live operations: the Open Media Platfor...