Sony Pixel Power calrec Sony

TOPS of the Class: Decoding AI Performance on RTX AI PCs and Workstations

12/06/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for RTX PC users.

The era of the AI PC is here, and it's powered by NVIDIA RTX and GeForce RTX technologies. With it comes a new way to evaluate performance for AI-accelerated tasks, and a new language that can be daunting to decipher when choosing between the desktops and laptops available.

While PC gamers understand frames per second (FPS) and similar stats, measuring AI performance requires new metrics.

Coming Out on TOPS The first baseline is TOPS, or trillions of operations per second. Trillions is the important word here - the processing numbers behind generative AI tasks are absolutely massive. Think of TOPS as a raw performance metric, similar to an engine's horsepower rating. More is better.

Compare, for example, the recently announced Copilot+ PC lineup by Microsoft, which includes neural processing units (NPUs) able to perform upwards of 40 TOPS. Performing 40 TOPS is sufficient for some light AI-assisted tasks, like asking a local chatbot where yesterday's notes are.

But many generative AI tasks are more demanding. NVIDIA RTX and GeForce RTX GPUs deliver unprecedented performance across all generative tasks - the GeForce RTX 4090 GPU offers more than 1,300 TOPS. This is the kind of horsepower needed to handle AI-assisted digital content creation, AI super resolution in PC gaming, generating images from text or video, querying local large language models (LLMs) and more.

Insert Tokens to Play TOPS is only the beginning of the story. LLM performance is measured in the number of tokens generated by the model.

Tokens are the output of the LLM. A token can be a word in a sentence, or even a smaller fragment like punctuation or whitespace. Performance for AI-accelerated tasks can be measured in tokens per second.

Another important factor is batch size, or the number of inputs processed simultaneously in a single inference pass. As an LLM will sit at the core of many modern AI systems, the ability to handle multiple inputs (e.g. from a single application or across multiple applications) will be a key differentiator. While larger batch sizes improve performance for concurrent inputs, they also require more memory, especially when combined with larger models.

The more you batch, the more (time) you save. RTX GPUs are exceptionally well-suited for LLMs due to their large amounts of dedicated video random access memory (VRAM), Tensor Cores and TensorRT-LLM software.

GeForce RTX GPUs offer up to 24GB of high-speed VRAM, and NVIDIA RTX GPUs up to 48GB, which can handle larger models and enable higher batch sizes. RTX GPUs also take advantage of Tensor Cores - dedicated AI accelerators that dramatically speed up the computationally intensive operations required for deep learning and generative AI models. That maximum performance is easily accessed when an application uses the NVIDIA TensorRT software development kit (SDK), which unlocks the highest-performance generative AI on the more than 100 million Windows PCs and workstations powered by RTX GPUs.

The combination of memory, dedicated AI accelerators and optimized software gives RTX GPUs massive throughput gains, especially as batch sizes increase.

Text-to-Image, Faster Than Ever Measuring image generation speed is another way to evaluate performance. One of the most straightforward ways uses Stable Diffusion, a popular image-based AI model that allows users to easily convert text descriptions into complex visual representations.

With Stable Diffusion, users can quickly create and refine images from text prompts to achieve their desired output. When using an RTX GPU, these results can be generated faster than processing the AI model on a CPU or NPU.

That performance is even higher when using the TensorRT extension for the popular Automatic1111 interface. RTX users can generate images from prompts up to 2x faster with the SDXL Base checkpoint - significantly streamlining Stable Diffusion workflows.

ComfyUI, another popular Stable Diffusion user interface, added TensorRT acceleration last week. RTX users can now generate images from prompts up to 60% faster, and can even convert these images to videos using Stable Video Diffuson up to 70% faster with TensorRT.

TensorRT acceleration can be put to the test in the new UL Procyon AI Image Generation benchmark, which delivers speedups of 50% on a GeForce RTX 4080 SUPER GPU compared with the fastest non-TensorRT implementation.

TensorRT acceleration will soon be released for Stable Diffusion 3 - Stability AI's new, highly anticipated text-to-image model - boosting performance by 50%. Plus, the new TensorRT-Model Optimizer enables accelerating performance even further. This results in a 70% speedup compared with the non-TensorRT implementation, along with a 50% reduction in memory consumption.

Of course, seeing is believing - the true test is in the real-world use case of iterating on an original prompt. Users can refine image generation by tweaking prompts significantly faster on RTX GPUs, taking seconds per iteration compared with minutes on a Macbook Pro M3 Max. Plus, users get both speed and security with everything remaining private when running locally on an RTX-powered PC or workstation.

The Results Are in and Open Sourced But don't just take our word for it. The team of AI researchers and engineers behind the open-source Jan.ai recently integrated TensorRT-LLM into its local chatbot app, then tested these optimizations for themselves.

Source: Jan.ai The researchers tested its implementation of TensorRT-LLM against the open-source llama.cpp inference engine across a variety of GPUs and CPUs used by the community. They found that TensorRT is 30-70% faster than llam
LINK: https://blogs.nvidia.com/blog/ai-decoded-tops/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

02/04/2026

Scripps Completes Sale of WRTV to Circle City Broadcasting

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

GoVertical! AiDi Powers Real-Time 9:16 Autocropping for I...

Already deployed extensively by NBC Sports, FOR-A Corporation will demonstrate GoVertical! AiDi, the real-time 9:16 autocropping feature of viztrick AiDi, durin...

02/04/2026

Elite Media Technologies Selects Interra Systems BATON Fi...

Interra Systems, a provider of end-to-end quality assurance solutions for the digital media industry, announced that Elite Media Technologies has selected its B...

02/04/2026

TDF Expands Broadcast Channel Lineup with Harmonic

Harmonic's Media Processing Solutions Maximize Bandwidth Efficiency for Terrestrial Broadcast Delivery Harmonic (NASDAQ: HLIT) today announced that TDF, a...

02/04/2026

FOR-A's Software-Defined, AI-Powered Development Advances...

NBC Sports Deploys viztrick AiDi to Stream Live Events in 9:16 Mobile-First Formats with Auto Tracking, Development Signals Strategic Shift for FOR-A Long reco...

02/04/2026

Evergent showcases innovations in sports streaming and mo...

Evergent will showcase new innovations in subscriber lifecycle management and monetization at NAB Show 2026 (Las Vegas, April 18 22), including: New advances i...

02/04/2026

Binghamton University Strengthens Student Run Productions...

Riedel Communications is proud to be part of Binghamton University, State University of New York, Athletics' milestone year, celebrating the university'...

02/04/2026

Techex and Encompass Launch Industry-Leading Cloud-Based...

Encompass Digital Media and Techex have today announced new, fully managed, cloud-native Master Control services designed to meet the growing operational demand...

02/04/2026

Winning in the new media economy - Avid debuts fully avai...

Avid today announced it will showcase new innovations designed to help media companies win in the new media economy at NAB Show 2026 (April 18 22, Las Vegas Co...

02/04/2026

PlayBox Neo reinforces MIMO Tech with new Playout capabil...

PlayBox Neo helps AIS PLAY kick-off premier football content direct to fans PlayBox Neo has provided MIMO Tech with a brand-new major installation to extend it...

02/04/2026

Globo transitions primary distribution to SRT over IP wit...

Globo has transitioned its primary content distribution to Secure Reliable Transport over a fully IP-based managed backbone using Synamedia's Quortex PowerV...

02/04/2026

Nexstar Says Pausing Tegna Merger Creates 'Impossible' Challenges

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

FCC Launches Efforts to Strengthen U.S. Drone Ecosystem

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

WAPA+ to Launch on Dish, DishLatino, Sling TV and Sling Freestream

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

Student Spotlight: Al-Fadl Salem

Student Spotlight: Al-Fadl Salem The Danish singer recently performed for the queen of Denmark. April 1, 2026 By Editorial Staff Image by Junia Morrow Wh...

02/04/2026

Taku Hirano's Career Is Defined by Identity

Taku Hirano's Career Is Defined by Identity Whether he's performing, composing, teaching, or developing instruments, the do-it-all percussionist sees ...

02/04/2026

Design Perspective Intelligent Hybrid Software Platforms to Survive the Evolutionary Avalanche

By Lance Maurer, CEO Image generated by AI Engineering is supposed to be fun....

02/04/2026

Continuing to connect with Young Ireland: 2FM Announces Brand-New Daytime Schedule

2FM Breakfast to extend on weekday mornings from 6am to 10am Doireann Garrihy m...

02/04/2026

RT NEWS ANNOUNCES BARRY LENIHAN AS NEW POLITICAL CORRESPONDENT

RT News & Current Affairs is pleased to announce the appointment of RT Radio 1 reporter, Barry Lenihan, as Political Correspondent. Barry has reported across...

02/04/2026

Press Start on April: GeForce NOW Brings 10 Games to the Cloud

No joke - GFN Thursday is skipping the tricks and heading straight into the games. April kicks off with ten new titles, bringing fresh adventures to GeForce NOW...

01/04/2026

SVG New Sponsor Spotlight: Flowstate AI's Sahil Shah on Transforming Video Content with Intelligent AI Agents

As sports media organization continue to seek out new ways to streamline their p...

01/04/2026

SVG GFX Forum 2026: Sessions Now Available to Watch on SVG PLAY

The SVG GFX Forum hit New York City earlier this month for a day packed with sessions focused on the creative strategy and technology behind today's cutting...

01/04/2026

From Buenos Aires to Mexico City, EQUAL Days Bring Latin America Together for Women in Audio

This year, Spotify celebrates the five-year anniversary of EQUAL, our global pro...

01/04/2026

FourFingers announce Tape Splice Pro plug-in

Analogue-style tape splicing in the digital domain In this era of digital recording and multiple layers of Undo, it seems that the fading art of tape splici...

01/04/2026

Zero G introduce Morphology Evolved

Latest release introduces new Orbita Engine Zero G's latest release marks the start of a new series of libraries, as well as introducing an all-new engi...

01/04/2026

Warm Audio introduce the WA-8TRX

Until now, one format has largely been left behind Warm Audio's extensive product range includes modern-day recreations of all manner of sought-after s...

01/04/2026

The Crow Hill Company announce Crystal Pianos

A piano with glass vessels for strings! The Crow Hill Company's recently released Gong Piano offered a refreshing new take on piano libraries, harnessin...

01/04/2026

ESSENCE RS from Aim Audio

Remote Streaming Studio Condenser Aim Audio have just revealed their latest creation, the ESSENCE RS Remote Streaming Studio Condenser, which becomes the wo...

01/04/2026

Call for NFVF funding applications to attend Film Festivals and Markets taking place from 08 - 31 May 2026

The National Film and Video Foundation (NFVF) is pleased to announce that the ca...

01/04/2026

AgileTV powers Liwest's next-generation TV experience with the launch of next IPTV platform in Austria

Bilbao, April 1st, 2026 - AgileTV, a leading provider of end-to-end TV technolog...

01/04/2026

Green Hippo Debuts Hands on Hippotizer Media Server Train...

Green Hippo is excited to announce the launch of its new Hippotizer Media Server training courses at Pixel Academy, a purpose built AV learning hub combining ha...

01/04/2026

TAG Video Systems and Oracle Cloud Infrastructure Partner...

TAG Video Systems, a global leader in IP-native broadcast monitoring, multiviewing, and quality control, today announced a collaboration with Oracle Cloud Infra...

01/04/2026

Professional Wireless Systems PWS Takes on Intercom and R...

Professional Wireless Systems (PWS), a leading provider of wireless audio solutions and RF management, was on site at the Caesars Superdome in New Orleans, wher...

01/04/2026

AgileTV powers Liwest next generation TV experience with...

AgileTV, a leading provider of end-to-end TV technology solutions, has deployed next , the new IPTV platform of the Austrian telco LIWEST, marking the first st...

01/04/2026

LTN and Ateme partner to deliver integrated video process...

LTN, a leader in fully managed IP video transport, and Ateme, a global leader in video compression and delivery solutions, today announced a collaboration integ...

01/04/2026

Adobe Unveils Powerful New Innovations for Creative Pros in Adobe Illustrator

Adobe Unveils Powerful New Innovations for Creative Pros in Adobe Illustrator Deepa Subramaniam April 1, 2026 0 Comments I'm excited to share that...

01/04/2026

Boland Communications Introduces QD4K315HDR10 QD-OLED Series Monitors for Live Production, Film, Post, and Broadcast

Boland Communications Introduces QD4K315HDR10 QD-OLED Series Monitors for Live P...

01/04/2026

2026 NAB Show Exhibitor Insight: Evertz

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Judge Blocks Order Barring NPR and PBS From Funding

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Nikon to Sell Mark Roberts Motion Control

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Mediagenix Showcases Semantic Intelligence-Powered Title Management, Schedule Optimization, and Personalization at NAB 2026

Mediagenix Showcases Semantic Intelligence-Powered Title Management, Schedule Op...

01/04/2026

FCC Approves WJAX-TV License Transfer to Cox

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Scripps Sports Ink Deal for Ion to Air 2026 Teal Rising Cup

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

UK Group Companies Unveil NAB Show Plans

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Victoria Mont Brings the Multi-Hyphenate Mindset to Career Jam 2026

Victoria Mon t Brings the Multi-Hyphenate Mindset to Career Jam 2026 The Grammy-winning singer, songwriter, and producer shared how versatility and self-inves...

01/04/2026

UKTV announces expanded remit for Jonathan Newman and appoints David Swetman as Director of Content Partnerships & Sales

UKTV today announces that Jonathan Newman has formally stepped into the role of ...