Sony Pixel Power calrec Sony

TOPS of the Class: Decoding AI Performance on RTX AI PCs and Workstations

12/06/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for RTX PC users.

The era of the AI PC is here, and it's powered by NVIDIA RTX and GeForce RTX technologies. With it comes a new way to evaluate performance for AI-accelerated tasks, and a new language that can be daunting to decipher when choosing between the desktops and laptops available.

While PC gamers understand frames per second (FPS) and similar stats, measuring AI performance requires new metrics.

Coming Out on TOPS The first baseline is TOPS, or trillions of operations per second. Trillions is the important word here - the processing numbers behind generative AI tasks are absolutely massive. Think of TOPS as a raw performance metric, similar to an engine's horsepower rating. More is better.

Compare, for example, the recently announced Copilot+ PC lineup by Microsoft, which includes neural processing units (NPUs) able to perform upwards of 40 TOPS. Performing 40 TOPS is sufficient for some light AI-assisted tasks, like asking a local chatbot where yesterday's notes are.

But many generative AI tasks are more demanding. NVIDIA RTX and GeForce RTX GPUs deliver unprecedented performance across all generative tasks - the GeForce RTX 4090 GPU offers more than 1,300 TOPS. This is the kind of horsepower needed to handle AI-assisted digital content creation, AI super resolution in PC gaming, generating images from text or video, querying local large language models (LLMs) and more.

Insert Tokens to Play TOPS is only the beginning of the story. LLM performance is measured in the number of tokens generated by the model.

Tokens are the output of the LLM. A token can be a word in a sentence, or even a smaller fragment like punctuation or whitespace. Performance for AI-accelerated tasks can be measured in tokens per second.

Another important factor is batch size, or the number of inputs processed simultaneously in a single inference pass. As an LLM will sit at the core of many modern AI systems, the ability to handle multiple inputs (e.g. from a single application or across multiple applications) will be a key differentiator. While larger batch sizes improve performance for concurrent inputs, they also require more memory, especially when combined with larger models.

The more you batch, the more (time) you save. RTX GPUs are exceptionally well-suited for LLMs due to their large amounts of dedicated video random access memory (VRAM), Tensor Cores and TensorRT-LLM software.

GeForce RTX GPUs offer up to 24GB of high-speed VRAM, and NVIDIA RTX GPUs up to 48GB, which can handle larger models and enable higher batch sizes. RTX GPUs also take advantage of Tensor Cores - dedicated AI accelerators that dramatically speed up the computationally intensive operations required for deep learning and generative AI models. That maximum performance is easily accessed when an application uses the NVIDIA TensorRT software development kit (SDK), which unlocks the highest-performance generative AI on the more than 100 million Windows PCs and workstations powered by RTX GPUs.

The combination of memory, dedicated AI accelerators and optimized software gives RTX GPUs massive throughput gains, especially as batch sizes increase.

Text-to-Image, Faster Than Ever Measuring image generation speed is another way to evaluate performance. One of the most straightforward ways uses Stable Diffusion, a popular image-based AI model that allows users to easily convert text descriptions into complex visual representations.

With Stable Diffusion, users can quickly create and refine images from text prompts to achieve their desired output. When using an RTX GPU, these results can be generated faster than processing the AI model on a CPU or NPU.

That performance is even higher when using the TensorRT extension for the popular Automatic1111 interface. RTX users can generate images from prompts up to 2x faster with the SDXL Base checkpoint - significantly streamlining Stable Diffusion workflows.

ComfyUI, another popular Stable Diffusion user interface, added TensorRT acceleration last week. RTX users can now generate images from prompts up to 60% faster, and can even convert these images to videos using Stable Video Diffuson up to 70% faster with TensorRT.

TensorRT acceleration can be put to the test in the new UL Procyon AI Image Generation benchmark, which delivers speedups of 50% on a GeForce RTX 4080 SUPER GPU compared with the fastest non-TensorRT implementation.

TensorRT acceleration will soon be released for Stable Diffusion 3 - Stability AI's new, highly anticipated text-to-image model - boosting performance by 50%. Plus, the new TensorRT-Model Optimizer enables accelerating performance even further. This results in a 70% speedup compared with the non-TensorRT implementation, along with a 50% reduction in memory consumption.

Of course, seeing is believing - the true test is in the real-world use case of iterating on an original prompt. Users can refine image generation by tweaking prompts significantly faster on RTX GPUs, taking seconds per iteration compared with minutes on a Macbook Pro M3 Max. Plus, users get both speed and security with everything remaining private when running locally on an RTX-powered PC or workstation.

The Results Are in and Open Sourced But don't just take our word for it. The team of AI researchers and engineers behind the open-source Jan.ai recently integrated TensorRT-LLM into its local chatbot app, then tested these optimizations for themselves.

Source: Jan.ai The researchers tested its implementation of TensorRT-LLM against the open-source llama.cpp inference engine across a variety of GPUs and CPUs used by the community. They found that TensorRT is 30-70% faster than llam
LINK: https://blogs.nvidia.com/blog/ai-decoded-tops/...
See more stories from nvidia

North America Stories

15/02/2026

Live From NBA All-Star 2026: Entertainment Takes the Court in a Big Way

With new partnership between the league and NBC, workflows distinguish more between live, broadcast sound There'll be a lot new for the 75th NBA All-Star W...

15/02/2026

Live From NBA All-Star 2026: NBC Sports Director Pierre Moossa Previews NBC's Return to the Event

After 24-year absence, NBC Sports returns to NBA All-Star Weekend with unique ca...

15/02/2026

Live From NBA All-Star 2026: Peacock, NBC Sports Offer Viewers a Front-Row Seat With Courtside Live'

New to NBA coverage, the viewer experience offers several angles in addition to ...

15/02/2026

Live From NBA All-Star 2026: NBC Sports Returns With Plenty of Tech Toys in Tow

Coverage features 4X-slo-mo Supracam and Steadicam, Nucleus 4K cameras, closer play-by-play angle, 10 player mics NBC Sports is in the midst of its first NBA A...

14/02/2026

Cineverse Acquires TV Monetization Platform IndiCue

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

ESPN's Audiences for College Basketball On Track for Major Growth

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

TCL Display Technologies Deployed at Winter Olympics

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

Boston Conservatory Orchestra Helps Peter and Leonardo Dugan Complete Their Dream Piece

Boston Conservatory Orchestra Helps Peter and Leonardo Dugan Complete Their Dre...

13/02/2026

OBS Accelerates Shift to Cloud

Olympic Broadcasting Services (OBS) has provided an update on its adoption of the cloud as it continues on its journey to fully migrate to IT-based systems by 2...

13/02/2026

France Tlvisions Launches France 2 UHD with Dolby Vision and Dolby Atmos to Max out AC-4 Experiences for Winter Olympics Fans

France T l visions has successfully launched France 2 UHD featuring Dolby Vision...

13/02/2026

OBS Expands Athlete Moment,' Family Reunions to Capture Human Side of Winter Games

Partnering with Worldwide Olympic Partner TCL, OBS deploys connected Athlete Mom...

13/02/2026

Men's Figure Skating Photo Gallery

The men's figure skating long-form program is tonight, and it promises to be an exciting night for fans in the stands, fans at home, and even the production...

13/02/2026

Entertainment Takes the NBA Court in a Big Way

With new partnership between the league and NBC, workflows distinguish more between live, broadcast sound There'll be a lot new for the 75th NBA All-Star W...

13/02/2026

SVG GameDay, Episode 3: Sean Tabler - Producing Hockey in the City of Angels

In-venue and creative video staffers at the professional and collegiate level have one major thing in common: the intensity and attention to detail ramps up dur...

13/02/2026

Teradek Introduces RF-X, Revolutionizing Mission-Critical Signal Redundancy

Teradek announces the launch of RF-X Auto Switcher, a revolutionary appliance designed to deliver flawless, uncompromised signal integrity for the world's m...

13/02/2026

Synamedia & Globecast Selected for FA Cup Cloud Distribution

Globecast and Synamedia announces that Pitch International (Pitch), the leading London-based sports marketing agency, has gone live with cloud-based distributi...

13/02/2026

Ratings Roundup: NBC Sports' Legendary February Hits Record Viewership Levels

Ratings Roundup is a rundown of recent rating news and is derived from press rel...

13/02/2026

NBC Olympics' Amy Rosenfeld on the Drone Craze, Friends & Family Moments, Stamford's Role for Milano Cortina

Far from the action in the snow and on the ice, the team controls the production...

13/02/2026

2026 Daytona 500: FOX Sports' Mike Davies, George Grill on Working Within an IP-Based Compound, Solving the Ops Puzzle of the Super Bowl of Racing

The Daytona 500 is called The Super Bowl of Racing for a reason. Whether it's the culmination to five days of action on the track, the sheer size and scop...

13/02/2026

OBS Expands AI-Powered Content Workflows

For the Milano Cortina Games, Olympic Broadcasting Services (OBS) is delivering more than 6,500 hours of content, with more than 900 hours of live action, sprea...

13/02/2026

NBC Sports Director Pierre Moossa Previews NBC's First NBA All-Star Production in 24 Years

After 24-year absence, NBC Sports returns to NBA All-Star Weekend with unique ca...

13/02/2026

Film Festival Watch: 18 Sundance Institute-Supported Projects To Watch at the 2026 Berlin International Film Festival

By Jessica Herndon We may have just wrapped an unforgettable 2026 Sundance Film...

13/02/2026

Give Me the Backstory: Get to Know Amanda Kramer, the Writer-Director Behind By Design

By Jessica Herndon One of the most exciting things about the Sundance Film Fest...

13/02/2026

L3Harris Successfully Completes First Phase of P25 Transition for Florida SLERS

The upgrade to a Project 25 network provides state agencies communicating on the Statewide Law Enforcement Radio System flexibility to tailor the network to the...

13/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24...

Riedel Communications has officially opened a new office in Kuala Lumpur, Malaysia, marking a strategic expansion of its global Customer Success and IT software...

13/02/2026

ES Broadcast Hire duo celebrate 10-year anniversary with...

Two of ES Broadcast Hire's longest-serving employees recently celebrated a decade working for the company. Annie Breislin, Operations Manager, and Charles ...

13/02/2026

Disguise Opens Experience Center and Office in Atlanta

Disguise, the award-winning technology company powering global experiences, today unveils a new 8,000-square-foot office and Experience Center in Atlanta, creat...

13/02/2026

Mavis Expands External Camera Support with Accsoon SeeMo...

At BSC Expo 2026, Mavis announced full support for the Accsoon SeeMo series of iOS camera adapters across Mavis Camera and Mavis Monitor apps. This new integrat...

13/02/2026

Butcher Bird Studios Keeps Signals Flowing Seamlessly Acr...

Executing technically ambitious live streams, virtual productions, and immersive media today requires talent, creativity, and the right supporting technology. L...

13/02/2026

LTN makes key appointments and introduces new Technology...

Michal Miskin-Amir, Jonathan Stanton and Bobby Bond to lead technical advances amid surge in demand for LTN's IP video transport services as satellite capac...

13/02/2026

NATO Upgrades Broadcast Studio with Grass Valley Cameras

Grass Valley, the pioneering media and entertainment technology innovator, has won a competitive NATO-wide tender to provide the new camera system for NATO'...

13/02/2026

Digital Azul strengthens remote production strategy with...

Wireless IP intercom underpins agile, multi-location live production workflows Digital Azul, the independent production powerhouse specialising in complex liv...

13/02/2026

Actus Digital Sets a New Standard for QA Monitoring and C...

Actus Digital, a LiveU company, will unveil major new enhancements to its Actus X Intelligent Monitoring Platform at NAB Show (LiveU booth N1740), reinforcing i...

13/02/2026

FA Cup goes IP with Pitch International plus Synamedia an...

Globecast, a worldwide leader in broadcast services, and leading video software provider, Synamedia, today announced that Pitch International (Pitch), the leadi...

13/02/2026

Rai Selects Imagine Selenio Network Processor for IP Migration

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

CIMM Details Research Plans for 2026 and New Board Appointments

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Teradek Unveils RF-A Auto Switcher

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Spectrum Launches 'Invincible Wifi'

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Actus Digital to Introduce Actus X Platform Enhancements At NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Sennheiser Wireless Spectera Solution Tackles Super Bowl LX With Ease

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Nate Bargatze to Receive 2026 NAB Television Chairman's Award

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Chyron Merges Live Web Content and CG Graphics with PRIME 5.3

Chyron unveils PRIME 5.3, the latest software release of the company's powerful engine for live production graphics. PRIME 5.3 delivers the first official i...

12/02/2026

SVG New Sponsor Spotlight: Interra Systems' Anupama Anantharaman on Protecting Live Sports Quality Across IP and OTT Workflows

The vendor's VP of Product Management explains how quality assurance, monito...

12/02/2026

LTN Makes Key Appointments and Introduces New Technology Organization

LTN announces the appointment of three experienced executives to lead its new Technology organization: Michal Miskin-Amir as EVP and Head of Technology, Jonatha...

12/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24/7 Software and IT Support

Riedel Communications has officially opened a new office in Kuala Lumpur, Malays...

12/02/2026

NATO Upgrades Brussels HQ Broadcast Studio with Grass Valley LDX 135 Cameras

Grass Valley has won a competitive NATO-wide tender to provide the new camera system for NATO's main broadcast studio at its Brussels headquarters. The proj...

12/02/2026

Canon Announces Big Game Broadcast Lens Use Data

Canon U.S.A announces that the vast majority of broadcast lenses utilized on the NBC live broadcast for the Big Game between New England and Seattle on Sunday w...

12/02/2026

Three-Time Grammy-Winner Ludacris to Headline Performances at NBA All-Star 2026 in LA

The National Basketball Association (NBA) and NBC Sports announce the entertainm...

12/02/2026

IOC Awards Broadcast Rights in Middle East and North Africa to beIN MEDIA GROUP

The International Olympic Committee (IOC) announces that beIN MEDIA GROUP ( beIN ), the leading global sports, entertainment and media organisation, has secured...

12/02/2026

Big 12 Conference Unveils ASB GlassFloor for Upcoming Tournaments in March

The Big 12 Conference and ASB GlassFloor introduces a full LED video sports floor that will debut at the 2026 Phillips 66 Big 12 Men's and Women's Baske...