Sony Pixel Power calrec Sony

NVIDIA Takes Inference to New Heights Across MLPerf Tests

05/04/2023

MLPerf remains the definitive measurement for AI performance as an independent, third-party benchmark. NVIDIA's AI platform has consistently shown leadership across both training and inference since the inception of MLPerf, including the MLPerf Inference 3.0 benchmarks released today.

Three years ago when we introduced A100, the AI world was dominated by computer vision. Generative AI has arrived, said NVIDIA founder and CEO Jensen Huang.

This is exactly why we built Hopper, specifically optimized for GPT with the Transformer Engine. Today's MLPerf 3.0 highlights Hopper delivering 4x more performance than A100.

The next level of Generative AI requires new AI infrastructure to train large language models with great energy efficiency. Customers are ramping Hopper at scale, building AI infrastructure with tens of thousands of Hopper GPUs connected by NVIDIA NVLink and InfiniBand.

The industry is working hard on new advances in safe and trustworthy Generative AI. Hopper is enabling this essential work, he said.

The latest MLPerf results show NVIDIA taking AI inference to new levels of performance and efficiency from the cloud to the edge.

Specifically, NVIDIA H100 Tensor Core GPUs running in DGX H100 systems delivered the highest performance in every test of AI inference, the job of running neural networks in production. Thanks to software optimizations, the GPUs delivered up to 54% performance gains from their debut in September.

In healthcare, H100 GPUs delivered a 31% performance increase since September on 3D-UNet, the MLPerf benchmark for medical imaging.

Powered by its Transformer Engine, the H100 GPU, based on the Hopper architecture, excelled on BERT, a transformer-based large language model that paved the way for today's broad use of generative AI.

Generative AI lets users quickly create text, images, 3D models and more. It's a capability companies from startups to cloud service providers are rapidly adopting to enable new business models and accelerate existing ones.

Hundreds of millions of people are now using generative AI tools like ChatGPT - also a transformer model - expecting instant responses.

At this iPhone moment of AI, performance on inference is vital. Deep learning is now being deployed nearly everywhere, driving an insatiable need for inference performance from factory floors to online recommendation systems.

L4 GPUs Speed Out of the Gate NVIDIA L4 Tensor Core GPUs made their debut in the MLPerf tests at over 3x the speed of prior-generation T4 GPUs. Packaged in a low-profile form factor, these accelerators are designed to deliver high throughput and low latency in almost any server.

L4 GPUs ran all MLPerf workloads. Thanks to their support for the key FP8 format, their results were particularly stunning on the performance-hungry BERT model.

In addition to stellar AI performance, L4 GPUs deliver up to 10x faster image decode, up to 3.2x faster video processing and over 4x faster graphics and real-time rendering performance.

Announced two weeks ago at GTC, these accelerators are already available from major systems makers and cloud service providers. L4 GPUs are the latest addition to NVIDIA's portfolio of AI inference platforms launched at GTC.

Software, Networks Shine in System Test NVIDIA's full-stack AI platform showed its leadership in a new MLPerf test.

The so-called network-division benchmark streams data to a remote inference server. It reflects the popular scenario of enterprise users running AI jobs in the cloud with data stored behind corporate firewalls.

On BERT, remote NVIDIA DGX A100 systems delivered up to 96% of their maximum local performance, slowed in part because they needed to wait for CPUs to complete some tasks. On the ResNet-50 test for computer vision, handled solely by GPUs, they hit the full 100%.

Both results are thanks, in large part, to NVIDIA Quantum Infiniband networking, NVIDIA ConnectX SmartNICs and software such as NVIDIA GPUDirect.

Orin Shows 3.2x Gains at the Edge Separately, the NVIDIA Jetson AGX Orin system-on-module delivered gains of up to 63% in energy efficiency and 81% in performance compared with its results a year ago. Jetson AGX Orin supplies inference when AI is needed in confined spaces at low power levels, including on systems powered by batteries.

For applications needing even smaller modules drawing less power, the Jetson Orin NX 16G shined in its debut in the benchmarks. It delivered up to 3.2x the performance of the prior-generation Jetson Xavier NX processor.

A Broad NVIDIA AI Ecosystem The MLPerf results show NVIDIA AI is backed by the industry's broadest ecosystem in machine learning.

Ten companies submitted results on the NVIDIA platform in this round. They came from the Microsoft Azure cloud service and system makers including ASUS, Dell Technologies, GIGABYTE, H3C, Lenovo, Nettrix, Supermicro and xFusion.

Their work shows users can get great performance with NVIDIA AI both in the cloud and in servers running in their own data centers.

NVIDIA partners participate in MLPerf because they know it's a valuable tool for customers evaluating AI platforms and vendors. Results in the latest round demonstrate that the performance they deliver today will grow with the NVIDIA platform.

Users Need Versatile Performance NVIDIA AI is the only platform to run all MLPerf inference workloads and scenarios in data center and edge computing. Its versatile performance and efficiency make users the real winners.

Real-world applications typically employ many neural networks of different kinds that often need to deliver answers in real time.

For example, an AI application may need to understand a user's spoken request, classify an image, make a recommendation and then deliver a response as a spoken message in a human-sounding voice. Each step requires a different type
LINK: https://blogs.nvidia.com/blog/2023/04/05/inference-mlperf-ai/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

13/02/2026

Rai Selects Imagine Selenio Network Processor for IP Migration

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

CIMM Details Research Plans for 2026 and New Board Appointments

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Teradek Unveils RF-A Auto Switcher

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Spectrum Launches 'Invincible Wifi'

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Actus Digital to Introduce Actus X Platform Enhancements At NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Sennheiser Wireless Spectera Solution Tackles Super Bowl LX With Ease

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Nate Bargatze to Receive 2026 NAB Television Chairman's Award

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Chyron Merges Live Web Content and CG Graphics with PRIME 5.3

Chyron unveils PRIME 5.3, the latest software release of the company's powerful engine for live production graphics. PRIME 5.3 delivers the first official i...

12/02/2026

SVG New Sponsor Spotlight: Interra Systems' Anupama Anantharaman on Protecting Live Sports Quality Across IP and OTT Workflows

The vendor's VP of Product Management explains how quality assurance, monito...

12/02/2026

LTN Makes Key Appointments and Introduces New Technology Organization

LTN announces the appointment of three experienced executives to lead its new Technology organization: Michal Miskin-Amir as EVP and Head of Technology, Jonatha...

12/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24/7 Software and IT Support

Riedel Communications has officially opened a new office in Kuala Lumpur, Malays...

12/02/2026

NATO Upgrades Brussels HQ Broadcast Studio with Grass Valley LDX 135 Cameras

Grass Valley has won a competitive NATO-wide tender to provide the new camera system for NATO's main broadcast studio at its Brussels headquarters. The proj...

12/02/2026

Canon Announces Big Game Broadcast Lens Use Data

Canon U.S.A announces that the vast majority of broadcast lenses utilized on the NBC live broadcast for the Big Game between New England and Seattle on Sunday w...

12/02/2026

Three-Time Grammy-Winner Ludacris to Headline Performances at NBA All-Star 2026 in LA

The National Basketball Association (NBA) and NBC Sports announce the entertainm...

12/02/2026

IOC Awards Broadcast Rights in Middle East and North Africa to beIN MEDIA GROUP

The International Olympic Committee (IOC) announces that beIN MEDIA GROUP ( beIN ), the leading global sports, entertainment and media organisation, has secured...

12/02/2026

Big 12 Conference Unveils ASB GlassFloor for Upcoming Tournaments in March

The Big 12 Conference and ASB GlassFloor introduces a full LED video sports floor that will debut at the 2026 Phillips 66 Big 12 Men's and Women's Baske...

12/02/2026

TNDV Showcases Aspiration35 at National Religious Broadcasters Convention 2026

Continuing its commitment to serving the faith-based broadcast and live event community, mobile production company TNDV, a division of Live Media Group, will hi...

12/02/2026

Marshall Electronics POVs Power Hidden-Camera Investigations on German TV Series, ACHTUNG ABZOCKE

The production team of the long-running German investigative series Achtung Abz...

12/02/2026

Vizrt Launches Sports Production Bundles to Empower US Students to Produce Like the Pros

Vizrt announces the launch of four Campus Stadium Production Bundles, designed t...

12/02/2026

LiveU Spotlights Three Broadcast Priorities with Digital-First, Workflow Automation, IP Contribution Resilience

At NAB Show, LiveU will showcase its broadest IP-video EcoSystem to date, design...

12/02/2026

Follow the Money, Episode 5: Analyzing Sports Media Deals With Sam McCleery and PwC's Lori Bistis

Welcome to the Sports Video Group's new interview series, Follow the Money, ...

12/02/2026

NBC Sports Engineers a Super Bowl Transmission Plan That Reached From Alcatraz to Stamford - and Out Onto San Francisco Bay

400 Gbps of bandwidth, layered redundancy, and mobile-first connectivity powered...

12/02/2026

5 Ideas for Bringing the Music You Love Into Valentine's Day

Valentine's Day often comes with a soundtrack. In fact, Spotify data shows that more people used Blend, our shared playlist feature, on February 14, 2025, t...

12/02/2026

Prompted Playlist in Beta Coming to Premium Listeners in More Markets

Some days you want your music to reflect a specific feeling, memory, or vibe that goes beyond a single artist or genre. You want to do more than listen. You wan...

12/02/2026

Our Medicine S2: Frontline Medicine Through A Blak Lens

Our Medicine S2: Frontline Medicine Through A Blak Lens 12 February, 2026 Media releases A Bigger, Bolder Second Series showcasing First Nations Frontline ...

12/02/2026

L3Harris' VAMPIRE System Successfully Fires Thales' Belgian-Made 70 mm Rockets

L3Harris' VAMPIRE system fires Thales Belgian-made 70 MM rocket from an FZ60...

12/02/2026

LTN Names 3 Executives to Lead Technology Group

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

On the Ice, There's a Third Team at Work

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Mo Rocca to Receive the 2026 LABF Insight Award at NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Marshall Electronics POVs Power Hidden Camera Investigat...

The production team of the long-running German investigative series Achtung Abzocke recently upgraded its cameras for the show's 12th season. The objectiv...

12/02/2026

Bitmovin Appoints Ian Baglow as Co-Chief Executive Office...

Leading provider of video streaming solutions, Bitmovin, has appointed Ian Baglow as Co-CEO alongside existing CEO and Co-Founder Stefan Lederer. Under this str...

12/02/2026

Vizrt Launches Sports Production Bundles to Empower US St...

Vizrt, a leading viewer engagement platform and a trusted expert in live production technologies, today announces the launch of four Campus Stadium Production B...

12/02/2026

Ailanto and Cubbit launch sovereign cloud storage for Swi...

Strategic agreement to deliver S3 cloud storage in Switzerland with full data sovereignty and local control including at the level of individual cantons plu...

12/02/2026

Mad About Video counts on Lightware MX2 matrix switcher a...

Mad About Video is a leading specialist in video for live events and installations throughout Malta. In operation since 2011, it has evolved from a company focu...

12/02/2026

JAGGAER supports Betsson Group in further strengthening i...

JAGGAER, a global leader in digital procurement and supplier collaboration solutions, today announced the successful delivery of a procurement digitalization pr...

12/02/2026

LiveU Spotlights Three Broadcast Priorities at NAB Show 2...

At NAB Show, LiveU will showcase its broadest IP-video EcoSystem to date, designed to help broadcasters and content creators embrace digital first operations, d...

12/02/2026

Spectrum News Acquires New England Cable News

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Vizrt Unveils Campus Stadium Production Bundles

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Hulu + Live TV Adds Fubo Sports Network to Channel Line-up

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

FCC To Hold Open Commission Meeting on Feb. 18

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Ralph M. Oakley to Receive NAB's Chuck Sherman TV Leadership Award

Share Copy link Facebook X Linkedin Bluesky Email...

12/02/2026

Sky Original drama Under Salt Marsh hits 1.8 million viewers in its first seven days

The six-part crime drama, created by Claire Oakley and produced by Little Door P...

12/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24/7 Software and IT Support

Wuppertal February 12, 2026 Riedel Opens Kuala Lumpur Office to Strengthen Glo...

12/02/2026

Netflix unveils the trailer for 'That Night'

Back to All News Netflix unveils the trailer for That Night Entertainment 12 February 2026 GlobalSpain Link copied to clipboard WATCH THE TRAILER DOWNLOA...