Sony Pixel Power calrec Sony

You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper

21/03/2024

Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by NVIDIA founder and CEO Jensen Huang with seven of the authors of the legendary research paper that introduced the aptly named transformer - a neural network architecture that went on to change the deep learning landscape and enable today's era of generative AI.

Everything that we're enjoying today can be traced back to that moment, Huang said to a packed room with hundreds of attendees, who heard him speak with the authors of Attention Is All You Need.

Sharing the stage for the first time, the research luminaries reflected on the factors that led to their original paper, which has been cited more than 100,000 times since it was first published and presented at the NeurIPS AI conference. They also discussed their latest projects and offered insights into future directions for the field of generative AI.

While they started as Google researchers, the collaborators are now spread across the industry, most as founders of their own AI companies.

We have a whole industry that is grateful for the work that you guys did, Huang said.

From L to R: Lukasz Kaiser, Noam Shazeer, Aidan Gomez, Jensen Huang, Llion Jones, Jakob Uszkoreit, Ashish Vaswani and Illia Polosukhin. Origins of the Transformer Model The research team initially sought to overcome the limitations of recurrent neural networks, or RNNs, which were then the state of the art for processing language data.

Noam Shazeer, cofounder and CEO of Character.AI, compared RNNs to the steam engine and transformers to the improved efficiency of internal combustion.

We could have done the industrial revolution on the steam engine, but it would just have been a pain, he said. Things went way, way better with internal combustion.

Now we're just waiting for the fusion, quipped Illia Polosukhin, cofounder of blockchain company NEAR Protocol.

The paper's title came from a realization that attention mechanisms - an element of neural networks that enable them to determine the relationship between different parts of input data - were the most critical component of their model's performance.

We had very recently started throwing bits of the model away, just to see how much worse it would get. And to our surprise it started getting better, said Llion Jones, cofounder and chief technology officer at Sakana AI.

Having a name as general as transformers spoke to the team's ambitions to build AI models that could process and transform every data type - including text, images, audio, tensors and biological data.

That North Star, it was there on day zero, and so it's been really exciting and gratifying to watch that come to fruition, said Aidan Gomez, cofounder and CEO of Cohere. We're actually seeing it happen now.

Packed house at the San Jose Convention Center. Envisioning the Road Ahead Adaptive computation, where a model adjusts how much computing power is used based on the complexity of a given problem, is a key factor the researchers see improving in future AI models.

It's really about spending the right amount of effort and ultimately energy on a given problem, said Jakob Uszkoreit, cofounder and CEO of biological software company Inceptive. You don't want to spend too much on a problem that's easy or too little on a problem that's hard.

A math problem like two plus two, for example, shouldn't be run through a trillion-parameter transformer model - it should run on a basic calculator, the group agreed.

They're also looking forward to the next generation of AI models.

I think the world needs something better than the transformer, said Gomez. I think all of us here hope it gets succeeded by something that will carry us to a new plateau of performance.

You don't want to miss these next 10 years, Huang said. Unbelievable new capabilities will be invented.

The conversation concluded with Huang presenting each researcher with a framed cover plate of the NVIDIA DGX-1 AI supercomputer, signed with the message, You transformed the world.

Jensen presents lead author Ashish Vaswani with a signed DGX-1 cover. There's still time to catch the session replay by registering for a virtual GTC pass - it's free.

To discover the latest in generative AI, watch Huang's GTC keynote address:
LINK: https://blogs.nvidia.com/blog/gtc-2024-transformer-ai-research-panel-j...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

14/02/2026

Cineverse Acquires TV Monetization Platform IndiCue

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

ESPN's Audiences for College Basketball On Track for Major Growth

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

TCL Display Technologies Deployed at Winter Olympics

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

Boston Conservatory Orchestra Helps Peter and Leonardo Dugan Complete Their Dream Piece

Boston Conservatory Orchestra Helps Peter and Leonardo Dugan Complete Their Dre...

13/02/2026

OBS Accelerates Shift to Cloud

Olympic Broadcasting Services (OBS) has provided an update on its adoption of the cloud as it continues on its journey to fully migrate to IT-based systems by 2...

13/02/2026

France Tlvisions Launches France 2 UHD with Dolby Vision and Dolby Atmos to Max out AC-4 Experiences for Winter Olympics Fans

France T l visions has successfully launched France 2 UHD featuring Dolby Vision...

13/02/2026

OBS Expands Athlete Moment,' Family Reunions to Capture Human Side of Winter Games

Partnering with Worldwide Olympic Partner TCL, OBS deploys connected Athlete Mom...

13/02/2026

Men's Figure Skating Photo Gallery

The men's figure skating long-form program is tonight, and it promises to be an exciting night for fans in the stands, fans at home, and even the production...

13/02/2026

Entertainment Takes the NBA Court in a Big Way

With new partnership between the league and NBC, workflows distinguish more between live, broadcast sound There'll be a lot new for the 75th NBA All-Star W...

13/02/2026

SVG GameDay, Episode 3: Sean Tabler - Producing Hockey in the City of Angels

In-venue and creative video staffers at the professional and collegiate level have one major thing in common: the intensity and attention to detail ramps up dur...

13/02/2026

Teradek Introduces RF-X, Revolutionizing Mission-Critical Signal Redundancy

Teradek announces the launch of RF-X Auto Switcher, a revolutionary appliance designed to deliver flawless, uncompromised signal integrity for the world's m...

13/02/2026

Synamedia & Globecast Selected for FA Cup Cloud Distribution

Globecast and Synamedia announces that Pitch International (Pitch), the leading London-based sports marketing agency, has gone live with cloud-based distributi...

13/02/2026

Ratings Roundup: NBC Sports' Legendary February Hits Record Viewership Levels

Ratings Roundup is a rundown of recent rating news and is derived from press rel...

13/02/2026

NBC Olympics' Amy Rosenfeld on the Drone Craze, Friends & Family Moments, Stamford's Role for Milano Cortina

Far from the action in the snow and on the ice, the team controls the production...

13/02/2026

2026 Daytona 500: FOX Sports' Mike Davies, George Grill on Working Within an IP-Based Compound, Solving the Ops Puzzle of the Super Bowl of Racing

The Daytona 500 is called The Super Bowl of Racing for a reason. Whether it's the culmination to five days of action on the track, the sheer size and scop...

13/02/2026

OBS Expands AI-Powered Content Workflows

For the Milano Cortina Games, Olympic Broadcasting Services (OBS) is delivering more than 6,500 hours of content, with more than 900 hours of live action, sprea...

13/02/2026

NBC Sports Director Pierre Moossa Previews NBC's First NBA All-Star Production in 24 Years

After 24-year absence, NBC Sports returns to NBA All-Star Weekend with unique ca...

13/02/2026

Film Festival Watch: 18 Sundance Institute-Supported Projects To Watch at the 2026 Berlin International Film Festival

By Jessica Herndon We may have just wrapped an unforgettable 2026 Sundance Film...

13/02/2026

Give Me the Backstory: Get to Know Amanda Kramer, the Writer-Director Behind By Design

By Jessica Herndon One of the most exciting things about the Sundance Film Fest...

13/02/2026

Women in Podcasting Craft New Connections at Spotify's Galentine's Day Celebration in LA

This Wednesday in Los Angeles, Spotify brought together a group of podcast creat...

13/02/2026

Spotify and LoveShackFancy Bring Galentine's Glam to NYC, Featuring Special Performance by Joshua Basset

Yesterday, Spotify and LoveShackFancy hosted a Galentine's and Gents Lunch a...

13/02/2026

L3Harris Successfully Completes First Phase of P25 Transition for Florida SLERS

The upgrade to a Project 25 network provides state agencies communicating on the Statewide Law Enforcement Radio System flexibility to tailor the network to the...

13/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24...

Riedel Communications has officially opened a new office in Kuala Lumpur, Malaysia, marking a strategic expansion of its global Customer Success and IT software...

13/02/2026

ES Broadcast Hire duo celebrate 10-year anniversary with...

Two of ES Broadcast Hire's longest-serving employees recently celebrated a decade working for the company. Annie Breislin, Operations Manager, and Charles ...

13/02/2026

Disguise Opens Experience Center and Office in Atlanta

Disguise, the award-winning technology company powering global experiences, today unveils a new 8,000-square-foot office and Experience Center in Atlanta, creat...

13/02/2026

Mavis Expands External Camera Support with Accsoon SeeMo...

At BSC Expo 2026, Mavis announced full support for the Accsoon SeeMo series of iOS camera adapters across Mavis Camera and Mavis Monitor apps. This new integrat...

13/02/2026

Butcher Bird Studios Keeps Signals Flowing Seamlessly Acr...

Executing technically ambitious live streams, virtual productions, and immersive media today requires talent, creativity, and the right supporting technology. L...

13/02/2026

LTN makes key appointments and introduces new Technology...

Michal Miskin-Amir, Jonathan Stanton and Bobby Bond to lead technical advances amid surge in demand for LTN's IP video transport services as satellite capac...

13/02/2026

NATO Upgrades Broadcast Studio with Grass Valley Cameras

Grass Valley, the pioneering media and entertainment technology innovator, has won a competitive NATO-wide tender to provide the new camera system for NATO'...

13/02/2026

Digital Azul strengthens remote production strategy with...

Wireless IP intercom underpins agile, multi-location live production workflows Digital Azul, the independent production powerhouse specialising in complex liv...

13/02/2026

Actus Digital Sets a New Standard for QA Monitoring and C...

Actus Digital, a LiveU company, will unveil major new enhancements to its Actus X Intelligent Monitoring Platform at NAB Show (LiveU booth N1740), reinforcing i...

13/02/2026

FA Cup goes IP with Pitch International plus Synamedia an...

Globecast, a worldwide leader in broadcast services, and leading video software provider, Synamedia, today announced that Pitch International (Pitch), the leadi...

13/02/2026

Rai Selects Imagine Selenio Network Processor for IP Migration

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

CIMM Details Research Plans for 2026 and New Board Appointments

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Teradek Unveils RF-A Auto Switcher

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Spectrum Launches 'Invincible Wifi'

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Actus Digital to Introduce Actus X Platform Enhancements At NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Sennheiser Wireless Spectera Solution Tackles Super Bowl LX With Ease

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Nate Bargatze to Receive 2026 NAB Television Chairman's Award

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

UKTV Highlights Saturday February 14th - Friday February 20th 2026

What can I watch on UKTV this week?What can I stream on U this week? This guide highlights romantic dramas for Valentine's Day, alternative relationship t...

13/02/2026

New RT series tells stranger-than-fiction stories of Irish con artists

New RT series tells stranger-than-fiction stories of Irish con artists Swindlers airs Wednesday 18 February, 9.35pm on RT One and RT Player Swindlers, a...

12/02/2026

Chyron Merges Live Web Content and CG Graphics with PRIME 5.3

Chyron unveils PRIME 5.3, the latest software release of the company's powerful engine for live production graphics. PRIME 5.3 delivers the first official i...

12/02/2026

SVG New Sponsor Spotlight: Interra Systems' Anupama Anantharaman on Protecting Live Sports Quality Across IP and OTT Workflows

The vendor's VP of Product Management explains how quality assurance, monito...

12/02/2026

LTN Makes Key Appointments and Introduces New Technology Organization

LTN announces the appointment of three experienced executives to lead its new Technology organization: Michal Miskin-Amir as EVP and Head of Technology, Jonatha...

12/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24/7 Software and IT Support

Riedel Communications has officially opened a new office in Kuala Lumpur, Malays...