Sony Pixel Power calrec Sony

You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper

21/03/2024

Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by NVIDIA founder and CEO Jensen Huang with seven of the authors of the legendary research paper that introduced the aptly named transformer - a neural network architecture that went on to change the deep learning landscape and enable today's era of generative AI.

Everything that we're enjoying today can be traced back to that moment, Huang said to a packed room with hundreds of attendees, who heard him speak with the authors of Attention Is All You Need.

Sharing the stage for the first time, the research luminaries reflected on the factors that led to their original paper, which has been cited more than 100,000 times since it was first published and presented at the NeurIPS AI conference. They also discussed their latest projects and offered insights into future directions for the field of generative AI.

While they started as Google researchers, the collaborators are now spread across the industry, most as founders of their own AI companies.

We have a whole industry that is grateful for the work that you guys did, Huang said.

From L to R: Lukasz Kaiser, Noam Shazeer, Aidan Gomez, Jensen Huang, Llion Jones, Jakob Uszkoreit, Ashish Vaswani and Illia Polosukhin. Origins of the Transformer Model The research team initially sought to overcome the limitations of recurrent neural networks, or RNNs, which were then the state of the art for processing language data.

Noam Shazeer, cofounder and CEO of Character.AI, compared RNNs to the steam engine and transformers to the improved efficiency of internal combustion.

We could have done the industrial revolution on the steam engine, but it would just have been a pain, he said. Things went way, way better with internal combustion.

Now we're just waiting for the fusion, quipped Illia Polosukhin, cofounder of blockchain company NEAR Protocol.

The paper's title came from a realization that attention mechanisms - an element of neural networks that enable them to determine the relationship between different parts of input data - were the most critical component of their model's performance.

We had very recently started throwing bits of the model away, just to see how much worse it would get. And to our surprise it started getting better, said Llion Jones, cofounder and chief technology officer at Sakana AI.

Having a name as general as transformers spoke to the team's ambitions to build AI models that could process and transform every data type - including text, images, audio, tensors and biological data.

That North Star, it was there on day zero, and so it's been really exciting and gratifying to watch that come to fruition, said Aidan Gomez, cofounder and CEO of Cohere. We're actually seeing it happen now.

Packed house at the San Jose Convention Center. Envisioning the Road Ahead Adaptive computation, where a model adjusts how much computing power is used based on the complexity of a given problem, is a key factor the researchers see improving in future AI models.

It's really about spending the right amount of effort and ultimately energy on a given problem, said Jakob Uszkoreit, cofounder and CEO of biological software company Inceptive. You don't want to spend too much on a problem that's easy or too little on a problem that's hard.

A math problem like two plus two, for example, shouldn't be run through a trillion-parameter transformer model - it should run on a basic calculator, the group agreed.

They're also looking forward to the next generation of AI models.

I think the world needs something better than the transformer, said Gomez. I think all of us here hope it gets succeeded by something that will carry us to a new plateau of performance.

You don't want to miss these next 10 years, Huang said. Unbelievable new capabilities will be invented.

The conversation concluded with Huang presenting each researcher with a framed cover plate of the NVIDIA DGX-1 AI supercomputer, signed with the message, You transformed the world.

Jensen presents lead author Ashish Vaswani with a signed DGX-1 cover. There's still time to catch the session replay by registering for a virtual GTC pass - it's free.

To discover the latest in generative AI, watch Huang's GTC keynote address:
LINK: https://blogs.nvidia.com/blog/gtc-2024-transformer-ai-research-panel-j...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

04/08/2026

Dalet Announces Commercial Availability of Dalia, Bringing Media-Aware Agentic AI to Enterprise Productions

Dalet, a leading technology and service provider for media-rich organizations, t...

04/07/2026

Detective Conan: Fallen Angel of the Highway Opens in Dolby Cinemas Across Japan, Presented in Dolby Atmos and Dolby ...

April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

15/05/2026

IAB Releases Campaign Data Standards 1.0 for Public Comment

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

ARRI Expands Management Board

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Gray Media Names Joanie Vasiliadis SVP of Transformation

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Study: Data and Measurement Problems Reduce CTV Ad Budgets

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Upfronts: WBD Expands Advanced Ad Capabilities and AI Ad Tech

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

VLAST Powers PLAVEs Asia Tour Encore with AJA Gear

Delivering a live, arena-scale production of a massively popular band is no small feat. Between expansive in-arena LED walls and a global live stream fed to onl...

15/05/2026

Sun Broadcast Futureproofs Dayalbaghs Multimedia Van with...

Connection is the heartbeat of any strong community, and with live streaming becoming more accessible in the modern era, it's much easier for faith-based or...

15/05/2026

Disguise and Creative Technology Power Eurovision for the...

Powered by GX 3 media servers, optimised IP-VFC workflows and on-site engineering expertise, the production delivers high-performance visuals for one of the wor...

14/05/2026

Sweetwater and Airstream Unveil Mobile Dolby Atmos Recording Studio

Sweetwater and Airstream have announced a custom-built Dolby Atmos mobile recording studio inside an Airstream trailer, set to tour music festivals, schools, tr...

14/05/2026

American Association of Professional Baseball Expands Broadcast Distribution for 2026 Season

The American Association of Professional Baseball (AAPB) has announced a new par...

14/05/2026

ESPN to Establish Week-Long Super Bowl LXI Broadcast Center on Santa Monica Beach

ESPN has announced plans to transform Santa Monica Beach into a broadcast hub du...

14/05/2026

Amagi Announces Major Upgrade to CLOUDPORT Cloud Broadcast Platform

Amagi has announced a significant update to Amagi CLOUDPORT, its cloud-based broadcast playout platform. The update includes 250-plus features shipped in FY25-2...

14/05/2026

Clear-Com to Showcase New Products and Platform Updates at InfoComm 2026

Clear-Com will exhibit at InfoComm 2026 (Booth N7005, June 17-19, Las Vegas Convention Center), introducing a new product that builds on Arcadia Central Station...

14/05/2026

Ikegami to Exhibit at BroadcastAsia 2026 with Two New Viewfinder Premieres

Ikegami will exhibit at BroadcastAsia 2026 (Stand 5D3-1, Singapore Expo, May 20-22), introducing two new viewfinders alongside its existing camera, control, and...

14/05/2026

dB Broadcast Delivers IP-Based OB Trucks for Cloudbass Featuring Grass Valley LDX 100 Cameras

Grass Valley has announced that dB Broadcast has delivered new IP-based outside ...

14/05/2026

NAGRAVISION and WPBSA Launch Play Snooker Digital Platform

NAGRAVISION, a Kudelski Group company, has announced a partnership with the World Professional Billiards and Snooker Association (WPBSA) to launch Play Snooker,...

14/05/2026

Belden To Acquire RUCKUS Networks for Approximately $1.85 Billion

Belden Inc. has announced a definitive agreement to acquire RUCKUS Networks from Vistance Networks for approximately $1.85 billion. The transaction has been app...

14/05/2026

NVIDIA Releases Content Localization Blueprint for AI-Assisted Dubbing and Speech Translation

NVIDIA has released the Content Localization Blueprint, a modular reference arch...

14/05/2026

Disney+ To Stream Banana Bowl Championship Live This October

Disney has announced that Disney will be the exclusive U.S. streaming home of the Banana Bowl, the Banana Ball league season championship, streaming live this ...

14/05/2026

Arkona and Manifold To Exhibit on Magna Systems Stand at BroadcastAsia 2026

Arkona technologies and technology partner manifold will demonstrate their production solutions on the Magna Systems and Engineering stand (Booth 5D1-1) at Broa...

14/05/2026

Haivision Webinar: New Broadcast Contribution Products Featuring Minor League Baseball Case Study

Haivision will host a webinar on Thursday, May 21 at 10 a.m. ET / 4 p.m. CET cov...

14/05/2026

The CW Network and ESPN Announce ACC Sublicense Agreement Through 2030-31 Season

The CW Network and ESPN have announced a sublicense broadcast agreement for The CW to televise ACC football and men's and women's college basketball gam...

14/05/2026

Detroit Pistons Ink Local Media Rights Deal With Scripps Sports

The agreement marks Scripps Sports' first NBA local rights deal...

14/05/2026

Report: Broadcasting Among Hardest-Hit Industries as AI Reshapes the Workforce

A new report from education-technology company Wiingy testing post-ChatGPT predictions against three years of real-world data has identified broadcasting as one...

14/05/2026

Madonna, Shakira, and BTS to Headline First-Ever FIFA World Cup Final Halftime Show

Global Citizen and FIFA have announced that Madonna, Shakira, and BTS will headl...

14/05/2026

Sundance Institute Names 2026 Episodic Lab Fellows

LOS ANGELES, CA, May 14, 2026 - The nonprofit Sundance Institute announced today the cohort selected for the 2026 Episodic Lab program, taking place at Dunaway ...

14/05/2026

Spotify Expands Music Access for Young Listeners, Extending Managed Accounts to Free Tier

At Spotify, we're focused on making every listening experience feel intentio...

14/05/2026

Spotify Brings Nashville's Songwriting Community Together for Mental Health Summit

Spotify recently welcomed songwriters, artists, executives, and music students t...

14/05/2026

Sonuscore update Lux Orchestral Strings

New articulations, ostinatos, Motion Scoring Articulation Sets & more Sonuscore's flagship cinematic string library has just been treated to a significa...

14/05/2026

Pr Recording & Residence

World-class studio opens on T rkiye's Aegean coast P r Recording & Residence have announced their official opening, introducing a new world-class reside...

14/05/2026

Nugen Audio update DialogCheck

Now supports channel layouts up to 9.1.6 Nugen Audio have just released an update for their AI-powered dialogue intelligibility and compliance tool. Set to ...

14/05/2026

UVI release Orchestral Suite 2

New recordings & one-key chord tool UVI have just announced the release of Orchestral Suite 2, a ground-up redesign of their all-in-one symphonic orchestra ...

14/05/2026

Rob Papen launch eXplorer 11

Two new arrivals & expanded factory content Rob Papen's all-encompassing plug-in and virtual instrument collection has just been treated to another upda...

14/05/2026

FSK Audio release QRonicle

Embed QR codes into DAW sessions FSK Audio's latest plug-in doesn't process audio, but serves as an organisational tool that allows QR codes to be e...

14/05/2026

SBS Board appoints Jane Palfreyman Managing Director

SBS Board appoints Jane Palfreyman Managing Director 13 May, 2026 Media releases The Special Broadcasting Service (SBS) Board of Directors is pleased to an...

14/05/2026

Transforming bold ideas into market-ready productions: Digital Originals returns

Transforming bold ideas into market-ready productions: Digital Originals returns 14 May, 2026 Media releases SBS, NITV and Screen Australia have announced ...

14/05/2026

Australia Uncovered Returns to SBS with Bold New Season Featuring John Safran on Free Speech and Powerful Stories of Justice and Survival

Australia Uncovered Returns to SBS with Bold New Season Featuring John Safran on...

14/05/2026

Rohde & Schwarz transforms spectrum complexity into situational awareness and effective countermeasures at AOC Europe 2026

Rohde & Schwarz transforms spectrum complexity into situational awareness and ef...

14/05/2026

Rohde & Schwarz and Quantum Systems join forces to redefine EW and C-UAS-enabled uncrewed operations

Rohde & Schwarz and Quantum Systems join forces to redefine EW and C-UAS-enabled...

14/05/2026

Rohde & Schwarz showcases STANAG aligned ARDRONIS Counter UAS capability at NATO Technical Interoperability Exercise 2026

Rohde & Schwarz showcases STANAG aligned ARDRONIS Counter UAS capability at NATO...

14/05/2026

Code of Silence wins Best Drama Series at the 2026 BAFTA's!

Code of Silence has won the BAFTA for Best Drama Series at Sunday night's ceremony at the Royal Festival Hall. The series, starring Rose Ayling-Ellis and w...

14/05/2026

The Wraith Shield Advantage: Transforming L3Harris Radios into AI-Enabled Counter-UAS Sensors

Soldiers equipped with Falcon IV radios will soon gain a sense-and-protect capa...

14/05/2026

Getting into the Space Nuclear Power Game with Next-Generation Technology

Artists concept of the L3Harris Next Gen RTG in flight configuration, designed to provide 250 watts of reliable power for decades-long missions in deep space....

14/05/2026

Vivid Broadcast builds remote production network around Calrec

Vivid Broadcast was embracing remote production long before it became the industry norm. Now, with Calrec's True Control 2.0-enabled Argo M and Type R conso...

14/05/2026

Nielsen data shows NZ vehicle advertisers are shifting gears as fuel pressures make EVs and hybrids an increasingly attractive option

Car ad spend rises sharply in March as more auto buyers turn to electric, hybrid...

14/05/2026

Is This the Year for Agentic AI's Breakout in Broadcast?

Share Copy link Facebook X Linkedin Bluesky Email...