
Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by NVIDIA founder and CEO Jensen Huang with seven of the authors of the legendary research paper that introduced the aptly named transformer - a neural network architecture that went on to change the deep learning landscape and enable today's era of generative AI.
Everything that we're enjoying today can be traced back to that moment, Huang said to a packed room with hundreds of attendees, who heard him speak with the authors of Attention Is All You Need.
Sharing the stage for the first time, the research luminaries reflected on the factors that led to their original paper, which has been cited more than 100,000 times since it was first published and presented at the NeurIPS AI conference. They also discussed their latest projects and offered insights into future directions for the field of generative AI.
While they started as Google researchers, the collaborators are now spread across the industry, most as founders of their own AI companies.
We have a whole industry that is grateful for the work that you guys did, Huang said.
From L to R: Lukasz Kaiser, Noam Shazeer, Aidan Gomez, Jensen Huang, Llion Jones, Jakob Uszkoreit, Ashish Vaswani and Illia Polosukhin. Origins of the Transformer Model The research team initially sought to overcome the limitations of recurrent neural networks, or RNNs, which were then the state of the art for processing language data.
Noam Shazeer, cofounder and CEO of Character.AI, compared RNNs to the steam engine and transformers to the improved efficiency of internal combustion.
We could have done the industrial revolution on the steam engine, but it would just have been a pain, he said. Things went way, way better with internal combustion.
Now we're just waiting for the fusion, quipped Illia Polosukhin, cofounder of blockchain company NEAR Protocol.
The paper's title came from a realization that attention mechanisms - an element of neural networks that enable them to determine the relationship between different parts of input data - were the most critical component of their model's performance.
We had very recently started throwing bits of the model away, just to see how much worse it would get. And to our surprise it started getting better, said Llion Jones, cofounder and chief technology officer at Sakana AI.
Having a name as general as transformers spoke to the team's ambitions to build AI models that could process and transform every data type - including text, images, audio, tensors and biological data.
That North Star, it was there on day zero, and so it's been really exciting and gratifying to watch that come to fruition, said Aidan Gomez, cofounder and CEO of Cohere. We're actually seeing it happen now.
Packed house at the San Jose Convention Center. Envisioning the Road Ahead Adaptive computation, where a model adjusts how much computing power is used based on the complexity of a given problem, is a key factor the researchers see improving in future AI models.
It's really about spending the right amount of effort and ultimately energy on a given problem, said Jakob Uszkoreit, cofounder and CEO of biological software company Inceptive. You don't want to spend too much on a problem that's easy or too little on a problem that's hard.
A math problem like two plus two, for example, shouldn't be run through a trillion-parameter transformer model - it should run on a basic calculator, the group agreed.
They're also looking forward to the next generation of AI models.
I think the world needs something better than the transformer, said Gomez. I think all of us here hope it gets succeeded by something that will carry us to a new plateau of performance.
You don't want to miss these next 10 years, Huang said. Unbelievable new capabilities will be invented.
The conversation concluded with Huang presenting each researcher with a framed cover plate of the NVIDIA DGX-1 AI supercomputer, signed with the message, You transformed the world.
Jensen presents lead author Ashish Vaswani with a signed DGX-1 cover. There's still time to catch the session replay by registering for a virtual GTC pass - it's free.
To discover the latest in generative AI, watch Huang's GTC keynote address:
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
01/06/2026
January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026
Throughout the week, Dolby brings to life the latest innovatio...
02/05/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
01/05/2026
January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...
01/04/2026
January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION
Douyin Users Can Now Create And Share Videos With Stun...
08/02/2026
A look inside the tech, tools, and the team that make the Super Bowl into true e...
08/02/2026
Studio programming is leveraging multiple sets at Levi's Stadium in Santa Cl...
08/02/2026
The annual production is a partnership between the NFL and the host venue
NFL f...
08/02/2026
A software-defined IP backbone and centralized signal control hub redefine champ...
08/02/2026
Networks continue to raise the bar when it comes to producing an eye-catching pr...
08/02/2026
NEP's TFC Fabric empowers NBC to treat a 22-Unit production as one seamless ...
08/02/2026
Game coverage will feature nearly 100 cameras, a deep well of replay channels, a...
08/02/2026
The Alum Behind the Sound of the Super Bowl Joshua Sutherland BM '19 was recently interviewed by the Boston Globe about his role as music supervisor for t...
08/02/2026
08 Feb 2026
VEON and JazzWorld Launch Invest in Pakistan, NOW! Inviting Inter...
07/02/2026
X Games will host a live event taking place Thursday, March 12 at Cosm Los Angeles, marking the first-ever draft in X Games history and the official launch poin...
07/02/2026
Onsite experts and shop inside the stadium are providing quick solutions
Before any live event production, there is a high probability that unpredictable issue...
07/02/2026
Three Sony HDC-F5500s will capture this look on the broadcaster's concourse ...
07/02/2026
Inclusion in control-room revamp proves pivotal to supporting this tentpole even...
07/02/2026
The digital giant also produced episodes of The Edge with Micah Parsons' be...
07/02/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
07/02/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
07/02/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
06/02/2026
Appear, which specializes in live production technology, announces the appointme...
06/02/2026
Baller League US announces CBS Sports and its 24/7 soccer streaming channel CBS Sports Golazo Network will air the league's programming in the United States...
06/02/2026
Gravity Media, which concentrates in production, content, media services, and fa...
06/02/2026
The Alliance for IP Media Solutions (AIMS), together with the Video Services Forum (VSF), the Advanced Media Workflow Association (AMWA), and the European Broad...
06/02/2026
Bitmovin, a provider of video streaming solutions, announces that 1001, an OTT service in Iraq, has chosen the Bitmovin Player to improve its video streaming pe...
06/02/2026
Combate Global and content creator Shane Fazen announce a licensing agreement to distribute the Hispanic-focused franchise's first three live MMA events in ...
06/02/2026
Cisco is powering the invisible backbone of Super Bowl LX at Levi's Stadium as the technology giant delivers secure, high-capacity connectivity for over 70,...
06/02/2026
Over the past decade, the NFL and Amazon Web Services have changed how football analytics are analyzed and presented through Next Gen Stats. There's real-ti...
06/02/2026
In-venue and creative video staffers at the professional and collegiate level ha...
06/02/2026
In-venue and creative video staffers at the professional and collegiate level ha...
06/02/2026
Ratings Roundup is a rundown of recent rating news and is derived from press rel...
06/02/2026
How the podcast-turned-studio-show Boston Has Entered The Chat became an anchor ...
06/02/2026
ORF, the public service broadcaster for Austria, is in Italy for Milano Cortina 2026, ready to bring the country's most popular winter sports direct to view...
06/02/2026
Milano Cortina 2026 is now underway and Austrian public service broadcaster, ORF...
06/02/2026
Warner Bros. Discovery (WBD) has lifted the curtain on its studios in Italy that...
06/02/2026
Milano Cortina marks the first time since London 2012 that NRK has had the full ...
06/02/2026
Winter sports are wildly popular in Norway, with cross-country skiing and biathl...
06/02/2026
Norwegian broadcaster NRK has the free-to-air rights to the Olympics back for th...
06/02/2026
The production of the mega-esports event also leverages facilities at EA headqua...
06/02/2026
Here's a preview of NBC's massive game and pregame production operation as Super Bowl Sunday approaches....
06/02/2026
Music fans know the feeling: A song stops you in your tracks, and you immediately want to know more. What inspired it, and what's the meaning behind it? We ...
06/02/2026
The National Film and Video Foundation (NFVF), an agency of the Department of Sp...
06/02/2026
Calrec Wins Best of Show at ISE 2026 for Orchestrating Distributed IP Production
Calrec is delighted to announce that its IP Ecosystem Powered by True Control...
06/02/2026
Despite most never having strapped on skis or skates, Aussies are keen for some ...
06/02/2026
MNC Software, a global leader in network management and operational support systems tailored to the broadcast and media industry, today announced the launch of ...
06/02/2026
The annual Junior Eurovision Song Contest arrived at Tbilisi's Gymnastic Hall in Olympic City, presenting an international stage for young talent with rich,...
06/02/2026
NAB Show 2026 | April 19 22 | Booth # N2471
At this year s NAB Show, Sonnet will showcase new Thunderbolt 5 products, including desktop and rackmount PCIe card...
06/02/2026
The Alliance for IP Media Solutions (AIMS), together with the Video Services Forum (VSF), the Advanced Media Workflow Association (AMWA), and the European Broad...