You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper
21/03/2024
Everything that we're enjoying today can be traced back to that moment, Huang said to a packed room with hundreds of attendees, who heard him speak with the authors of Attention Is All You Need.
Sharing the stage for the first time, the research luminaries reflected on the factors that led to their original paper, which has been cited more than 100,000 times since it was first published and presented at the NeurIPS AI conference. They also discussed their latest projects and offered insights into future directions for the field of generative AI.
While they started as Google researchers, the collaborators are now spread across the industry, most as founders of their own AI companies.
We have a whole industry that is grateful for the work that you guys did, Huang said.
From L to R: Lukasz Kaiser, Noam Shazeer, Aidan Gomez, Jensen Huang, Llion Jones, Jakob Uszkoreit, Ashish Vaswani and Illia Polosukhin. Origins of the Transformer Model The research team initially sought to overcome the limitations of recurrent neural networks, or RNNs, which were then the state of the art for processing language data.
Noam Shazeer, cofounder and CEO of Character.AI, compared RNNs to the steam engine and transformers to the improved efficiency of internal combustion.
We could have done the industrial revolution on the steam engine, but it would just have been a pain, he said. Things went way, way better with internal combustion.
Now we're just waiting for the fusion, quipped Illia Polosukhin, cofounder of blockchain company NEAR Protocol.
The paper's title came from a realization that attention mechanisms - an element of neural networks that enable them to determine the relationship between different parts of input data - were the most critical component of their model's performance.
We had very recently started throwing bits of the model away, just to see how much worse it would get. And to our surprise it started getting better, said Llion Jones, cofounder and chief technology officer at Sakana AI.
Having a name as general as transformers spoke to the team's ambitions to build AI models that could process and transform every data type - including text, images, audio, tensors and biological data.
That North Star, it was there on day zero, and so it's been really exciting and gratifying to watch that come to fruition, said Aidan Gomez, cofounder and CEO of Cohere. We're actually seeing it happen now.
Packed house at the San Jose Convention Center. Envisioning the Road Ahead Adaptive computation, where a model adjusts how much computing power is used based on the complexity of a given problem, is a key factor the researchers see improving in future AI models.
It's really about spending the right amount of effort and ultimately energy on a given problem, said Jakob Uszkoreit, cofounder and CEO of biological software company Inceptive. You don't want to spend too much on a problem that's easy or too little on a problem that's hard.
A math problem like two plus two, for example, shouldn't be run through a trillion-parameter transformer model - it should run on a basic calculator, the group agreed.
They're also looking forward to the next generation of AI models.
I think the world needs something better than the transformer, said Gomez. I think all of us here hope it gets succeeded by something that will carry us to a new plateau of performance.
You don't want to miss these next 10 years, Huang said. Unbelievable new capabilities will be invented.
The conversation concluded with Huang presenting each researcher with a framed cover plate of the NVIDIA DGX-1 AI supercomputer, signed with the message, You transformed the world.
Jensen presents lead author Ashish Vaswani with a signed DGX-1 cover. There's still time to catch the session replay by registering for a virtual GTC pass - it's free.
To discover the latest in generative AI, watch Huang's GTC keynote address:
LINK: | https://blogs.nvidia.com/blog/gtc-2024-transformer-ai-research-panel-j... |
See more stories from nvidia |
More from Nvidia
25/04/2024
AI Drives Future of Transportation at Asia's Largest Automotive Show
The latest trends and technologies in the automotive industry are in the spotlight at the Beijing International Automotive Exhibition, aka Auto China, which ope...
25/04/2024
Into the Omniverse: Unlocking the Future of Manufacturing With OpenUSD on Siemens Teamcenter X
Editor's note: This post is part of Into the Omniverse, a series focused on ...
25/04/2024
Blast From the Past: Stream StarCraft' and Diablo' on GeForce NOW
Support for Battle.net on GeForce NOW expands this GFN Thursday, as titles from the iconic StarCraft and Diablo series come to the cloud. StarCraft Remastered,...
24/04/2024
Rays Up: Decoding AI-Powered DLSS 3.5 Ray Reconstruction
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...
24/04/2024
Forecasting the Future: AI2's Christopher Bretherton Discusses Using Machine Learning for Climate Modeling
Can machine learning help predict extreme weather events and climate change? Chr...
24/04/2024
NVIDIA to Acquire GPU Orchestration Software Provider Run:ai
To help customers make more efficient use of their AI computing resources, NVIDIA today announced it has entered into a definitive agreement to acquire Run:ai, ...
24/04/2024
How Virtual Factories Are Making Industrial Digitalization a Reality
To address the shift to electric vehicles, increased semiconductor demand, manufacturing onshoring, and ambitions for greater sustainability, manufacturers are ...
23/04/2024
Small and Mighty: NVIDIA Accelerates Microsoft's Open Phi-3 Mini Language Models
NVIDIA announced today its acceleration of Microsoft's new Phi-3 Mini open l...
22/04/2024
Climate Tech Startups Integrate NVIDIA AI for Sustainability Applications
Whether they're monitoring miniscule insects or delivering insights from satellites in space, NVIDIA-accelerated startups are making every day Earth Day. S...
18/04/2024
Wide Open: NVIDIA Accelerates Inference on Meta Llama 3
NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM). The open mod...
18/04/2024
Up to No Good: No Rest for the Wicked' Early Access Launches on GeForce NOW
It's time to get a little wicked. Members can now stream No Rest for the Wicked from the cloud. It leads six new games joining the GeForce NOW library of m...
18/04/2024
NVIDIA Honors Partners of the Year in Europe, Middle East, Africa
NVIDIA today recognized 18 partners in Europe, the Middle East and Africa for their achievements and commitment to driving AI adoption. The recipients were hon...
17/04/2024
Seeing Beyond: Living Optics CEO Robin Wang on Democratizing Hyperspectral Imaging
Step into the realm of the unseen with Robin Wang, CEO of Living Optics. The sta...
17/04/2024
Moving Pictures: Transform Images Into 3D Scenes With NVIDIA Instant NeRF
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...
16/04/2024
New NVIDIA RTX A400 and A1000 GPUs Enhance AI-Powered Design and Productivity Workflows
AI integration across design and productivity applications is becoming the new s...
16/04/2024
To Cut a Long Story Short: Video Editors Benefit From DaVinci Resolve's New AI Features Powered by RTX
Editor's note: This post is part of our In the NVIDIA Studio series, which c...
15/04/2024
AI Is Tech's Greatest Contribution to Social Elevation,' NVIDIA CEO Tells Oregon State Students
AI promises to bring the full benefits of the digital revolution to billions acr...
10/04/2024
The Building Blocks of AI: Decoding the Role and Significance of Foundation Models
Editor's note: This post is part of the AI Decoded series, which demystifies...
10/04/2024
Combating Corruption With Data: Cleanlab and Berkeley Research Group on Using AI-Powered Investigative Analytics
Talk about scrubbing data. Curtis Northcutt, cofounder and CEO of Cleanlab, and ...
09/04/2024
NVIDIA Joins $110 Million Partnership to Help Universities Teach AI Skills
The Biden Administration has announced a new $110 million AI partnership between Japan and the United States that includes an initiative to fund research throug...
09/04/2024
Broadcasting Breakthroughs: NVIDIA Holoscan for Media, Available Now, Transforms Live Media With Easy AI Integration
Whether delivering live sports programming, streaming services, network broadcas...
09/04/2024
Start Up Your Engines: NVIDIA and Google Cloud Collaborate to Accelerate AI Development
NVIDIA and Google Cloud have announced a new collaboration to help startups arou...
04/04/2024
NVIDIA Ranked by Fortune at No. 3 on 100 Best Companies to Work For' List
NVIDIA jumped to No. 3 on the latest list of America's 100 Best Companies to Work For by Fortune magazine and Great Place to Work. It's the company'...
04/04/2024
The Elder Scrolls Online' Joins GeForce NOW for Game's 10th Anniversary
Rain or shine, a new month means new games. GeForce NOW kicks off April with nearly 20 new games, seven of which are available to play this week. GFN Thursday ...
03/04/2024
A New Lens: Dotlumen CEO Cornel Amariei on Assistive Technology for the Visually Impaired
Dotlumen is illuminating a new technology to help people with visual impairments...
03/04/2024
Coming Up ACEs: Decoding the AI Technology That's Enhancing Games With Realistic Digital Humans
Editor's note: This post is part of the AI Decoded series, which demystifies...
28/03/2024
Greater Scope: Doctors Get Inside Look at Gut Health With AI-Powered Endoscopy
From humble beginnings as a university spinoff to an acquisition by the leading global medtech company in its field, Odin Vision has been on an accelerated jour...
28/03/2024
Get Cozy With Palia' on GeForce NOW
Ease into spring with the warm, cozy vibes of Palia, coming to the cloud this GFN Thursday. It's part of six new titles joining the GeForce NOW library of ...
27/03/2024
Software Developers Launch OpenUSD and Generative AI-Powered Product Configurators Built on NVIDIA Omniverse
From designing dream cars to customizing clothing, 3D product configurators are ...
27/03/2024
NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf
It's official: NVIDIA delivered the world's fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks...
27/03/2024
Viome's Guru Banavar Discusses AI for Personalized Health
In the latest episode of NVIDIA's AI Podcast, Viome Chief Technology Officer Guru Banavar spoke with host Noah Kravitz about how AI and RNA sequencing are r...
27/03/2024
Unlocking Peak Generations: TensorRT Accelerates AI on RTX PCs and Workstations
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...
26/03/2024
Boom in AI-Enabled Medical Devices Transforms Healthcare
The future of healthcare is software-defined and AI-enabled. Around 700 FDA-cleared, AI-enabled medical devices are now on the market - more than 10x the number...
26/03/2024
Model Innovators: How Digital Twins Are Making Industries More Efficient
A manufacturing plant near Hsinchu, Taiwan's Silicon Valley, is among facilities worldwide boosting energy efficiency with AI-enabled digital twins. A virt...
26/03/2024
Into the Omniverse: Groundbreaking OpenUSD Advancements Put NVIDIA GTC Spotlight on Developers
Editor's note: This post is part of Into the Omniverse, a series focused on ...
25/03/2024
NVIDIA Blackwell and Automotive Industry Innovators Dazzle at NVIDIA GTC
Generative AI, in the data center and in the car, is making vehicle experiences safer and more enjoyable. The latest advancements in automotive technology were...
21/03/2024
AI's New Frontier: From Daydreams to Digital Deeds
Imagine a world where you can whisper your digital wishes into your device, and poof, it happens. That world may be coming sooner than you think. But if you...
21/03/2024
You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper
Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by...
21/03/2024
Instant Latte: NVIDIA Gen AI Research Brews 3D Shapes in Under a Second
NVIDIA researchers have pumped a double shot of acceleration into their latest text-to-3D generative AI model, dubbed LATTE3D. Like a virtual 3D printer, LATTE...
21/03/2024
Here Be Dragons: Dragon's Dogma 2' Comes to GeForce NOW
Arise for a new adventure with Dragon's Dogma 2, leading two new titles joining the GeForce NOW library this week. Set Forth, Arisen Fulfill a forgotten de...
20/03/2024
AI Decoded From GTC: The Latest Developer Tools and Apps Accelerating AI on PC and Workstation
Editor's note: This post is part of the AI Decoded series, which demystifies...
19/03/2024
NVIDIA Celebrates Americas Partners Driving AI-Powered Transformation
NVIDIA recognized 14 partners in the Americas for their achievements in transforming businesses with AI, this week at GTC. The winners of the NVIDIA Partner Ne...
19/03/2024
Climate Pioneers: 3 Startups Harnessing NVIDIA's AI and Earth-2 Platforms
To help mitigate climate change - one of humanity's greatest challenges - researchers are turning to AI and sustainable computing to accelerate and operatio...
19/03/2024
Secure by Design: NVIDIA AIOps Partner Ecosystem Blends AI for Businesses
In today's complex business environments, IT teams face a constant flow of challenges, from simple issues like employee account lockouts to critical securit...
19/03/2024
Generation Sensation: New Generative AI and RTX Tools Boost Content Creation
Editor's note: This post is part of our In the NVIDIA Studio series, which celebrates featured artists, offers creative tips and tricks, and demonstrates ho...
19/03/2024
NVIDIA, Huang Win Top Honors in Innovation, Engineering
NVIDIA today was named the world's most innovative company by Fast Company magazine. The accolade comes on the heels of company founder and CEO Jensen Huan...
18/03/2024
NVIDIA Edify Unlocks 3D Generative AI, New Image Controls for Visual Content Providers
NVIDIA Edify, a multimodal architecture for visual generative AI, is entering a ...
18/03/2024
From Atoms to Supercomputers: NVIDIA, Partners Scale Quantum Computing
The latest advances in quantum computing include investigating molecules, deploying giant supercomputers and building the quantum workforce with a new academic ...
18/03/2024
New NVIDIA Storage Partner Validation Program Streamlines Enterprise AI Deployments
A sharp increase in generative AI deployments is driving business innovation for...
18/03/2024
NVIDIA Unveils Digital Blueprint for Building Next-Gen Data Centers
Designing, simulating and bringing up modern data centers is incredibly complex, involving multiple considerations like performance, energy efficiency and scala...