
Idea generation, not hardware or software, needs to be the bottleneck to the advancement of AI, Bryan Catanzaro, vice president of applied deep learning research at NVIDIA, said this week at the AI Hardware Summit.
We want the inventors, the researchers and the engineers that are coming up with future AI to be limited only by their own thoughts, Catanzaro told the audience.
Catanzaro leads a team of researchers working to apply the power of deep learning to everything from video games to chip design. At the annual event held in Silicon Valley, he described the work that NVIDIA is doing to enable advancements in AI, with a focus on large language modeling.
CUDA Is for the Dreamers Training and deploying large neural networks is a tough computational problem, so hardware that's both incredibly fast and highly efficient is a necessity, according to Catanzaro.
But, he explained, the software that accompanies that hardware might be even more important to unlocking further advancements in AI.
The core of the work that we do involves optimizing hardware and software together, all the way from chips, to systems, to software, frameworks, libraries, compilers, algorithms and applications, he said. We optimize all of these things to give transformational capabilities to scientists, researchers and engineers around the world.
This end-to-end approach yields chart-topping performance in industry-standard benchmarks, such as MLPerf. It also ensures that developers aren't constrained by the platform as they aim to advance AI.
CUDA is for the dreamers, CUDA is for the people who are thinking new thoughts, said Catanzaro. How do they think those thoughts and test them efficiently? They need something general and flexible, and that's why we build what we build.
Large Language Models Are Changing the World One of the most exciting areas of AI is language modeling, which is enabling groundbreaking applications in natural language understanding and conversational AI.
The complexity of large language models is growing at an incredible rate, with parameter counts doubling every two months.
A well-known example of a large and powerful language model is GPT-3, developed by OpenAI. Packing 175 billion parameters, it required 314 zettaflops (1021 floating point operations) to train.
It's a staggering amount of compute, Catanzaro said. And that means language modeling is now becoming constrained by economics.
Estimates suggest that GPT-3 would cost about $12 million to train and, Catanzaro observed, the rapid growth in model complexity means that, despite NVIDIA's tireless work to advance the performance and efficiency of its hardware and software, the cost to train these models is set to grow.
And, according to Catanzaro, this trend suggests that it might not be too long before a single model might require more than a billion dollars' worth of computer time to train.
What would it look like to build a model that took a billion dollars to train a single model? Well, it would need to reinvent an entire company, and you'd need to be able to use it in a lot of different contexts, Catanzaro explained.
Catanzaro expects that these models will unlock an incredible amount of value, inspiring continued innovation. During his talk, Catanzaro showed an example of the surprising capabilities of large language models to solve new tasks without being explicitly trained to do so.
After inputting just a few examples into a large language model - four sentences, with two written in English and their corresponding translations into Spanish - he then entered an English sentence, which the model then translated into Spanish properly.
The model was able to do this despite never being trained to do translation. Instead, it was trained - using, as Catanzaro described, an enormous amount of data from the internet - to predict the next word that should follow a given sequence of text.
To perform that very generic task, the model needed to come up with higher-level representations of concepts, such as the existence of languages in general, English and Spanish vocabularies and grammar, and the concept of a translation task, in order to understand the query and properly respond.
These language models are first steps towards generalized artificial intelligence with few shot learning, and that is enormously valuable and very exciting, explained Catanzaro.
A Full-Stack Approach to Language Modeling Catanzaro then went on to describe NVIDIA Megatron, a framework created by NVIDIA using PyTorch for efficiently training the world's largest, transformer-based language models.
A key feature of NVIDIA Megatron, which Catanzaro notes has already been used by various companies and organizations to train large transformer-based models, is model parallelism.
Megatron supports both inter-layer (pipeline) parallelism, which allows different layers of a model to be processed on different devices, as well as intra-layer (tensor) parallelism, which allows a single layer to be processed by multiple different devices.
Catanzaro further described some of the optimizations that NVIDIA applies to maximize the efficiency of pipeline parallelism and minimize so-called pipeline bubbles, during which a GPU is not performing useful work.
A batch is split into microbatches, the execution of which is pipelined. This boosts the utilization of the GPU resources in a system during training. With further optimizations, pipeline bubbles can be reduced even more.
Catanzaro described an optimization, recently published, that entails round-robining each (pipeline) stage among multiple GPUs so that we can further reduce the amount of pipeline bubble overhead in this schedule.
Although this optimization puts additional stress on the communication fabric within the system, Catanzaro showed that, by l
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
04/08/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
04/07/2026
April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...
01/06/2026
January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026
Throughout the week, Dolby brings to life the latest innovatio...
20/05/2026
Cable Captures Only Monthly Increase Among Viewing Categories in March, Earns it...
20/05/2026
April brought a symbolic decrease in the overall time spent in front of television screens. On average, Poles watched video content for 3 hours and 51 minutes a...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
The Royal Television Society Technology Centre today announces the launch of the RTS Technology Awards 2026, celebrating excellence, innovation and achievement ...
20/05/2026
In the heart of London's financial district, the new purpose-built Troubadour Canary Wharf Theatre invites audiences to experience Suzanne Collins' inte...
20/05/2026
LiveU, the leader in live IP-video solutions, today announced that production powerhouse BCC Live successfully deployed the new LU900Q intelligent production un...
20/05/2026
Nella Mente di Narciso Docuseries Uses Blackmagic Design Workflow
Brie Clayton May 19, 2026
0 Comments
PYXIS 6K full frame camera and DaVinci Resolve ...
20/05/2026
Beeble launches Canvas, a node-based AI compositor for VFX and Virtual Productio...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
First Nations Factual Co-production Development Fund launched to elevate Indigen...
20/05/2026
May 20 2026, 06:00 (PDT) Dolby Recognized as 2025 Supplier of the Year and Over...
20/05/2026
Mayo's Dee Freney and Margaret Leahy from Galway have reached the final of RT Today's TV Home Cook competition.
Both contestants will cook again live...
20/05/2026
RT IN FULL BLOOM AT BORD BIA BLOOM 2026 WITH LIVE BROADCASTS, MUSIC, CHAT AND M...
19/05/2026
The winner of Thomson Foundation's Young Journalist of the Year 2025, Tracy Bonareri Onchoke, and runner up Wangu Kanuri enjoyed a three-day trip to London ...
19/05/2026
Cisco and the USGA have announced a multiyear extension of their partnership, which began in 2018. Cisco serves as the Official Technology Partner of the USGA, ...
19/05/2026
Urban Edge Network (UEN), a streaming platform for NAIA sports, has announced a partnership with Spiideo to provide streaming and production tools to UEN's ...
19/05/2026
Warner Bros. Discovery (WBD) will provide live coverage of all 900 Roland-Garros matches across its platforms beginning with qualifiers on May 18. In Europe, 21...
19/05/2026
Tubi, Fox Corporation's free streaming service, has announced the launch of the FIFA World Cup 2026 FOX Hub, a dedicated destination for World Cup programmi...
19/05/2026
Telef nica, in collaboration with Sony, has conducted a 5G connectivity trial at the Movistar Arena in Spain using the 26 GHz millimetre wave (mmWave) band. The...
19/05/2026
Ross Production Services (RPS) has installed a Calrec Argo M console into its new Hypermax-1 remote production truck, replacing one of three Argo S consoles pre...
19/05/2026
Panasonic Projector and Display Corporation has announced the acquisition of 100% of the shares of UK-based media technology company Hive Media Control Ltd. (HI...
19/05/2026
Globecast has announced the completion of a nine-month renovation of its Singapore facility, converting it from a traditional linear broadcast operation into a ...
19/05/2026
Grass Valley has announced a three-year enterprise agreement with Phoenix Broadc...
19/05/2026
Bitmovin has announced that Watch Brasil, a streaming platform operating across Brazil and Europe since 2018, has replaced its legacy systems with Bitmovin'...
19/05/2026
Ateme has announced the migration of Dish Home Nepal's Nepal Premier League (NPL) streaming infrastructure to Ateme's TITAN Live solution deployed on Ak...
19/05/2026
CMSI provided workflow, media management, and HDR support for ESPN during coverage of the NCAA Gymnastics Semifinals and Championships. The company supported fi...
19/05/2026
In advance of this year's Sports Emmy Awards, SVG is taking a deep dive into the six production-technologies nominated for this year's George Wensel Tec...
19/05/2026
In advance of this year's Sports Emmy Awards, SVG is taking a deep dive into the six production-technologies nominated for this year's George Wensel Tec...
19/05/2026
Featuring a fully IP infrastructure, Supershooter 11 is intended for large-scale events. Enabling remote and distributed workflows, Supershooter 65 joins the RE...
19/05/2026
By Jessica Herndon
The line wrapped around the building outside Denver's La...
19/05/2026
Podcasting continues to evolve, and so does Spotify. As we build what comes next, one thing remains constant: This is a medium built on connection. It lives in ...
19/05/2026
Popular design joins Inherit cartridge line-up
When GC Audio introduced their modular Inherit system, it was available with a selection of the company's...
19/05/2026
Resonance-suppression plug-in gets ground-up rebuild
Following on from its 10-year anniversary, oeksound's flagship plug-in has just reached its third m...
19/05/2026
Dedicated FL Studio controller keyboard range refreshed
Novation's dedicated FL Studio controller family has just been upgraded, with four new models ex...
19/05/2026
Rohde & schwarz strengthens its in-vehicle networks test portfolio with the laun...
19/05/2026
Lawful Intelligence: Rohde & Schwarz stellt neues Portfolio f r moderne Polizeia...
19/05/2026
Press Release
18 May 2026, Johannesburg South Africa has been selected as the...
19/05/2026
When all companies in the market are allowed to compete under the same set of ru...
19/05/2026
Battle-proven technology withstands electronic warfare threats across air, maritime surface and subsea domains....
19/05/2026
WESCAM MX-Series EO/IR systems provide high-precision targeting across domains, including counter-UAS applications...
19/05/2026
Champaign, IL - March 16, 2026 Cobalt Digital, the leading designer and manufa...
19/05/2026
CHAMPAIGN, Ill. April 16, 2026 - Cobalt Digital today announced a partnership ...
19/05/2026
Las Vegas - April 18, 2026 Cobalt Digital, the leading designer and manufactur...
19/05/2026
LAS VEGAS April 18, 2026 - Advanced HDR by Technicolor and Cobalt Digital are ...