The Building Blocks of AI: Decoding the Role and Significance of Foundation Models
10/04/2024
Skyscrapers start with strong foundations. The same goes for apps powered by AI.
A foundation model is an AI neural network trained on immense amounts of raw data, generally with unsupervised learning.
It's a type of artificial intelligence model trained to understand and generate human-like language. Imagine giving a computer a huge library of books to read and learn from, so it can understand the context and meaning behind words and sentences, just like a human does.
Foundation models. A foundation model's deep knowledge base and ability to communicate in natural language make it useful for a broad range of applications, including text generation and summarization, copilot production and computer code analysis, image and video creation, and audio transcription and speech synthesis.
ChatGPT, one of the most notable generative AI applications, is a chatbot built with OpenAI's GPT foundation model. Now in its fourth version, GPT-4 is a large multimodal model that can ingest text or images and generate text or image responses.
Online apps built on foundation models typically access the models from a data center. But many of these models, and the applications they power, can now run locally on PCs and workstations with NVIDIA GeForce and NVIDIA RTX GPUs.
Foundation Model Uses Foundation models can perform a variety of functions, including:
Language processing: understanding and generating text
Code generation: analyzing and debugging computer code in many programming languages
Visual processing: analyzing and generating images
Speech: generating text to speech and transcribing speech to text
They can be used as is or with further refinement. Rather than training an entirely new AI model for each generative AI application - a costly and time-consuming endeavor - users commonly fine-tune foundation models for specialized use cases.
Pretrained foundation models are remarkably capable, thanks to prompts and data-retrieval techniques like retrieval-augmented generation, or RAG. Foundation models also excel at transfer learning, which means they can be trained to perform a second task related to their original purpose.
For example, a general-purpose large language model (LLM) designed to converse with humans can be further trained to act as a customer service chatbot capable of answering inquiries using a corporate knowledge base.
Enterprises across industries are fine-tuning foundation models to get the best performance from their AI applications.
Types of Foundation Models More than 100 foundation models are in use - a number that continues to grow. LLMs and image generators are the two most popular types of foundation models. And many of them are free for anyone to try - on any hardware - in the NVIDIA API Catalog.
LLMs are models that understand natural language and can respond to queries. Google's Gemma is one example; it excels at text comprehension, transformation and code generation. When asked about the astronomer Cornelius Gemma, it shared that his contributions to celestial navigation and astronomy significantly impacted scientific progress. It also provided information on his key achievements, legacy and other facts.
Extending the collaboration of the Gemma models, accelerated with the NVIDIA TensorRT-LLM on RTX GPUs, Google's CodeGemma brings powerful yet lightweight coding capabilities to the community. CodeGemma models are available as 7B and 2B pretrained variants that specialize in code completion and code generation tasks.
MistralAI's Mistral LLM can follow instructions, complete requests and generate creative text. In fact, it helped brainstorm the headline for this blog, including the requirement that it use a variation of the series' name AI Decoded, and it assisted in writing the definition of a foundation model.
Hello, world, indeed. Meta's Llama 2 is a cutting-edge LLM that generates text and code in response to prompts.
Mistral and Llama 2 are available in the NVIDIA ChatRTX tech demo, running on RTX PCs and workstations. ChatRTX lets users personalize these foundation models by connecting them to personal content - such as documents, doctors' notes and other data - through RAG. It's accelerated by TensorRT-LLM for quick, contextually relevant answers. And because it runs locally, results are fast and secure.
Image generators like StabilityAI's Stable Diffusion XL and SDXL Turbo let users generate images and stunning, realistic visuals. StabilityAI's video generator, Stable Video Diffusion, uses a generative diffusion model to synthesize video sequences with a single image as a conditioning frame.
Multimodal foundation models can simultaneously process more than one type of data - such as text and images - to generate more sophisticated outputs.
A multimodal model that works with both text and images could let users upload an image and ask questions about it. These types of models are quickly working their way into real-world applications like customer service, where they can serve as faster, more user-friendly versions of traditional manuals.
Many foundation models are free to try - on any hardware - in the NVIDIA API Catalog. Kosmos 2 is Microsoft's groundbreaking multimodal model designed to understand and reason about visual elements in images.
Think Globally, Run AI Models Locally GeForce RTX and NVIDIA RTX GPUs can run foundation models locally.
The results are fast and secure. Rather than relying on cloud-based services, users can harness apps like ChatRTX to process sensitive data on their local PC without sharing the data with a third party or needing an internet connection.
More from Nvidia
29/04/2024
SEA.AI Navigates the Future With AI at the Helm
Talk about commitment. When startup SEA.AI, an NVIDIA Metropolis partner, set out to create a system that would use AI to scan the seas to enhance maritime safe...
25/04/2024
AI Drives Future of Transportation at Asia's Largest Automotive Show
The latest trends and technologies in the automotive industry are in the spotlight at the Beijing International Automotive Exhibition, aka Auto China, which ope...
25/04/2024
Into the Omniverse: Unlocking the Future of Manufacturing With OpenUSD on Siemens Teamcenter X
Editor's note: This post is part of Into the Omniverse, a series focused on ...
25/04/2024
Blast From the Past: Stream StarCraft' and Diablo' on GeForce NOW
Support for Battle.net on GeForce NOW expands this GFN Thursday, as titles from the iconic StarCraft and Diablo series come to the cloud. StarCraft Remastered,...
24/04/2024
Rays Up: Decoding AI-Powered DLSS 3.5 Ray Reconstruction
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...
24/04/2024
Forecasting the Future: AI2's Christopher Bretherton Discusses Using Machine Learning for Climate Modeling
Can machine learning help predict extreme weather events and climate change? Chr...
24/04/2024
NVIDIA to Acquire GPU Orchestration Software Provider Run:ai
To help customers make more efficient use of their AI computing resources, NVIDIA today announced it has entered into a definitive agreement to acquire Run:ai, ...
24/04/2024
How Virtual Factories Are Making Industrial Digitalization a Reality
To address the shift to electric vehicles, increased semiconductor demand, manufacturing onshoring, and ambitions for greater sustainability, manufacturers are ...
23/04/2024
Small and Mighty: NVIDIA Accelerates Microsoft's Open Phi-3 Mini Language Models
NVIDIA announced today its acceleration of Microsoft's new Phi-3 Mini open l...
22/04/2024
Climate Tech Startups Integrate NVIDIA AI for Sustainability Applications
Whether they're monitoring miniscule insects or delivering insights from satellites in space, NVIDIA-accelerated startups are making every day Earth Day. S...
18/04/2024
Wide Open: NVIDIA Accelerates Inference on Meta Llama 3
NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM). The open mod...
18/04/2024
Up to No Good: No Rest for the Wicked' Early Access Launches on GeForce NOW
It's time to get a little wicked. Members can now stream No Rest for the Wicked from the cloud. It leads six new games joining the GeForce NOW library of m...
18/04/2024
NVIDIA Honors Partners of the Year in Europe, Middle East, Africa
NVIDIA today recognized 18 partners in Europe, the Middle East and Africa for their achievements and commitment to driving AI adoption. The recipients were hon...
17/04/2024
Seeing Beyond: Living Optics CEO Robin Wang on Democratizing Hyperspectral Imaging
Step into the realm of the unseen with Robin Wang, CEO of Living Optics. The sta...
17/04/2024
Moving Pictures: Transform Images Into 3D Scenes With NVIDIA Instant NeRF
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...
16/04/2024
New NVIDIA RTX A400 and A1000 GPUs Enhance AI-Powered Design and Productivity Workflows
AI integration across design and productivity applications is becoming the new s...
16/04/2024
To Cut a Long Story Short: Video Editors Benefit From DaVinci Resolve's New AI Features Powered by RTX
Editor's note: This post is part of our In the NVIDIA Studio series, which c...
15/04/2024
AI Is Tech's Greatest Contribution to Social Elevation,' NVIDIA CEO Tells Oregon State Students
AI promises to bring the full benefits of the digital revolution to billions acr...
10/04/2024
The Building Blocks of AI: Decoding the Role and Significance of Foundation Models
Editor's note: This post is part of the AI Decoded series, which demystifies...
10/04/2024
Combating Corruption With Data: Cleanlab and Berkeley Research Group on Using AI-Powered Investigative Analytics
Talk about scrubbing data. Curtis Northcutt, cofounder and CEO of Cleanlab, and ...
09/04/2024
NVIDIA Joins $110 Million Partnership to Help Universities Teach AI Skills
The Biden Administration has announced a new $110 million AI partnership between Japan and the United States that includes an initiative to fund research throug...
09/04/2024
Broadcasting Breakthroughs: NVIDIA Holoscan for Media, Available Now, Transforms Live Media With Easy AI Integration
Whether delivering live sports programming, streaming services, network broadcas...
09/04/2024
Start Up Your Engines: NVIDIA and Google Cloud Collaborate to Accelerate AI Development
NVIDIA and Google Cloud have announced a new collaboration to help startups arou...
04/04/2024
NVIDIA Ranked by Fortune at No. 3 on 100 Best Companies to Work For' List
NVIDIA jumped to No. 3 on the latest list of America's 100 Best Companies to Work For by Fortune magazine and Great Place to Work. It's the company'...
04/04/2024
The Elder Scrolls Online' Joins GeForce NOW for Game's 10th Anniversary
Rain or shine, a new month means new games. GeForce NOW kicks off April with nearly 20 new games, seven of which are available to play this week. GFN Thursday ...
03/04/2024
A New Lens: Dotlumen CEO Cornel Amariei on Assistive Technology for the Visually Impaired
Dotlumen is illuminating a new technology to help people with visual impairments...
03/04/2024
Coming Up ACEs: Decoding the AI Technology That's Enhancing Games With Realistic Digital Humans
Editor's note: This post is part of the AI Decoded series, which demystifies...
28/03/2024
Greater Scope: Doctors Get Inside Look at Gut Health With AI-Powered Endoscopy
From humble beginnings as a university spinoff to an acquisition by the leading global medtech company in its field, Odin Vision has been on an accelerated jour...
28/03/2024
Get Cozy With Palia' on GeForce NOW
Ease into spring with the warm, cozy vibes of Palia, coming to the cloud this GFN Thursday. It's part of six new titles joining the GeForce NOW library of ...
27/03/2024
Software Developers Launch OpenUSD and Generative AI-Powered Product Configurators Built on NVIDIA Omniverse
From designing dream cars to customizing clothing, 3D product configurators are ...
27/03/2024
NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf
It's official: NVIDIA delivered the world's fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks...
27/03/2024
Viome's Guru Banavar Discusses AI for Personalized Health
In the latest episode of NVIDIA's AI Podcast, Viome Chief Technology Officer Guru Banavar spoke with host Noah Kravitz about how AI and RNA sequencing are r...
27/03/2024
Unlocking Peak Generations: TensorRT Accelerates AI on RTX PCs and Workstations
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...
26/03/2024
Boom in AI-Enabled Medical Devices Transforms Healthcare
The future of healthcare is software-defined and AI-enabled. Around 700 FDA-cleared, AI-enabled medical devices are now on the market - more than 10x the number...
26/03/2024
Model Innovators: How Digital Twins Are Making Industries More Efficient
A manufacturing plant near Hsinchu, Taiwan's Silicon Valley, is among facilities worldwide boosting energy efficiency with AI-enabled digital twins. A virt...
26/03/2024
Into the Omniverse: Groundbreaking OpenUSD Advancements Put NVIDIA GTC Spotlight on Developers
Editor's note: This post is part of Into the Omniverse, a series focused on ...
25/03/2024
NVIDIA Blackwell and Automotive Industry Innovators Dazzle at NVIDIA GTC
Generative AI, in the data center and in the car, is making vehicle experiences safer and more enjoyable. The latest advancements in automotive technology were...
21/03/2024
AI's New Frontier: From Daydreams to Digital Deeds
Imagine a world where you can whisper your digital wishes into your device, and poof, it happens. That world may be coming sooner than you think. But if you...
21/03/2024
You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper
Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by...
21/03/2024
Instant Latte: NVIDIA Gen AI Research Brews 3D Shapes in Under a Second
NVIDIA researchers have pumped a double shot of acceleration into their latest text-to-3D generative AI model, dubbed LATTE3D. Like a virtual 3D printer, LATTE...
21/03/2024
Here Be Dragons: Dragon's Dogma 2' Comes to GeForce NOW
Arise for a new adventure with Dragon's Dogma 2, leading two new titles joining the GeForce NOW library this week. Set Forth, Arisen Fulfill a forgotten de...
20/03/2024
AI Decoded From GTC: The Latest Developer Tools and Apps Accelerating AI on PC and Workstation
Editor's note: This post is part of the AI Decoded series, which demystifies...
19/03/2024
NVIDIA Celebrates Americas Partners Driving AI-Powered Transformation
NVIDIA recognized 14 partners in the Americas for their achievements in transforming businesses with AI, this week at GTC. The winners of the NVIDIA Partner Ne...
19/03/2024
Climate Pioneers: 3 Startups Harnessing NVIDIA's AI and Earth-2 Platforms
To help mitigate climate change - one of humanity's greatest challenges - researchers are turning to AI and sustainable computing to accelerate and operatio...
19/03/2024
Secure by Design: NVIDIA AIOps Partner Ecosystem Blends AI for Businesses
In today's complex business environments, IT teams face a constant flow of challenges, from simple issues like employee account lockouts to critical securit...
19/03/2024
Generation Sensation: New Generative AI and RTX Tools Boost Content Creation
Editor's note: This post is part of our In the NVIDIA Studio series, which celebrates featured artists, offers creative tips and tricks, and demonstrates ho...
19/03/2024
NVIDIA, Huang Win Top Honors in Innovation, Engineering
NVIDIA today was named the world's most innovative company by Fast Company magazine. The accolade comes on the heels of company founder and CEO Jensen Huan...
18/03/2024
NVIDIA Edify Unlocks 3D Generative AI, New Image Controls for Visual Content Providers
NVIDIA Edify, a multimodal architecture for visual generative AI, is entering a ...
18/03/2024
From Atoms to Supercomputers: NVIDIA, Partners Scale Quantum Computing
The latest advances in quantum computing include investigating molecules, deploying giant supercomputers and building the quantum workforce with a new academic ...
18/03/2024
New NVIDIA Storage Partner Validation Program Streamlines Enterprise AI Deployments
A sharp increase in generative AI deployments is driving business innovation for...