
Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible and showcases new hardware, software, tools and accelerations for NVIDIA RTX PC and workstation users.
The demand for tools to simplify and optimize generative AI development is skyrocketing. Applications based on retrieval-augmented generation (RAG) - a technique for enhancing the accuracy and reliability of generative AI models with facts fetched from specified external sources - and customized models are enabling developers to tune AI models to their specific needs.
While such work may have required a complex setup in the past, new tools are making it easier than ever.
NVIDIA AI Workbench simplifies AI developer workflows by helping users build their own RAG projects, customize models and more. It's part of the RTX AI Toolkit - a suite of tools and software development kits for customizing, optimizing and deploying AI capabilities - launched at COMPUTEX earlier this month. AI Workbench removes the complexity of technical tasks that can derail experts and halt beginners.
What Is NVIDIA AI Workbench? Available for free, NVIDIA AI Workbench enables users to develop, experiment with, test and prototype AI applications across GPU systems of their choice - from laptops and workstations to data center and cloud. It offers a new approach for creating, using and sharing GPU-enabled development environments across people and systems.
A simple installation gets users up and running with AI Workbench on a local or remote machine in just minutes. Users can then start a new project or replicate one from the examples on GitHub. Everything works through GitHub or GitLab, so users can easily collaborate and distribute work. Learn more about getting started with AI Workbench.
How AI Workbench Helps Address AI Project Challenges Developing AI workloads can require manual, often complex processes, right from the start.
Setting up GPUs, updating drivers and managing versioning incompatibilities can be cumbersome. Reproducing projects across different systems can require replicating manual processes over and over. Inconsistencies when replicating projects, like issues with data fragmentation and version control, can hinder collaboration. Varied setup processes, moving credentials and secrets, and changes in the environment, data, models and file locations can all limit the portability of projects.
AI Workbench makes it easier for data scientists and developers to manage their work and collaborate across heterogeneous platforms. It integrates and automates various aspects of the development process, offering:
Ease of setup: AI Workbench streamlines the process of setting up a developer environment that's GPU-accelerated, even for users with limited technical knowledge.
Seamless collaboration: AI Workbench integrates with version-control and project-management tools like GitHub and GitLab, reducing friction when collaborating.
Consistency when scaling from local to cloud: AI Workbench ensures consistency across multiple environments, supporting scaling up or down from local workstations or PCs to data centers or the cloud.
RAG for Documents, Easier Than Ever NVIDIA offers sample development Workbench Projects to help users get started with AI Workbench. The hybrid RAG Workbench Project is one example: It runs a custom, text-based RAG web application with a user's documents on their local workstation, PC or remote system.
Every Workbench Project runs in a container - software that includes all the necessary components to run the AI application. The hybrid RAG sample pairs a Gradio chat interface frontend on the host machine with a containerized RAG server - the backend that services a user's request and routes queries to and from the vector database and the selected large language model.
This Workbench Project supports a wide variety of LLMs available on NVIDIA's GitHub page. Plus, the hybrid nature of the project lets users select where to run inference.
Workbench Projects let users version the development environment and code. Developers can run the embedding model on the host machine and run inference locally on a Hugging Face Text Generation Inference server, on target cloud resources using NVIDIA inference endpoints like the NVIDIA API catalog, or with self-hosting microservices such as NVIDIA NIM or third-party services.
The hybrid RAG Workbench Project also includes:
Performance metrics: Users can evaluate how RAG- and non-RAG-based user queries perform across each inference mode. Tracked metrics include Retrieval Time, Time to First Token (TTFT) and Token Velocity.
Retrieval transparency: A panel shows the exact snippets of text - retrieved from the most contextually relevant content in the vector database - that are being fed into the LLM and improving the response's relevance to a user's query.
Response customization: Responses can be tweaked with a variety of parameters, such as maximum tokens to generate, temperature and frequency penalty.
To get started with this project, simply install AI Workbench on a local system. The hybrid RAG Workbench Project can be brought from GitHub into the user's account and duplicated to the local system.
More resources are available in the AI Decoded user guide. In addition, community members provide helpful video tutorials, like the one from Joe Freeman below.
Customize, Optimize, Deploy Developers often seek to customize AI models for specific use cases. Fine-tuning, a technique that changes the model by training it with additional data, can be useful for style transfer or changing model behavior. AI Workbench helps with fine-tuning, as well.
The Llama-factory AI Workbench Project enables QLoRa, a fine-tuning method that minimizes memory requirements, for a variety of models, as well as
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
01/06/2026
January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026
Throughout the week, Dolby brings to life the latest innovatio...
02/05/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
01/05/2026
January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...
01/04/2026
January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION
Douyin Users Can Now Create And Share Videos With Stun...
24/03/2026
Utah Scientific to Showcase a Variety of Hybrid SDI/IP Innovations at the 2026 N...
24/03/2026
Mediaproxy will showcase an updated version of its LogPlayer interface at NAB Sh...
24/03/2026
C gep de Jonqui re, a public college in Qu bec, has upgraded its TV studio complex with five Calrec Brio consoles and a Dante IP network. The installation was c...
24/03/2026
The American Association of Professional Baseball (AAPB) has announced new and r...
24/03/2026
National Telecom Public Company Limited (NT) used Ateme's KYRION platform fo...
24/03/2026
Grass Valley has expanded its relationship with ImSoPROD, the TV production entity of FDJ United, with the deployment of LDX C98 compact cameras across its nati...
24/03/2026
The E.W. Scripps Company is launching Scripps Sports Network (SSN), a free, ad-s...
24/03/2026
NEP Group has launched NEP Platform, a new software orchestration system that un...
24/03/2026
Daktronics has partnered with the Arizona Diamondbacks to design, manufacture and install a new LED video display and three ribbon boards at Chase Field in Phoe...
24/03/2026
Grass Valley has completed an upgrade of University of Pittsburgh Athletics'...
24/03/2026
Daktronics is upgrading eight LED displays at Wrigley Field in Chicago, totaling more than 8,300 square feet, using the Daktronics Renew product line. The upgra...
24/03/2026
Advanced Systems Group has appointed Kevin Poole as Senior Project and Support Manager within the ASG Workflow and Tools practice. The position is newly created...
24/03/2026
In an increasingly streaming-centric world, protecting high-value live sports co...
24/03/2026
Mediaproxy will showcase an updated version of its LogPlayer interface at NAB Show, April 19-22 in Las Vegas, booth W1423. The update includes AI tools, native ...
24/03/2026
The production studio has worked with content juggernauts such as MrBeast, The R...
24/03/2026
The event will also feature a State of the Industry address from sports-media veteran and strategic advisor Patrick Crakes....
24/03/2026
Victory , a free sports streaming service owned by A Parent Media Co. (APMC), has launched its NWSL Sunday Night Soccer franchise. The platform says the launch ...
24/03/2026
Sports content in free streaming is still relatively scant, but it's growing, right along with its monetization capabilities
Streaming-tech company Wurl re...
24/03/2026
Directed by Greg DeHart, the doc will feature Sports Broadcasting Hall of Famers...
24/03/2026
For the first time ever, Tubi - Fox's free streaming service - is working w...
24/03/2026
As the music industry gathers for Canada's Juno Awards, the local edition of...
24/03/2026
Behind every great song is a complex web of people, stories, and inspirations th...
24/03/2026
Few filmmakers have shaped popular cinema as profoundly as Steven Spielberg. So ...
24/03/2026
Latest Eurorack modules revealed
ALM/Busy Circuits have just introduced another two Eurorack modules, delivering a powerful new modulator that builds on the...
24/03/2026
Smallest, most affordable MPC to date
Akai Pro have just introduced their most accessible standalone sampler to date, making the iconic MPC experience avail...
24/03/2026
Create & manage presets on Windows & macOS
Electro-Harmonix have announced that their free Mac/Windows application now allows users of their Oceans Abyss re...
24/03/2026
Rohde & Schwarz amplifiers enable high-field immunity testing expansion at IB Le...
24/03/2026
The National Film and Video Foundation (NFVF) invites eligible Tier 1 and 2 South African production companies to submit proposals to serve as the Facilitating ...
24/03/2026
L3Harris ramps up production of its VAMPIRE C-UxS system....
24/03/2026
For more than 70 years, L3Harris has delivered essential fuzes and safe-and-arm solutions that bring the might to multi-domain battlefields across the globe, ...
24/03/2026
eds3_5_jq(document).ready(function($) { $(#eds_sliderM519).chameleonSlider_2_1({...
24/03/2026
Rick Bernier reflects on a career that has taken him to the very top of live broadcast audio. As a Senior Broadcast Audio A1 Engineer and Music Director, he has...
24/03/2026
Grass Valley has strengthened its long-standing relationship with ImSoPROD, the TV production entity of the FDJ United, with the deployment of Grass Valley LDX ...
24/03/2026
At the 2026 NAB Show in Las Vegas, Utah Scientific will spotlight its expanding portfolio of hybrid SDI/IP routing, conversion, control, and signal management s...
24/03/2026
Fresh off its global reveal at ISE 2026 in Barcelona, the new V410 live 4K encoder/decoder from Miri Technologies Inc. will make its North American debut at the...
24/03/2026
Cinegy GmbH, premier provider of software-defined television technology, will show how its solutions answer the real challenges facing content creators and broa...
24/03/2026
Canberra-based radio and online broadcast network ArtSound FM (https://artsound.fm/) has invested in DHD audio mixing consoles for two new studios at its headqu...
24/03/2026
IBC today announced the nine projects selected for its 2026 Accelerator Media Innovation Programme, bringing together organisations from across broadcast, strea...
24/03/2026
Mediaproxy, the global standard for software-based IP compliance monitoring and multiviewing solutions, will showcase the next generation of its industry-leadin...
24/03/2026
Underscoring its global reputation as one of Canada's most renowned colleges for television production, the C gep de Jonqui re public college in Qu bec has ...
24/03/2026
Leading video software provider Synamedia today announced the launch of the industry's first edge watermarking solution that enables faster disruption of pi...
24/03/2026
Experience Commerce, a leading full-service digital marketing agency and part of the Cheil SWA Group, has secured the annual social media mandate for Fiat Payme...
24/03/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
24/03/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
24/03/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...