Sony Pixel Power calrec Sony

Decoding How NVIDIA AI Workbench Powers App Development

19/06/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible and showcases new hardware, software, tools and accelerations for NVIDIA RTX PC and workstation users.

The demand for tools to simplify and optimize generative AI development is skyrocketing. Applications based on retrieval-augmented generation (RAG) - a technique for enhancing the accuracy and reliability of generative AI models with facts fetched from specified external sources - and customized models are enabling developers to tune AI models to their specific needs.

While such work may have required a complex setup in the past, new tools are making it easier than ever.

NVIDIA AI Workbench simplifies AI developer workflows by helping users build their own RAG projects, customize models and more. It's part of the RTX AI Toolkit - a suite of tools and software development kits for customizing, optimizing and deploying AI capabilities - launched at COMPUTEX earlier this month. AI Workbench removes the complexity of technical tasks that can derail experts and halt beginners.

What Is NVIDIA AI Workbench? Available for free, NVIDIA AI Workbench enables users to develop, experiment with, test and prototype AI applications across GPU systems of their choice - from laptops and workstations to data center and cloud. It offers a new approach for creating, using and sharing GPU-enabled development environments across people and systems.

A simple installation gets users up and running with AI Workbench on a local or remote machine in just minutes. Users can then start a new project or replicate one from the examples on GitHub. Everything works through GitHub or GitLab, so users can easily collaborate and distribute work. Learn more about getting started with AI Workbench.

How AI Workbench Helps Address AI Project Challenges Developing AI workloads can require manual, often complex processes, right from the start.

Setting up GPUs, updating drivers and managing versioning incompatibilities can be cumbersome. Reproducing projects across different systems can require replicating manual processes over and over. Inconsistencies when replicating projects, like issues with data fragmentation and version control, can hinder collaboration. Varied setup processes, moving credentials and secrets, and changes in the environment, data, models and file locations can all limit the portability of projects.

AI Workbench makes it easier for data scientists and developers to manage their work and collaborate across heterogeneous platforms. It integrates and automates various aspects of the development process, offering:

Ease of setup: AI Workbench streamlines the process of setting up a developer environment that's GPU-accelerated, even for users with limited technical knowledge.

Seamless collaboration: AI Workbench integrates with version-control and project-management tools like GitHub and GitLab, reducing friction when collaborating.

Consistency when scaling from local to cloud: AI Workbench ensures consistency across multiple environments, supporting scaling up or down from local workstations or PCs to data centers or the cloud.

RAG for Documents, Easier Than Ever NVIDIA offers sample development Workbench Projects to help users get started with AI Workbench. The hybrid RAG Workbench Project is one example: It runs a custom, text-based RAG web application with a user's documents on their local workstation, PC or remote system.

Every Workbench Project runs in a container - software that includes all the necessary components to run the AI application. The hybrid RAG sample pairs a Gradio chat interface frontend on the host machine with a containerized RAG server - the backend that services a user's request and routes queries to and from the vector database and the selected large language model.

This Workbench Project supports a wide variety of LLMs available on NVIDIA's GitHub page. Plus, the hybrid nature of the project lets users select where to run inference.

Workbench Projects let users version the development environment and code. Developers can run the embedding model on the host machine and run inference locally on a Hugging Face Text Generation Inference server, on target cloud resources using NVIDIA inference endpoints like the NVIDIA API catalog, or with self-hosting microservices such as NVIDIA NIM or third-party services.

The hybrid RAG Workbench Project also includes:

Performance metrics: Users can evaluate how RAG- and non-RAG-based user queries perform across each inference mode. Tracked metrics include Retrieval Time, Time to First Token (TTFT) and Token Velocity.

Retrieval transparency: A panel shows the exact snippets of text - retrieved from the most contextually relevant content in the vector database - that are being fed into the LLM and improving the response's relevance to a user's query.

Response customization: Responses can be tweaked with a variety of parameters, such as maximum tokens to generate, temperature and frequency penalty.

To get started with this project, simply install AI Workbench on a local system. The hybrid RAG Workbench Project can be brought from GitHub into the user's account and duplicated to the local system.

More resources are available in the AI Decoded user guide. In addition, community members provide helpful video tutorials, like the one from Joe Freeman below.

Customize, Optimize, Deploy Developers often seek to customize AI models for specific use cases. Fine-tuning, a technique that changes the model by training it with additional data, can be useful for style transfer or changing model behavior. AI Workbench helps with fine-tuning, as well.

The Llama-factory AI Workbench Project enables QLoRa, a fine-tuning method that minimizes memory requirements, for a variety of models, as well as
LINK: https://blogs.nvidia.com/blog/ai-decoded-workbench-hybrid-rag/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

18/02/2026

Milano Cortina 2026: How OBS is Utilizing Audio QC at the First Winter Games to Use Immersive Sound

Audio quality control (QC) is becoming ever more crucial for Olympic Broadcastin...

18/02/2026

Milano Cortina 2026: OBS Demonstrates its Commitment to an Inclusive Sports Media Landscape

The Olympic Games are not only a showcase of athletic excellence, they are also ...

18/02/2026

Ronda Rousey vs. Gina Carano Headlines Netflix's First Live MMA Event

Netflix is entering the MMA game with a matchup between two of the biggest names ever to compete on the women's side. Most Valuable Promotions and Netflix ...

18/02/2026

Network18 Media & Investments Selects Grass Valley's Playout X to Power Unified News Operations

Grass Valley announces that Network18 Media & Investments Ltd., one of India'...

18/02/2026

Cobalt Digital Products Receive IPMX Product Certification Following Inaugural Testing Event

Cobalt Digital Inc., a designer and manufacturer of video and audio conversion, ...

18/02/2026

Channel Nine's Wide World of Sports Hits Livigno; XR From North Sydney Drives Coverage

Production is divided between a studio in the mountains and a brand-new studio i...

18/02/2026

Milano Cortina 2026: Focusing in on the Parabolic' Sound of Skis on Snow at NBC Sports

At the Winter Games IBC in Milan, NBC Sports and Olympics' director of audio...

18/02/2026

2026 Milano Cortina Women's Alpine Photo Gallery

The Women's Alpine events at the Tofane Alpine Skiing Center, about a 10-minute drive from Cortina d'Ampezzo, featured some of the most exciting Olympic...

18/02/2026

Generated in the Italian Alps, NBC's Olympics Audio Is Mixed in Stamford

At the Broadcast Center, 14 audio-control rooms handle the sound in a complex routing and processing regimen We are exactly where we want to be, Karl Malone,...

18/02/2026

Cortina Curling Olympic Stadium Photo Gallery

With 300 hours of curling competition in two weeks, it's a safe bet that even the most curling-hungry fan will be satiated. It also requires a production te...

18/02/2026

Cortina Sliding Center Two-Man Bobsleigh Photo Gallery

Always seen as one of the more crazy Olympic events, Bobsleigh is a sport in which athletes must have nerves of steel and pilots must navigate high-tech sleds...

18/02/2026

Milano Cortina 2026: Inside the Milan-Based Warner Bros. Discovery Casa Italia Studio with Eurosport

A typical Winter Olympics day - or should I say evening - inside the Warner Bros...

18/02/2026

Milano Cortina 2026: How NOS is Delivering a Speed Skating-First Feed for Dutch Audiences

The Netherlands has dominated Milan's ice rinks, scooping six speed skating ...

18/02/2026

NBC Sports Sets Its Own Record With Olympics Comms

At the Broadcast Center in Stamford, three discrete intercom systems combine into a centralized infrastructure Into the second week of Milan Cortina 2026, eigh...

18/02/2026

Deloitte Releases 2026 Global Sports Industry Outlook

AI is reshaping operations, capital is scaling ownership, sports are converging with media and entertainment, and venues are evolving into year-round platforms...

18/02/2026

SVG New Sponsor Spotlight: DNA Inc.'s Thomas Engel on Evolving Digital Experiences and the Power of In-House Collaboration

DNA Inc. specializes in crafting next level digital experiences in the Media, St...

18/02/2026

Copyright Issues Are a Crack in the Ice for Olympic Figure Skating

Conflicts have increased, but so have solutions, driven chiefly by pragmatism and the threat of AI music The only persons on the Figure Skating ice at the 2026...

18/02/2026

Canada's CBC Navigates Multi-Zone Winter Olympics With Bi-Lingual Production Model, Remote Studios, and Custom Content Hubs

Canadian rightsholder deploys its most complex setup at an Olympics ever with ...

18/02/2026

Live From Stamford: NBC Sports Broadcast Center Is Anchor Point of the Entire Olympic Broadcast'

A crew of 1,685 people and 13 control rooms produce nearly every on-air minute f...

18/02/2026

Bad Bunny to Make History in Tokyo With First-Ever Billions Club Live Performance in Asia

Following an extraordinary 2025 in which he was named Spotify's Global Top A...

18/02/2026

Bad Bunny har historia en Tokio con el primer Billions Club Live en Asia

Despu s de un extraordinario 2025 en el que fue nombrado Top Artista Global de Spotify por cuarta vez, algo sin precedentes (y tras su presentaci n en el halfti...

18/02/2026

SBS brings Australians together to mark Ramadan and Eid with communities nationwide

SBS brings Australians together to mark Ramadan and Eid with communities nationw...

18/02/2026

L3Harris Secures Full-Rate Production Contract for US Navy Submarine Communication Systems

The Virginia-class attack submarine USS Texas (SSN 775) underway....

18/02/2026

TV Viewing Hits 12-Month High

Share Copy link Facebook X Linkedin Bluesky Email...

18/02/2026

SinclairLaunches New True Crime Daily Video Podcast

Share Copy link Facebook X Linkedin Bluesky Email...

18/02/2026

E! Co-Founder Alan Mruvka Launches Filmology Labs Studios

Share Copy link Facebook X Linkedin Bluesky Email...

18/02/2026

Stephen Colbert, FCC Commissioner Gomez Blast FCC 'Censorship'

Share Copy link Facebook X Linkedin Bluesky Email...

18/02/2026

NBA All-Star Game Averages 8.8 Million Viewers

Share Copy link Facebook X Linkedin Bluesky Email...

18/02/2026

Speed - Innovation - LEDs - Whitaker ASC and Verbinski Di...

Good Luck, Have Fun, Don't Die in theaters February 2026 Photo courtesy of Graham Bartholomew, SMPSP James Whitaker, ASC reunited with visionary director G...

18/02/2026

TAG Video Systems Expands MCS With New Lens Visualization...

TAG Video Systems has expanded its Media Control System (MCS) version 1.7.0 release with Lens, a visual service health analysis interface that organizes monitor...

18/02/2026

Teradek Introduces RF-X Revolutionizing Mission-Critical...

Teradek, a global leader in advanced video transmission technology, today announced the launch of RF-X Auto Switcher, a revolutionary appliance designed to deli...

18/02/2026

DP Adolpho Veloso Captures Train Dreams with ZEISS Super...

Train Dreams. (L-R) Director of Photography Adolpho Veloso and Joel Edgerton as Robert Grainier on the set of Train Dreams. Cr. Daniel Schaefer/BBP Train Dreams...

18/02/2026

Cerberus Tech Expands Livelinks Multi-Cloud Capabilities...

Cerberus Tech, a leader in cloud-native IP video contribution and distribution, today announced the addition of Akamai Cloud as a supported infrastructure optio...

18/02/2026

ORTC deploys Synamedia Quortex Link for IP channel distri...

Leading video software provider, Synamedia, today announced that Office de Radio et T l vision des Comores (ORTC), the national public broadcaster for the Comor...

18/02/2026

Netflix Opens New Office in Mexico City

Back to All News Netflix Opens New Office in Mexico CityFrom left to right: Francisco Ramos, Vice President of Content for Latin America; Manola Zabalza, Secre...

18/02/2026

February 17, 2026

Calibr-Skaggs and Kainomyx launch collaboration to pioneer novel malaria treatments Partnering to speed up the discovery of new malaria drugs that target parasi...

17/02/2026

Eutelsat, Viewsat Renew Capacity Agreements to Support Development of Broadcast Market in MENA Region

Eutelsat announces the renewal of multiple capacity agreements with Viewsat, rei...

17/02/2026

The Evolution of Communications in Cloud Era of Global Broadcasting

At Milano Cortina 2026, one statistic stood out across industry reporting: 70% of global signal distribution is now happening over public cloud. What began as e...

17/02/2026

CP Communications Supports the Big Game Broadcast for 27 Years

CP Communications provided broadcast support for Super Bowl LX in Santa Clara, Calif. at Levi's Stadium, marking 27 years of working around the Big Game. Th...

17/02/2026

NBC Sports Graphics Team Goes Onsite, Embraces L.A. Motif for NBA All-Star

Street-level images of iconic Los Angeles landmarks are incorporated into the regular-season look With NBC Sports just months into its return to NBA coverage a...

17/02/2026

Jimpa Asks You to Choose Love

Sophie Hyde and John Lithgow backstage during the premiere of Jimpa. (Photo by George Pimentel / Shutterstock for Sundance Film Festival)...

17/02/2026

The 19th Annual South African Film and Television Awards (SAFTAs19) Nominees Announced

Johannesburg, 17 February 2026 - The National Film and Video Foundation (NFVF), ...

17/02/2026

L3Harris Marks Delivery of 3 Million Fuzes to US Army

A Green Beret hangs a 60mm mortar round from a M225A mortar system on a live-fire range. (Photo credit: U.S. Army)...

17/02/2026

The AI Wild West, and Why it Needs a Sheriff

AI Is Scaling Faster Than Governance And That's a Risk AI adoption hasn't rolled out through neat transformation programmes. It has spread organically...

17/02/2026

TV Viewing Hits 12-Month High in Nielsen's January Report of The Gauge

Cable Surges 9% on Strength of College Football Playoffs and News, while NFL Delivers Top 15 Broadcast Telecasts ESPN (+86%) and FOX News Channel (+17%) Combin...