Sony Pixel Power calrec Sony

How to Fine-Tune an LLM on NVIDIA GPUs With Unsloth

15/12/2025

Modern workflows showcase the endless possibilities of generative and agentic AI on PCs.

Of many, some examples include tuning a chatbot to handle product-support questions or building a personal assistant for managing one's schedule. A challenge remains, however, in getting a small language model to respond consistently with high accuracy for specialized agentic tasks.

That's where fine-tuning comes in.

Unsloth, one of the world's most widely used open-source frameworks for fine-tuning LLMs, provides an approachable way to customize models. It's optimized for efficient, low-memory training on NVIDIA GPUs - from GeForce RTX desktops and laptops to RTX PRO workstations and DGX Spark, the world's smallest AI supercomputer.

Another powerful starting point for fine-tuning is the just-announced NVIDIA Nemotron 3 family of open models, data and libraries. Nemotron 3 introduces the most efficient family of open models, ideal for agentic AI fine-tuning.

Teaching AI New Tricks Fine-tuning is like giving an AI model a focused training session. With examples tied to a specific topic or workflow, the model improves its accuracy by learning new patterns and adapting to the task at hand.

Choosing a fine-tuning method for a model depends on how much of the original model the developer wants to adjust. Based on their goals, developers can use one of three main fine-tuning methods:

Parameter-efficient fine-tuning (such as LoRA or QLoRA):

How it works: Updates only a small portion of the model for faster, lower-cost training. It's a smarter and efficient way to enhance a model without altering it drastically.

Target use case: Useful across nearly all scenarios where full fine-tuning would traditionally be applied - including adding domain knowledge, improving coding accuracy, adapting the model for legal or scientific tasks, refining reasoning, or aligning tone and behavior.

Requirements: Small- to medium-sized dataset (100-1,000 prompt-sample pairs).

Full fine-tuning:

How it works: Updates all of the model's parameters - useful for teaching the model to follow specific formats or styles.

Target use case: Advanced use cases, such as building AI agents and chatbots that must provide assistance about a specific topic, stay within a certain set of guardrails and respond in a particular manner.

Requirements: Large dataset (1,000+ prompt-sample pairs).

Reinforcement learning:

How it works: Adjusts the behavior of the model using feedback or preference signals. The model learns by interacting with its environment and uses the feedback to improve itself over time. This is a complex, advanced technique that interweaves training and inference - and can be used in tandem with parameter-efficient fine-tuning and full fine-tuning techniques. See Unsloth's Reinforcement Learning Guide for details.

Target use case: Improving the accuracy of a model in a particular domain - such as law or medicine - or building autonomous agents that can orchestrate actions on a user's behalf.

Requirements: A process that contains an action model, a reward model and an environment for the model to learn from.

Another factor to consider is the VRAM required per each method. The chart below provides an overview of the requirements to run each type of fine-tuning method on Unsloth.

Fine-tuning requirements on Unsloth. Unsloth: A Fast Path to Fine-Tuning on NVIDIA GPUs LLM fine-tuning is a memory- and compute-intensive workload that involves billions of matrix multiplications to update model weights at every training step. This type of heavy parallel workload requires the power of NVIDIA GPUs to complete the process quickly and efficiently.

Unsloth shines at this workload, translating complex mathematical operations into efficient, custom GPU kernels to accelerate AI training.

Unsloth helps boost the performance of the Hugging Face transformers library by 2.5x on NVIDIA GPUs. These GPU-specific optimizations, combined with Unsloth's ease of use, make fine-tuning accessible to a broader community of AI enthusiasts and developers.

The framework is built and optimized for NVIDIA hardware - from GeForce RTX laptops to RTX PRO workstations and DGX Spark - providing peak performance while reducing VRAM consumption.

Unsloth provides helpful guides on how to get started and manage different LLM configurations, hyperparameters and options, along with example notebooks and step-by-step workflows.

Check out some of these Unsloth guides:

Fine-Tuning LLMs With NVIDIA RTX 50 Series GPUs and Unsloth

Fine-Tuning LLMs With NVIDIA DGX Spark and Unsloth

Learn how to install Unsloth on NVIDIA DGX Spark. Read the NVIDIA technical blog for a deep dive of fine-tuning and reinforcement learning on the NVIDIA Blackwell platform.

For a hands-on local fine-tuning walkthrough, watch Matthew Berman showing reinforcement learning running on a NVIDIA GeForce RTX 5090 using Unsloth in the video below.

Available Now: NVIDIA Nemotron 3 Family of Open Models The new Nemotron 3 family of open models - in Nano, Super, and Ultra sizes - built on a new hybrid latent Mixture-of-Experts (MoE) architecture, introduces the most efficient family of open models with leading accuracy, ideal for building agentic AI applications.

Nemotron 3 Nano 30B-A3B, available now, is the most compute-efficient model in the lineup. It's optimized for tasks such as software debugging, content summarization, AI assistant workflows and information retrieval at low inference costs. Its hybrid MoE design delivers:

Up to 60% fewer reasoning tokens, significantly reducing inference cost.

A 1 million-token context window, allowing the model to retain far more information for long, multistep tasks.

Nemotron 3 Super is a high-accuracy reasoning model for multi-agent applications, while Nemotron 3 Ultra is for complex AI applications
LINK: https://blogs.nvidia.com/blog/rtx-ai-garage-fine-tuning-unsloth-dgx-sp...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

29/01/2026

Extension of Invitation to Submit Proposals for Micro-Budget Film Projects 2026 Deadline to 2 February 2026

The National Film and Video Foundation (NFVF), in collaboration with a distribut...

29/01/2026

Hitachi Europe Appoints Michele Fracchiolla as President

Michele Fracchiolla Succeeds Andrew Barr as President of EMEA region from April 1, 2026 London, January 29, 2026 Hitachi Europe Ltd. today announces the appoi...

29/01/2026

L3Harris Technologies Reports Strong Full Year and Fourth Quarter 2025 Results, Initiates 2026 Guidance

MELBOURNE, Fla., January 29, 2026 - L3Harris Technologies (NYSE: LHX) reports fu...

29/01/2026

Nielsen Announces 2025 ARTEY Award Winners Following Record-Breaking Year of Streaming

Bluey' Wins Second Consecutive Top Streaming Title of the Year with 45 Billi...

29/01/2026

Report: Performance TV Ties With Social Media in Driving Ad Results

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

ISE: NDI and OBSBOT Expand Partnership

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

NTCA Asks FCC to Block Nexstar, Tegna Deal

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

FCC Announces Tentative Agenda for February Open Meeting

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

CBS Sports AFC Championship Game Attracts 48.6 Million Viewers

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

Boston Conservatory Orchestra Presents East Coast Premiere of Peter and Leonardo Dugan Piano Concerto

Boston Conservatory Orchestra Presents East Coast Premiere of Peter and Leonardo...

29/01/2026

X-Rite Pantone Appoints Cindy Cooperman as Vice President and General Manager of Pantone

X-Rite Pantone Appoints Cindy Cooperman as Vice President and General Manager of...

29/01/2026

Outback Terror: The Falconio Murder

New two-part true crime documentary, OUTBACK TERROR: THE FALCONIO MURDER, aims to shed new light on a case that continues to intrigue on both sides of the world...

29/01/2026

'Love is Blind: Sweden' Returns for a Third Season - Premiering on March 12

Back to All News Love is Blind: Sweden Returns for a Third Season - Premiering ...

29/01/2026

Unmask Bridgerton' Season 4 With Our Complete Coverage Guide

Back to All News Unmask Bridgerton' Season 4 With Our Complete Coverage Guide Yerin Ha as Sophie Baek and Luke Thompson as Benedict Bridgerton in Season ...

29/01/2026

Extraordinary Crime Mysteries, Mythical Worlds and High-Stakes Psychological Thrillers: Inside Netflix's 2026 Chinese-Language Slate

Back to All News Extraordinary Crime Mysteries, Mythical Worlds and High-Stakes...

29/01/2026

FOX Sports Unveils Historic FIFA World Cup 2026 Broadcast Schedule

FOX Sports Unveils Historic FIFA World Cup 2026 Broadcast Schedule Monumental Slate Features 340 Hours of Live First-Run Programming Across FOX Sports Platfo...

29/01/2026

AI Assistants Head into 2026 on a High Note: Comscore Reports Triple-Digit Growth on Mobile

AI Assistants Head into 2026 on a High Note: Comscore Reports Triple-Digit Growt...

29/01/2026

Broadcom confirms Arvato Systems status as a VCSP partner

Broadcom Confirms Arvato Systems' Status as a VCSP Partner Broadcom Partner Program Update Arvato Systems confirmed as authorized VMware Cloud Service Pr...

29/01/2026

Into the Omniverse: Physical AI Open Models and Frameworks Advance Robots and Autonomous Systems

Editor's note: This post is part of Into the Omniverse, a series focused on ...

29/01/2026

Annette Malone appointed as Chief People Officer RT

RT has today announced that Annette Malone has been appointed to the role of Chief People Officer, RT following a public competition. As Chief People Officer...

29/01/2026

GeForce NOW Brings GeForce RTX Gaming to Linux PCs

Get ready to game - the native GeForce NOW app for Linux PCs is now available in beta, letting Linux desktops tap directly into GeForce RTX performance from the...

28/01/2026

2026 Sundance Film Festival Reveals Short Film Program Award Winners

Top L-R: The Liars, Jazz Infernal, Living with a Visionary Second Row L-R: Paper Trail, The Baddest Speechwriter of All, Crisis Actor Third Row: The Boys and ...

28/01/2026

3 Easy Ways to Discover Music That Fits Your Moment on Spotify

Music discovery should feel intuitive and personal. That's why we're continuing to give you more control, so you can ask for what you want, shape what y...

28/01/2026

From $11B in 2025 Payouts to What We're Building for Artists in 2026

Today, Charlie Hellman, Spotify's Head of Music, shared the following note on the Spotify for Artists blog that the company paid out more than $11 billion t...

28/01/2026

Sediba Scriptwriting Training Programme - Oudtshoorn Municipality (Second Call)

The National Film and Video Foundation (NFVF), in partnership with the Oudtshoorn Municipality, invites aspiring and emerging filmmakers to apply for the Sediba...

28/01/2026

MVP makes a tactical switch to Calrec Argo M

As demand for more complex live sports coverage grows, Balkan broadcast specialist MVP has upgraded its flagship HD1 progressive OB truck with the installation ...

28/01/2026

Aussies' love of travel sees 12% surge in ad investment according to Nielsen

Airlines, cruise and tour operators double down on ad spend as Australians' prioritise travel Sydney January 28, 2026 - New Nielsen Ad Intel data shows a...

28/01/2026

Daniel Finn Joins LABF in Philanthropy Role

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

28/01/2026

Tegna Expands Local News Offering with Revamped Mobile App

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

28/01/2026

Marshall Electronics Unveils CV420 27X UHD Camera at ISE...

Marshall Electronics launches the CV420-27X, its next-generation ultra-high-definition (UHD) IP camera, at ISE 2026 (Stand 4N900). Engineered for modern IP-base...

28/01/2026

TVM Selects Grass Valley Technology for OB Truck Refurbis...

Grass Valley has announced that Television Mobiles Ltd. (TVM), one of Europe's leading independent outside broadcast providers, has carried out a major refu...

28/01/2026

FOR-A to show cutting edge technology at FOMEX 2026

AI, graphics and virtual software power new production capabilities FOR-A is bringing remarkable new technologies to FOMEX, the Future of Media Exhibition (ex...

28/01/2026

Riedel and Media Tailor Deliver Unified Broadcast and AV...

Continuing a longstanding collaboration, Riedel Communications and Nordic media technology company Media Tailor have once again joined forces to deliver a state...

28/01/2026

Pebble appoints Paul Nagle-Smith to drive fulfilment

Pebble has appointed Paul Nagle-Smith as vice president for customer fulfilment, strengthening its senior leadership focus on customer delivery and operational ...

28/01/2026

TV Azteca Strengthens Disaster Recovery Capabilities with...

Cloud playout solutions provider, Veset has announced that leading Mexican broadcaster, TV Azteca is using Veset Nimbus on AWS as a disaster recovery (DR) playo...

28/01/2026

MVP kicks off major football tournament with a tactical s...

Ensuring it can keep pace with a rapidly evolving live sports market, Balkan broadcast facility provider MVP Most Valuable Production has upgraded its flags...

28/01/2026

Akamai and Yospace Deliver Seamless Personalized Ad Exper...

Akamai Technologies, Inc. (NASDAQ: AKAM), the cloud solutions provider that powers and protects life online, and Yospace, the leader in dynamic ad insertion tec...

28/01/2026

Clear-Com Empowers Reykjavik City Theatre with New Upgrad...

The renowned Reykjavik City Theatre (RCT) recently underwent a major intercom system upgrade using Clear-Com solutions. This milestone project utilizes Clear-C...

28/01/2026

SES Acknowledges Fitch's Rating Action and Reiterates Deleveraging Plan

Luxembourg, January 26, 2026 - SES S.A. ( SES or the Company ), a leading space solutions company, acknowledges the credit rating action announced by Fitch to...

28/01/2026

OpenDrives Announces New Funding, Appoints Trevor Morgan CEO

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

28/01/2026

AWARN Alliance Backs ATSC Sunset, NextGen TV Security Measures

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

28/01/2026

More Than Two Dozen Groups Tell FCC to Reject Nexstar-Tegna Deal

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

28/01/2026

Screen Australia refreshes Market & Audience approach to increase the impact of local content

28 01 2026 - Media release Screen Australia refreshes Market & Audience approach...

28/01/2026

Boston Conservatory Orchestra Premieres a New Piano Concerto by Peter and Leonardo Dugan

Boston Conservatory Orchestra Premieres a New Piano Concerto by Peter and Leonar...

28/01/2026

Netflix's 'Kohrra' Season 2 Unveils A Thrilling Whodunnit Trailer Where The Truth Looks Foggier Than It Is!

Back to All News Netflix's Kohrra Season 2 Unveils A Thrilling Whodunnit Tr...

28/01/2026

What Next? Netflix Presents the Latest German-Speaking Series, Films and Non-Fction Highlights, Live in Berlin

Back to All News What Next? Netflix Presents the Latest German-Speaking Series,...