Sony Pixel Power calrec Sony

How to Get Started With Large Language Models on NVIDIA RTX PCs

01/10/2025

Many users want to run large language models (LLMs) locally for more privacy and control, and without subscriptions, but until recently, this meant a trade-off in output quality. Newly released open-weight models, like OpenAI's gpt-oss and Alibaba's Qwen 3, can run directly on PCs, delivering useful high-quality outputs, especially for local agentic AI.

This opens up new opportunities for students, hobbyists and developers to explore generative AI applications locally. NVIDIA RTX PCs accelerate these experiences, delivering fast and snappy AI to users.

Getting Started With Local LLMs Optimized for RTX PCs NVIDIA has worked to optimize top LLM applications for RTX PCs, extracting maximum performance of Tensor Cores in RTX GPUs.

One of the easiest ways to get started with AI on a PC is with Ollama, an open-source tool that provides a simple interface for running and interacting with LLMs. It supports the ability to drag and drop PDFs into prompts, conversational chat and multimodal understanding workflows that include text and images.

It's easy to use Ollama to generate answers from a text simple prompt. NVIDIA has collaborated with Ollama to improve its performance and user experience on GeForce RTX GPUs. The most recent developments include:

50% performance improvements on OpenAI's gpt-oss-20B model

60% performance improvements on the new Gemma 3 270M and EmbeddingGemma models for hyper-efficient RAG

Improved model scheduling system to maximize and accurately report memory utilization

Stability enhancements to reduce the number of crashes

Ollama is a developer framework that can be used with other applications. For example, AnythingLLM - an open-source app that lets users build their own AI assistants powered by any LLM - can run on top of Ollama and benefit from all of its accelerations.

Enthusiasts can also get started with local LLMs using LM Studio, an app powered by the popular llama.cpp framework. The app provides a user-friendly interface for running models locally, letting users load different LLMs, chat with them in real time and even serve them as local application programming interface endpoints for integration into custom projects.

Example of using LM Studio to generate notes accelerated by NVIDIA RTX. NVIDIA has worked with llama.cpp to optimize performance on NVIDIA RTX GPUs. The latest updates include:

Support for the latest NVIDIA Nemotron Nano v2 9B model, which is based on the novel hybrid-mamba architecture

Flash Attention now turned on by default, offering an up to 20% performance improvement compared with Flash Attention being turned off

CUDA kernels optimizations for RMS Norm and fast-div based modulo, resulting in up to 9% performance improvements for popular model

Semantic versioning, making it easy for developers to adopt future releases

Learn more about gpt-oss on RTX and how NVIDIA has worked with LM Studio to accelerate LLM performance on RTX PCs.

Creating an AI-Powered Study Buddy With AnythingLLM In addition to greater privacy and performance, running LLMs locally removes restrictions on how many files can be loaded or how long they stay available, enabling context-aware AI conversations for a longer period of time. This creates more flexibility for building conversational and generative AI-powered assistants.

For students, managing a flood of slides, notes, labs and past exams can be overwhelming. Local LLMs make it possible to create a personal tutor that can adapt to individual learning needs.

The demo below shows how students can use local LLMs to build a generative-AI powered assistant:

AnythingLLM running on an RTX PC transforms study materials into interactive flashcards, creating a personalized AI-powered tutor. A simple way to do this is with AnythingLLM, which supports document uploads, custom knowledge bases and conversational interfaces. This makes it a flexible tool for anyone who wants to create a customizable AI to help with research, projects or day-to-day tasks. And with RTX acceleration, users can experience even faster responses.

By loading syllabi, assignments and textbooks into AnythingLLM on RTX PCs, students can gain an adaptive, interactive study companion. They can ask the agent, using plain text or speech, to help with tasks like:

Generating flashcards from lecture slides: Create flashcards from the Sound chapter lecture slides. Put key terms on one side and definitions on the other.

Asking contextual questions tied to their materials: Explain conservation of momentum using my Physics 8 notes.

Creating and grading quizzes for exam prep: Create a 10-question multiple choice quiz based on chapters 5-6 of my chemistry textbook and grade my answers.

Walking through tough problems step by step: Show me how to solve problem 4 from my coding homework, step by step.

Beyond the classroom, hobbyists and professionals can use AnythingLLM to prepare for certifications in new fields of study or for other similar purposes. And running locally on RTX GPUs ensures fast, private responses with no subscription costs or usage limits.

Project G-Assist Can Now Control Laptop Settings Project G-Assist is an experimental AI assistant that helps users tune, control and optimize their gaming PCs through simple voice or text commands - without needing to dig through menus. Over the next day, a new G-Assist update will roll out via the home page of the NVIDIA App.

Project G-Assist helps users tune, control and optimize their gaming PCs through simple voice or text commands. Building on its new, more efficient AI model and support for the majority of RTX GPUs released in August, the new G-Assist update adds commands to adjust laptop settings, including:

App profiles optimized for laptops: Automatically adjust games or apps for efficiency, quality or a balance when laptops aren't connected to chargers.

Bat
LINK: https://blogs.nvidia.com/blog/rtx-ai-garage-how-to-get-started-with-ll...
See more stories from nvidia

More from Nvidia

04/12/2025

Robots' Holiday Wishes Come True: NVIDIA Jetson Platform Offers High-Performance Edge AI at Festive Prices

Developers, researchers, hobbyists and students can take a byte out of holiday s...

04/12/2025

Game the Halls: GeForce NOW Brings Holiday Cheer With 30 New Games in the Cloud

Editor's note: The Game Pass edition of Hogwarts Legacy' will also be supported on GeForce NOW when the Steam and Epic Games Store versions launch on t...

03/12/2025

Mixture of Experts Powers the Most Intelligent Frontier AI Models, Runs 10x Faster on NVIDIA Blackwell NVL72

The top 10 most intelligent open-source models all use a mixture-of-experts arch...

02/12/2025

NVIDIA Partners With Mistral AI to Accelerate New Family of Open Models

Today, Mistral AI announced the Mistral 3 family of open-source multilingual, multimodal models, optimized across NVIDIA supercomputing and edge platforms. M...

02/12/2025

NVIDIA and AWS Expand Full-Stack Partnership, Providing the Secure, High-Performance Compute Platform Vital for Future Innovation

At AWS re:Invent, NVIDIA and Amazon Web Services expanded their strategic collab...

01/12/2025

At NeurIPS, NVIDIA Advances Open Model Development for Digital and Physical AI

Researchers worldwide rely on open-source technologies as the foundation of their work. To equip the community with the latest advancements in digital and physi...

27/11/2025

The Ultimate Black Friday Deal Is Here

Black Friday is leveling up. Get ready to score one of the biggest deals of the season - 50% off the first three months of a new GeForce NOW Ultimate membership...

25/11/2025

FLUX.2 Image Generation Models Now Released, Optimized for NVIDIA RTX GPUs

Black Forest Labs - the frontier AI research lab developing visual generative AI models - today released the FLUX.2 family of state-of-the-art image generation ...

24/11/2025

AI On: 3 Ways Specialized AI Agents Are Reshaping Businesses

Editor's note: This post is part of the AI On blog series, which explores the latest techniques and real-world applications of agentic AI, chatbots and copi...

20/11/2025

Into the Omniverse: How Smart City AI Agents Transform Urban Operations

Editor's note: This post is part of Into the Omniverse, a series focused on how developers, 3D practitioners and enterprises can transform their workflows u...

20/11/2025

Ultimate Cloud Gaming Is Everywhere With GeForce NOW

The NVIDIA Blackwell RTX upgrade is nearing the finish line, letting GeForce NOW Ultimate members across the globe experience true next-generation cloud gaming ...

20/11/2025

The Largest Digital Zoo: Biology Model Trained on NVIDIA GPUs Identifies Over a Million Species

Tanya Berger-Wolf's first computational biology project started as a bet wit...

18/11/2025

Powering AI Superfactories, NVIDIA and Microsoft Integrate Latest Technologies for Inference, Cybersecurity, Physical AI

Timed with the Microsoft Ignite conference running this week, NVIDIA is expandin...

18/11/2025

Microsoft, NVIDIA and Anthropic Announce Strategic Partnerships

Today, Microsoft, NVIDIA and Anthropic announced new strategic partnerships. Anthropic is scaling its rapidly growing Claude AI model on Microsoft Azure, powere...

18/11/2025

Delivering AI-Ready Enterprise Data With GPU-Accelerated AI Storage

AI agents have the potential to become indispensable tools for automating complex tasks. But bringing agents to production remains challenging. According to Ga...

17/11/2025

One Giant Leap for AI Physics: NVIDIA Apollo Unveiled as Open Model Family for Scientific Simulation

NVIDIA Apollo - a family of open models for accelerating industrial and computat...

17/11/2025

NVIDIA Accelerated Computing Enables Scientific Breakthroughs for Materials Discovery

To power future technologies including liquid-cooled data centers, high-resoluti...

17/11/2025

Accelerated Computing, Networking Drive Supercomputing in Age of AI

At SC25, NVIDIA unveiled advances across NVIDIA BlueField DPUs, next-generation networking, quantum computing, national research, AI physics and more - as accel...

17/11/2025

NVIDIA Accelerates AI for Over 80 New Science Systems Worldwide

Across quantum physics, digital biology and climate research, the world's researchers are harnessing a universal scientific instrument to chart new frontier...

17/11/2025

The Great Flip: How Accelerated Computing Redefined Scientific Systems - and What Comes Next

It used to be that computing power trickled down from hulking supercomputers to ...

14/11/2025

How to Unlock Accelerated AI Storage Performance With RDMA for S3-Compatible Storage

Today's AI workloads are data-intensive, requiring more scalable and afforda...

13/11/2025

AI On: 3 Ways to Bring Agentic AI to Computer Vision Applications

Editor's note: This post is part of the AI On blog series, which explores the latest techniques and real-world applications of agentic AI, chatbots and copi...

13/11/2025

GeForce NOW Enlists Call of Duty: Black Ops 7' for the Cloud

Chaos has entered the chat. It's GFN Thursday, and things are getting intense with the launch of Call of Duty: Black Ops 7, streaming at launch this week on...

12/11/2025

NVIDIA Wins Every MLPerf Training v5.1 Benchmark

In the age of AI reasoning, training smarter, more capable models is critical to scaling intelligence. Delivering the massive performance to meet this new age r...

12/11/2025

Faster Than a Click: Hyperlink Agent Search Now Available on NVIDIA RTX PCs

Large language model (LLM)-based AI assistants are powerful productivity tools, but without the right context and information, they can struggle to provide nuan...

10/11/2025

Think SMART: New NVIDIA Dynamo Integrations Simplify AI Inference at Data Center Scale

Editor's note: This post is part of Think SMART, a series focused on how lea...

06/11/2025

NVIDIA Founder and CEO Jensen Huang and Chief Scientist Bill Dally Awarded Prestigious Queen Elizabeth Prize for Engineering

NVIDIA founder and CEO Jensen Huang and chief scientist Bill Dally were honored ...

06/11/2025

Fall Into Gaming With 20+ Titles Joining GeForce NOW in November

Editor's note: This blog has been updated to reflect the correct launch date for Call of Duty: Black Ops 7', November 14. A crisp chill's in the...

04/11/2025

Deutsche Telekom and NVIDIA Launch Industrial AI Cloud - a New Era' for Germany's Industrial Transformation

In Berlin on Tuesday, Deutsche Telekom and NVIDIA unveiled the world's first...

04/11/2025

How NVIDIA GeForce RTX GPUs Power Modern Creative Workflows

When inspiration strikes, nothing kills momentum faster than a slow tool or a frozen timeline. Creative apps should feel fast and fluid - an extension of imagin...

03/11/2025

NVIDIA Partners Bring Physical AI, New Smart City Technologies to Dublin, Ho Chi Minh City, Raleigh and More

Two out of every three people are likely to be living in cities or other urban c...

31/10/2025

Korea Joins AI Industrial Revolution: NVIDIA CEO Jensen Huang Unveils Historic Partnership at APEC Summit

Amidst Gyeongju, South Korea's ancient temples and modern skylines, Jensen H...

30/10/2025

AI-Powered Mobile Clinics Deliver Breast Cancer Screening to India's Rural Communities

An unassuming van driving around rural India uses powerful AI technology that...

30/10/2025

Join the Resistance: ARC Raiders' Launches in the Cloud

Get ready, raiders - the wait is over. ARC Raiders is dropping onto GeForce NOW and bringing the fight from orbit to the screen. To celebrate the launch, gamer...

29/10/2025

Into the Omniverse: Open World Foundation Models Generate Synthetic Worlds for Physical AI Development

Editor's note: This post is part of Into the Omniverse, a series focused on ...

28/10/2025

NVIDIA and US Technology Leaders Unveil AI Factory Design to Modernize Government and Secure the Nation

Governments everywhere are racing to harness the power of AI - but legacy infras...

28/10/2025

NVIDIA IGX Thor Robotics Processor Brings Real-Time Physical AI to the Industrial and Medical Edge

AI is moving from the digital world into the physical one. Across factory floors...

28/10/2025

NVIDIA Open Sources Aerial Software to Accelerate AI-Native 6G

NVIDIA is delivering the telecom industry a major boost in open-source software for building AI-native 5G and 6G networks. NVIDIA Aerial software will soon be ...

28/10/2025

NVIDIA and General Atomics Advance Commercial Fusion Energy

The race to bottle a star now runs on AI. NVIDIA, General Atomics and a team of international partners have built a high-fidelity, AI-enabled digital twin for ...

28/10/2025

NVIDIA, NPS Commission the Navy's AI Flagship for Training Tomorrow's Leaders

Along the Pacific Ocean in Monterey, California, the Naval Postgraduate School (...

28/10/2025

Fueling Economic Development Across the US: How NVIDIA Is Empowering States, Municipalities and Universities to Drive Innovation

To democratize access to AI technology nationwide, AI education and deployment c...

28/10/2025

NVIDIA AI Physics Transforms Aerospace and Automotive Design, Accelerating Engineering by 500x

Leading technology companies in aerospace and automotive are accelerating their ...

26/10/2025

NVIDIA Contributes to Open Frameworks for Next-Generation Robotics Development

This year's ROSCon conference heads to Singapore, bringing together the global robotics developer community behind Robot Operating System (ROS) - the world&...

24/10/2025

NVIDIA GTC DC: Live Updates on What's Next in AI

Monday, Oct. 27, 12:30 p.m. How Medium-Sized Cities Are Tackling AI Readiness L to R: Mark Muro, senior fellow at Brookings Metro; Micah Runner, city manag...

23/10/2025

Fangs Out, Frames Up: Vampire: The Masquerade - Bloodlines 2' Leads a Killer GFN Thursday

The nights grow longer and the shadows get bolder with Vampire The Masquerade: B...

21/10/2025

UC Santa Cruz Maps Coastal Flooding With NVIDIA Accelerated Computing

Coastal communities in the U.S. have a 26% chance of flooding within a 30-year period. This percentage is expected to increase due to climate-change-driven sea-...

20/10/2025

NVIDIA and Google Cloud Accelerate Enterprise AI and Industrial Digitalization

NVIDIA and Google Cloud are expanding access to accelerated computing to transform the full spectrum of enterprise workloads, from visual computing to agentic a...

17/10/2025

Open Source AI Week - How Developers and Contributors Are Advancing AI Innovation

As Open Source AI Week comes to a close, we're celebrating the innovation, c...

17/10/2025

The Engines of American-Made Intelligence: NVIDIA and TSMC Celebrate First NVIDIA Blackwell Wafer Produced in the US

AI has ignited a new industrial revolution. NVIDIA and TSMC are working togethe...

16/10/2025

Ready, Set, Reward - GeForce NOW Membership Rewards Await

GeForce NOW is more than just a platform to stream fresh games every week - it offers celebrations for the gamers who make it epic, with member rewards to sweet...