Sony Pixel Power calrec Sony

Brave New World: Leo AI and Ollama Bring RTX-Accelerated Local LLMs to Brave Browser Users

02/10/2024

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and showcases new hardware, software, tools and accelerations for GeForce RTX PC and NVIDIA RTX workstation users.

From games and content creation apps to software development and productivity tools, AI is increasingly being integrated into applications to enhance user experiences and boost efficiency.

Those efficiency boosts extend to everyday tasks, like web browsing. Brave, a privacy-focused web browser, recently launched a smart AI assistant called Leo AI that, in addition to providing search results, helps users summarize articles and videos, surface insights from documents, answer questions and more.

Leo AI helps users summarize articles and videos, surface insights from documents, answer questions and more. The technology behind Brave and other AI-powered tools is a combination of hardware, libraries and ecosystem software that's optimized for the unique needs of AI.

Why Software Matters NVIDIA GPUs power the world's AI, whether running in the data center or on a local PC. They contain Tensor Cores, which are specifically designed to accelerate AI applications like Leo AI through massively parallel number crunching - rapidly processing the huge number of calculations needed for AI simultaneously, rather than doing them one at a time.

But great hardware only matters if applications can make efficient use of it. The software running on top of GPUs is just as critical for delivering the fastest, most responsive AI experience.

The first layer is the AI inference library, which acts like a translator that takes requests for common AI tasks and converts them to specific instructions for the hardware to run. Popular inference libraries include NVIDIA TensorRT, Microsoft's DirectML and the one used by Brave and Leo AI via Ollama, called llama.cpp.

Llama.cpp is an open-source library and framework. Through CUDA - the NVIDIA software application programming interface that enables developers to optimize for GeForce RTX and NVIDIA RTX GPUs - provides Tensor Core acceleration for hundreds of models, including popular large language models (LLMs) like Gemma, Llama 3, Mistral and Phi.

On top of the inference library, applications often use a local inference server to simplify integration. The inference server handles tasks like downloading and configuring specific AI models so that the application doesn't have to.

Ollama is an open-source project that sits on top of llama.cpp and provides access to the library's features. It supports an ecosystem of applications that deliver local AI capabilities. Across the entire technology stack, NVIDIA works to optimize tools like Ollama for NVIDIA hardware to deliver faster, more responsive AI experiences on RTX.

Applications like Brave's Leo AI can access RTX-powered AI acceleration to enhance user experiences. NVIDIA's focus on optimization spans the entire technology stack - from hardware to system software to the inference libraries and tools that enable applications to deliver faster, more responsive AI experiences on RTX.

Local vs. Cloud Brave's Leo AI can run in the cloud or locally on a PC through Ollama.

There are many benefits to processing inference using a local model. By not sending prompts to an outside server for processing, the experience is private and always available. For instance, Brave users can get help with their finances or medical questions without sending anything to the cloud. Running locally also eliminates the need to pay for unrestricted cloud access. With Ollama, users can take advantage of a wider variety of open-source models than most hosted services, which often support only one or two varieties of the same AI model.

Users can also interact with models that have different specializations, such as bilingual models, compact-sized models, code generation models and more.

RTX enables a fast, responsive experience when running AI locally. Using the Llama 3 8B model with llama.cpp, users can expect responses up to 149 tokens per second - or approximately 110 words per second. When using Brave with Leo AI and Ollama, this means snappier responses to questions, requests for content summaries and more.

NVIDIA internal throughput performance measurements on NVIDIA GeForce RTX GPUs, featuring a Llama 3 8B model with an input sequence length of 100 tokens, generating 100 tokens. Get Started With Brave With Leo AI and Ollama Installing Ollama is easy - download the installer from the project's website and let it run in the background. From a command prompt, users can download and install a wide variety of supported models, then interact with the local model from the command line.

For simple instructions on how to add local LLM support via Ollama, read the company's blog. Once configured to point to Ollama, Leo AI will use the locally hosted LLM for prompts and queries. Users can also switch between cloud and local models at any time.

Brave with Leo AI running on Ollama and accelerated by RTX is a great way to get more out of your browsing experience. You can even summarize and ask questions about AI Decoded blogs! Developers can learn more about how to use Ollama and llama.cpp in the NVIDIA Technical Blog.

Generative AI is transforming gaming, videoconferencing and interactive experiences of all kinds. Make sense of what's new and what's next by subscribing to the AI Decoded newsletter.
LINK: https://blogs.nvidia.com/blog/rtx-ai-brave-browser/...
See more stories from nvidia

North America Stories

25/06/2026

Sports and Dramas Drive April Viewing Patterns in Nielsen's Latest Gauge Reports

Cable Gains Share for Second Consecutive Month in Six-Month-High Finish, Boosted...

25/06/2026

Jeopardy!, Wheel of Fortune Tap Clear-Com for Comms Upgrade

Share Copy link Facebook X Linkedin Bluesky Email...

25/06/2026

ESC 2026 - Big Blue Marble cloud-based delivery powers fl...

The Eurovision Song Contest 2026 in Vienna was a significant success for the Austrian public broadcaster ORF. In Austria, more than 1.5 million viewers tuned in...

25/06/2026

WISYCOM DELIVERS NEW LEVELS OF RF EFFICIENCY AND INFRAST...

Wisycom has further strengthened its ecosystem of professional wireless solutions with the MPR60 Wideband IEM/IFB Receiver with expanded multichannel IFB mode, ...

25/06/2026

Ease Live Powers New Interactive Experience for Rally TV...

Ease Live, the interactivity expert, today announced that its graphics overlay platform is powering a new interactive experience on Rally.TV, the official video...

25/06/2026

VFX History: the origin of After Effects

VFX History: the origin of After Effects Graham Quince June 25, 2026 0 Comments Before it was Adobe, it was CoSA. This is the VFX history of Adobe Aft...

25/06/2026

Creative Remote to Open London Offline Facility

Creative Remote, the provider of remote and hybrid offline editing infrastructure, today announced the opening of 41, its new offline edit facility located at 4...

25/06/2026

Rise Announces 2026 Worldwide Mentoring Cohorts Supportin...

Rise, the award-winning advocacy group for gender diversity in the broadcast and media technology sector, is pleased to announce the global mentoring cohort for...

25/06/2026

Emergent Partners with ROCKET to Expand Canadian Operatio...

Emergent, a pioneer in browser-based, AI-enhanced content production environments, today announced a strategic partnership with ROCKET, a premier media-centric ...

25/06/2026

Mobile Television Group Launches MTVG Full-Stack Production Platform

Share Copy link Facebook X Linkedin Bluesky Email...

25/06/2026

NAB Updates FCC on ATSC 3.0 Alerting Advances

Share Copy link Facebook X Linkedin Bluesky Email...

25/06/2026

Tegna Elevates Four Executives to Senior VP

Share Copy link Facebook X Linkedin Bluesky Email...

25/06/2026

The Ultimate Summer Sale Pairing: Steam Sale Meets GeForce NOW Discounts

Summer savings are heating up. From the Steam Summer Sale to GeForce NOW membership discounts, this week's GFN Thursday delivers double the deals and more w...

25/06/2026

June 24, 2026

Immune molecule may drive excessive drinking in alcohol use disorder Scripps Research scientists showed that blocking an immune molecule tied to inflammation r...

24/06/2026

Nielsen's Q1 2026 Ad Supported Gauge

Streaming sets record high of 46.6% of ad supported TV viewing, driven by Super Bowl and Winter Olympics; overall share of ad supported TV remains steady NEW Y...

24/06/2026

FCC Flooded with Nearly 28K Comments Regarding Its Probe of 'The View'

Share Copy link Facebook X Linkedin Bluesky Email...

24/06/2026

Hearst Television Brings Ad Addressability to Local Broadcast TV

Share Copy link Facebook X Linkedin Bluesky Email...

24/06/2026

FCC Raises $3.5 Billion in AWS-3 Wireless Auction

Share Copy link Facebook X Linkedin Bluesky Email...

24/06/2026

RE:Vision Effects Announces Twixtor Standalone v 8.1 and a Sale!

RE:Vision Effects Announces Twixtor Standalone v 8.1 and a Sale! Brie Clayton June 24, 2026 0 Comments Twixtor v8.1 Standalone adds support for variab...

24/06/2026

Dreamtek Uses Full Blackmagic Workflow for Vercel Next JS Event

Dreamtek Uses Full Blackmagic Workflow for Vercel Next JS Event Brie Clayton June 24, 2026 0 Comments Blackmagic cameras, switchers, routers, recorder...

24/06/2026

Chyron LIVE Unveils New Features: Haivision StreamHub Integration, SCTE-35 Ad Insertion, and Refined Switching Tools

Chyron LIVE Unveils New Features: Haivision StreamHub Integration, SCTE-35 Ad In...

24/06/2026

Mapping an Education

Mapping an Education How composer Chloe Clarke Smith navigated her Boston Conservatory experience and brought new meaning to her work June 24, 2026 By Sara...

24/06/2026

The Next Act

The Next Act Dean Krisha Marcano's vision for a connected Theater Division, and the fund making it possible June 24, 2026 Photo by Eric Antoniou The Or...

24/06/2026

Announcing STAGES Magazine 2026

Announcing STAGES Magazine 2026 Marking a decade since Boston Conservatory and Berklee College of Music joined forces, this issue spotlights some of the groun...

24/06/2026

Rede Legislativa Chooses Appear to Support Brazil TV Ver...

In Brazil's TV 3.0 Trials, Appear's X5 is transporting live signals from Bras lia to S o Paulo over the public internet using secure, reliable next-gene...

24/06/2026

Mediaproxy partners with HVS for US broadcast market

Melbourne, Australia - 24 June 2026: Mediaproxy, the global standard for software-based IP compliance monitoring and multiviewing solutions, has named Heartland...

24/06/2026

Gray Media Launches Political 360 Digital Advertising Solution

Share Copy link Facebook X Linkedin Bluesky Email...

24/06/2026

Walmart to Pay $1.4 Billion to Acquire Ad Tech Firm Vibe.co

Share Copy link Facebook X Linkedin Bluesky Email...

24/06/2026

FCC Flooded with Nearly 28K Comments on 'The View'

Share Copy link Facebook X Linkedin Bluesky Email...

24/06/2026

First Rush Brings SDI Multicam ProRes Recording to Apple Silicon Macs

First Rush Brings SDI Multicam ProRes Recording to Apple Silicon Macs Brie Clayton June 23, 2026 0 Comments First Rush is a native macOS application d...

24/06/2026

Vertical Drama Beneath Crimson Sails Created with Blackmagic Design

Vertical Drama Beneath Crimson Sails Created with Blackmagic Design Brie Clayton June 23, 2026 0 Comments Thunder Child Productions relies on cameras&...

24/06/2026

Seven paradoxes shaping the next era of media production - Episode 1

Why Dynamic Media Facilities Matter In this series, we explore the technologies, architectures and operational realities shaping modern media operations. Along ...

23/06/2026

Case Study: YES Networks IP Transition Expands Production Possibilities and Redefines Workflows

When we began planning our transition from an SDI-based infrastructure to a new ...

23/06/2026

Imagine Communications Appoints Greg Garmon as SVP, Americas Video Sales

Imagine Communications has announced the appointment of Greg Garmon as Senior Vice President, Americas Video Sales. Garmon will oversee account growth and busin...

23/06/2026

Snap Promotes Emma Wakely to Head of Sports and Media Partnerships, Americas

Snap has promoted Emma Wakely to Head of Sports and Media Partnerships, Americas, succeeding Anmol Malhotra, who has been elevated to Global Head of Content and...

23/06/2026

YES Network and Gotham Sports App to Air MI New York Major League Cricket Matches

YES Network and The Gotham Sports App will air MI New York's Major League Cr...

23/06/2026

HAND Issues Persistent Digital IDs to 2026 NBA Draft Class

The Universal Talent Identifier (HAND) has issued HAND IDs to 34 top projected prospects in the 2026 NBA Draft class, including AJ Dybantsa, Cameron Boozer, and...

23/06/2026

World Boxing Launches World Boxing TV Streaming Platform

World Boxing has announced the launch of World Boxing TV, a subscription-based streaming platform built on the Joymo platform, offering live events, on-demand c...

23/06/2026

FloRacing to Stream 32 Off-Road Motorcycle Racing Events Including AMA Amateur National Motocross Championship

FloSports will stream 32 off-road motorcycle racing events on FloRacing, includi...

23/06/2026

SES Adds 14 Regional Channels and New Set-Top Boxes to ASTRA TV in Spain

SES has announced the expansion of its ASTRA TV platform in Spain with the addition of 14 regional channels in HD and UHD quality and the launch of new hybrid s...

23/06/2026

Appear Supports Rede Legislativas Contribution Workflow for Brazils TV 3.0 Trials

Appear ASA has announced its role in Rede Legislativa de R dio e TV's contri...

23/06/2026

PBS Selects LTN for Nationwide IP Video Network Across 330 Member Stations

LTN has announced that PBS has selected it as its IP video partner to modernize content distribution and contribution across more than 330 public television sta...

23/06/2026

Ease Live Powers Interactive Experience on Rally.TV for WRC

Ease Live has announced that its graphics overlay platform is powering an interactive fan experience on Rally.TV, the official streaming platform of the FIA Wor...

23/06/2026

Chyron LIVE Adds Haivision StreamHub Integration, SCTE-35 Ad Insertion, and Switcher Updates

Chyron has announced updates to Chyron LIVE, its cloud-native live production pl...

23/06/2026

ESPN Announces ESPN Fan House, Fan Engagement Hub Powered by Flowcode

ESPN has announced ESPN Fan House, a fan engagement hub powered by Flowcode, launching in August ahead of the 2026 college football season. Publicis Sports will...

23/06/2026

Sennheiser Relocates Americas Regional Hub to Nashville

The city's solid position in broadcast, entertainment, and sports attracted the major microphone manufacturer Sennheiser Group is moving its Americas Regio...

23/06/2026

IAB Tech Lab Releases SupplyChain v1.1

Share Copy link Facebook X Linkedin Bluesky Email...

23/06/2026

Besco to Represent PlayBox Neo in South Korea

PlayBox Neo appoints Besco as Channel Reseller to establish a firm foothold in Asia Pacific's thriving high-tech export-driven economic boom PlayBox Neo, t...

23/06/2026

PBS Selects LTN to Power Nationwide IP Video Network

Share Copy link Facebook X Linkedin Bluesky Email...

23/06/2026

PBS selects LTN for nationwide IP video network

LTN, a global leader in IP-based video transport and network services, today announced that PBS has selected LTN as its IP video partner to modernize and future...