
As of today, NVIDIA now supports the general availability of Gemma 3n on NVIDIA RTX and Jetson. Gemma, previewed by Google DeepMind at Google I/O last month, includes two new models optimized for multi-modal on-device deployment.
Gemma now includes audio in addition to the text and vision capabilities introduced in version 3.5. Each component integrates trusted research models: Universal Speech Model for audio, MobileNet v4 for vision, and MatFormer for text.
The biggest usage advancement is an innovation called Per-Lay Embeddings. It allows for significant reduction in RAM usage for parameters. The Gemma 3n E4B model has a raw parameter count of 8B parameters but can operate using a dynamic memory footprint that's comparable to a 4B model. This enables developers to use a higher quality model within a resource-constrained environment.
Model name Raw Parameters Input Context Length Output Context Length Size on Disk
E2B 5B 32K 32K subtracting request input 1.55GB
E4B 8B 32K 32K subtracting request input 2.82BB
Table 1: Gemma 3n model components for both the E2B and E4B model Powering robotics and edge AI with Jetson The Gemma family of models works well on NVIDIA Jetson devices that are geared at powering edge applications, such as next-generation robotics. The lightweight architecture and, now, dynamic memory usage fit in resource-constrained environments.
Jetson developers can participate in the Gemma 3n Impact Challenge hosted on Kaggle. The aim is to use this technology to create meaningful, positive change in the world in areas such as accessibility, education, healthcare, environmental sustainability, and crisis response. Several cash prizes, which start at $10,000, are available for submissions for overall placement and for using different technologies suited for on-device deployment, such as Jetson.
To get started, check out the live text and image demo from the Gemma 3 Developer Day in April and the GitHub repository for deploying Gemma locally using Ollama.
NVIDIA RTX for Windows developers and AI enthusiasts With NVIDIA RTX AI PCs, developers can easily deploy Gemma 3n models using Ollama. AI enthusiasts can use Gemma 3n models with RTX accelerations in their favorite apps like AnythingLLM and LM Studio.
Developers can deploy Gemma 3n locally to both RTX and Jetson devices with a few simple instructions using the Ollama CLI:
Download and install Ollama for Windows
Open a terminal window and complete the following commands:
ollama pull gemma3n:e4b ollama run gemma3n:e4b Summarize Shakespeare's Hamlet
NVIDIA collaborates with Ollama to provide performance optimizations for NVIDIA RTX GPUs, accelerating the latest models like Gemma 3n. For this model, Ollama leverages the Ollama engine in the backend, which builds upon the GGML library. Learn more about NVIDIA's contributions to the GGML library for maximum performance on NVIDIA RTX GPUs.
Customize Gemma for your data with the open NVIDIA NeMo Framework Developers can use the Gemma 3n models from Hugging Face with the open source NVIDIA NeMo Framework. It provides a comprehensive framework for post-training Llama models to achieve higher accuracy, specifically through fine-tuning with enterprise-specific data. The workflow within NeMo is designed to be end-to-end, covering data preparation, efficient fine-tuning, and model evaluation.
data-src=https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-png.webp alt=A diagram showing the workflow of NeMo Framework. It provides end-to-end support for developing large language models (LLMs) and multimodal models (MMs). class=lazyload wp-image-102645 data-srcset=https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-png.webp 1600w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-300x169-png.webp 300w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-625x352-png.webp 625w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-179x101-png.webp 179w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-768x432-png.webp 768w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-1536x864-png.webp 1536w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-645x363-png.webp 645w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-660x370-png.webp 660w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-500x281-png.webp 500w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-160x90-png.webp 160w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-362x204-png.webp 362w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-196x110-png.webp 196w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-1024x576-png.webp 1024w, https://developer-blogs.nvidia.com/wp-content/uploads/2025/06/powerade-fig-1-960x540-png.webp 960w data-sizes=(max-width: 1600px) 100vw, 1600px />
Figure 1. NeMo Framework provides end-to-end support for large language models and multimodal models.
The workflow includes:
Data curation (NeMo Curator): Curator prepares high-quality datasets for either pretraining or fine-tuning by offering tools to extract, filter, and deduplicate large volumes of structured and unstructured data. It ensures the quality of the input data for the model.
Fine-tuning (NeMo): Once the data is curated, NeMo enables efficient fine-tuning of Llama models. It supports various techniques to optimize this process, including LoRA (Low-Rank Adaptation), PEFT (Parameter-Efficient Fine-Tuning), and full parameter tuning for comprehensive customization.
Model evaluation (NeMo Evaluator): After fine-tuning, NeMo Evaluator is used
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
04/08/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
04/07/2026
April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...
01/06/2026
January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026
Throughout the week, Dolby brings to life the latest innovatio...
02/05/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
01/05/2026
January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...
30/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
30/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
30/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
30/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
30/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
30/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
30/04/2026
A flexible monitoring platform designed to simplify ST 2110 operations, consolidate vendor tools, and support modern live production environments. See it at NAB...
30/04/2026
Sundance-premiering Run Amok is an expressive, unconventional take on today's teen experience, replete with musical numbers. When cinematographer Shachar ...
30/04/2026
JACKSONVILLE, AL, APRIL 29, 2026 Jacksonville State University, known as Jax State and a proud NCAA Division I member of Conference USA, has transformed its a...
30/04/2026
Transforming viewer and content data into real-time intelligence to deliver relevant streaming experiences at scale
ThinkAnalytics, the global leader in AI-pow...
30/04/2026
Sports Production, Delivery is Big Biz at NAB
Andy Marken April 29, 2026
0 Comments
Hero image source: NAB
One of the neat things about trade shows i...
30/04/2026
Scripps Research ranks third in 2026 Cure Innovation Index
April 29, 2026
LA JOLLA, CA Scripps Research ranked third in the inaugural 2026 Cure Innovation In...
29/04/2026
It was a delicate job in a 150-year-old venue laden with traditions. Begun at th...
29/04/2026
In annual event, the league gives startup companies the opportunity to prove the...
29/04/2026
Panel discussions, networking, and a facility tour will take place in the renova...
29/04/2026
(L-R) Derek Drescher, Coss Marte, and Syretta Wright have each other's backs. (Micheal Hurcomb/Shutterstock for Sundance Film Festival)
By Veronika Lee Cla...
29/04/2026
Combines EQ and harmonic distortion
Techivation's latest release is a simple EQ designed to offer quick control over a source's overall tonal balanc...
29/04/2026
Two new MPE controllers announced
Expressive E caused quite a stir when they released the Osmose, making the sort of expression that was once reserved for p...
29/04/2026
New modules & enhanced machine-learning
The latest version of iZotope's flagship restoration suite is now available, and now offers over 50 tools design...
29/04/2026
Surgeon Dr Jasmina Kevric wins 2026 Les Murray Award
29 April, 2026
Media releases
Australia for UNHCR and SBS are proud to announce that Dr Jasmina Kevric...
29/04/2026
Some people stumble into their passion. Julissa Padilla walked straight into a film vault. For her, entertainment was never just about the movies themselves. It...
29/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
29/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
29/04/2026
Clear-Com has appointed Brian Grahn as Market Outreach Manager of the Americas and Ben Turnwell as Business Development Manager for EMEA live, expanding their ...
29/04/2026
nxtedition is bringing its range of consolidated production tools to MPTS 2026, with new developments spanning transcription, editing, graphics and AI-assisted ...
29/04/2026
Quortex Switch to boost the streaming experience for Telxius customers, reaching millions of viewers worldwide
Synamedia and Telxius, the leading global connec...
29/04/2026
freispace, the leading ERP-as-a-Service platform for media and entertainment production, and Projective, a leading provider of post-production collaboration tec...
29/04/2026
DHD reports strong interest in its broadcast audio product range, exhibited at the April 19th-22nd NAB Show in Las Vegas. The event attracted a claimed 58,000 a...
29/04/2026
Student Spotlight: Alan Catz The Argentine film and game composer talks about working on League of Legends, receiving Berklee's BMI Award, and the lifelon...
29/04/2026
Jay Jennings Builds the Worlds You Hear on Screen The supervising sound designer behind A Minecraft Movie, The Meg, Letters from Iwo Jima, and dozens of other...
29/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
29/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
29/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
29/04/2026
Voting opens at 4pm today
RT 's Today show have announced the eight finalists for their TV Home Cook competition. Amateur cooks from Cork, Dublin, Galway ...
29/04/2026
29 Apr 2026
VEON and Kyivstar Fulfill Commitment to Invest USD 1 Billion in Ukr...
29/04/2026
Rhod Gilbert, Harriet Kemsley, Kae Kurd, Sara Pascoe and Vicki Pattison to take part in brand new series on free streaming service U
London, 29th April 2026: F...
29/04/2026
Wednesday 29 April 2026
Katie Price: Nothing to Hide, a Sky Original documentar...
29/04/2026
Re-examining the case of Ellie Williams and the wider story of grooming in the town of BarrowWednesday 29 April 2026
Sky announces upcoming documentary series ...
29/04/2026
Wednesday 29 April 2026
Jennifer Garner to lead an all-star cast in new Sky Exc...
29/04/2026
Back to All News
SUPERNOVA: GENESIS Reached a Peak Audience of More Than 6.5 Mi...
29/04/2026
Students and staff from Hills Road Sixth Form College in Cambridge ran a 4.5km course around the roads of Cambridge as part of their annual programme of sustain...
29/04/2026
The Dawn Chorus airs Sunday 3 May from midnight to 7am on RT Radio 1 and RT ly...
29/04/2026
Jin-Quan Yu elected to the National Academy of Sciences Yu is recognized for his pioneering work in synthetic organic chemistry.
April 28, 2026
LA JOLLA, CA S...
28/04/2026
The audio team for the entertainment event must blend speech intelligibility with full-range music reproduction while considering the broadcast
Last week's...