
Mistral AI and NVIDIA today released a new state-of-the-art language model, Mistral NeMo 12B, that developers can easily customize and deploy for enterprise applications supporting chatbots, multilingual tasks, coding and summarization.
By combining Mistral AI's expertise in training data with NVIDIA's optimized hardware and software ecosystem, the Mistral NeMo model offers high performance for diverse applications.
We are fortunate to collaborate with the NVIDIA team, leveraging their top-tier hardware and software, said Guillaume Lample, cofounder and chief scientist of Mistral AI. Together, we have developed a model with unprecedented accuracy, flexibility, high-efficiency and enterprise-grade support and security thanks to NVIDIA AI Enterprise deployment.
Mistral NeMo was trained on the NVIDIA DGX Cloud AI platform, which offers dedicated, scalable access to the latest NVIDIA architecture.
NVIDIA TensorRT-LLM for accelerated inference performance on large language models and the NVIDIA NeMo development platform for building custom generative AI models were also used to advance and optimize the process.
This collaboration underscores NVIDIA's commitment to supporting the model-builder ecosystem.
Delivering Unprecedented Accuracy, Flexibility and Efficiency
Excelling in multi-turn conversations, math, common sense reasoning, world knowledge and coding, this enterprise-grade AI model delivers precise, reliable performance across diverse tasks.
With a 128K context length, Mistral NeMo processes extensive and complex information more coherently and accurately, ensuring contextually relevant outputs.
Released under the Apache 2.0 license, which fosters innovation and supports the broader AI community, Mistral NeMo is a 12-billion-parameter model. Additionally, the model uses the FP8 data format for model inference, which reduces memory size and speeds deployment without any degradation to accuracy.
That means the model learns tasks better and handles diverse scenarios more effectively, making it ideal for enterprise use cases.
Mistral NeMo comes packaged as an NVIDIA NIM inference microservice, offering performance-optimized inference with NVIDIA TensorRT-LLM engines.
This containerized format allows for easy deployment anywhere, providing enhanced flexibility for various applications.
As a result, models can be deployed anywhere in minutes, rather than several days.
NIM features enterprise-grade software that's part of NVIDIA AI Enterprise, with dedicated feature branches, rigorous validation processes, and enterprise-grade security and support.
It includes comprehensive support, direct access to an NVIDIA AI expert and defined service-level agreements, delivering reliable and consistent performance.
The open model license allows enterprises to integrate Mistral NeMo into commercial applications seamlessly.
Designed to fit on the memory of a single NVIDIA L40S, NVIDIA GeForce RTX 4090 or NVIDIA RTX 4500 GPU, the Mistral NeMo NIM offers high efficiency, low compute cost, and enhanced security and privacy.
Advanced Model Development and Customization
The combined expertise of Mistral AI and NVIDIA engineers has optimized training and inference for Mistral NeMo.
Trained with Mistral AI's expertise, especially on multilinguality, code and multi-turn content, the model benefits from accelerated training on NVIDIA's full stack.
It's designed for optimal performance, utilizing efficient model parallelism techniques, scalability and mixed precision with Megatron-LM.
The model was trained using Megatron-LM, part of NVIDIA NeMo, with 3,072 H100 80GB Tensor Core GPUs on DGX Cloud, composed of NVIDIA AI architecture, including accelerated computing, network fabric and software to increase training efficiency.
Availability and Deployment
With the flexibility to run anywhere - cloud, data center or RTX workstation - Mistral NeMo is ready to revolutionize AI applications across various platforms.
Experience Mistral NeMo as an NVIDIA NIM today via ai.nvidia.com, with a downloadable NIM coming soon.
See notice regarding software product information.
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
04/08/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
04/07/2026
April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...
10/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Kiswe announces an expanded long-term partnership with ONE Championship (ONE), t...
09/06/2026
SiriusXM will broadcast FOX Sports' English-language commentary for all 104 FIFA World Cup 2026 matches from June 11 through July 19, available to subscribe...
09/06/2026
EVS has announced it is changing its corporate name from EVS Broadcast Equipment to EVS, reflecting the company's expanded portfolio beyond broadcast equipm...
09/06/2026
Fox Corporation and the NFL have announced a multi-year agreement to bring NFL c...
09/06/2026
FOX Sports and ReachTV, an airport media network, have announced an agreement to...
09/06/2026
Cosm Atlanta, a 70,000-square-foot, three-level immersive entertainment venue located within Centennial Yards adjacent to State Farm Arena and Mercedes-Benz Sta...
09/06/2026
NESN captured five awards at the 2026 Boston/New England Emmy Awards, including four program honors and one individual award.
These awards reflect the passion...
09/06/2026
Ateme has announced that its support for Apple Immersive Video workflows was referenced by Apple during its 2026 Worldwide Developers Conference (WWDC26).
Atem...
09/06/2026
Bitmovin has announced that Axel Springer SE has deployed Player Web X, Bitmovin's web video player, to power audio readouts of news articles and an audio-o...
09/06/2026
Grass Valley and Lawo have announced a technology collaboration to validate orch...
09/06/2026
LiveU has announced a deployment with the Oceania Football Confederation (OFC) t...
09/06/2026
Globecast has announced the launch of its Content Exchange platform, powered by ...
09/06/2026
NBC Sports will present live coverage of Overtime's OT7 football league Championship Weekend from Sullivan Field at Loyola Marymount University in Los Angel...
09/06/2026
We're in three different locations, three different production teams. Coveri...
09/06/2026
The intro video for Men's Basketball won Outstanding In-Venue Video in the C...
09/06/2026
Today is match day minus two for FIFA and HBS. On Thursday, there will be two ma...
09/06/2026
Since their debut, the clock-and-score graphic has drawn the affinity and ire of...
09/06/2026
Last month, Spotify hosted PURE FLOWERS LIVE, a special event celebrating the re...
09/06/2026
Last month, RADAR U.K. artist Skye Newman took the stage in East London for a sp...
09/06/2026
Company to cease operating on 30 June 2026
Australian loudspeaker and amplifier manufacturer Wayne Jones Audio have announced that after much consideration,...
09/06/2026
Increases low-end weight and character
The latest plug-in release from Sheffield-based fedDSP aims to offer an all-in-one solution for users in search of mo...
09/06/2026
Introduces AI Studio Assistant, Moises Studio integration & more
Fender Studio have just announced the launch a significant update that brings an array of n...
09/06/2026
The National Film and Video Foundation (NFVF) is a statutory body mandated to spearhead the equitable growth and development of the South African film and video...
09/06/2026
The BBC has found its Hercule Poirot.
After Deadline revealed last month that t...
09/06/2026
Media buyers and sellers can now compare YouTube reach from computer, mobile, an...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
A project aimed at modernizing scheduling operations across Mediaset's channel portfolio
Mediagenix, a global leader in smart content solutions to profitab...
09/06/2026
tvONE, an ACT Entertainment brand, debuts its most powerful video processor ever built: the 4RU CALICO PRO (C7-PRO-4200), at InfoComm 2026 (Booth N6813). Engine...
09/06/2026
Composer and Conductor Eric Whitacre to Receive Honorary Doctorate at Berklee Va...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Technology collaboration to validate AMPP and Lawo HOME orchestration integration, showcasing Dynamic Media Facility principles in practice.
Grass Valley and ...
09/06/2026
Bitmovin, a leading provider of video streaming solutions, today announced that Axel Springer SE, international media and technology company, has deployed Playe...
09/06/2026
LiveU, the global leader in live IP-video solutions, today announced a landmark deployment with the Oceania Football Confederation (OFC) that has brought Video ...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
09/06/2026
Two Alumni Win Tony Awards Mike Morris and Cedric Leiba Jr. won awards for Best Orchestrations and Best Play, respectively.
June 8, 2026
By
Tori Donahue
...
09/06/2026
NVIDIA GPUs with Confidential Computing are now used for confidential inference in Apple's Private Cloud Compute (PCC), as it expands beyond Apple's dat...
09/06/2026
X-Rite Pantone Launches Offset360 to Modernize Color Control Across Existing Pre...
09/06/2026
09 Jun 2026
VEON Appoints Serkan Ozturk as Chief of Staff & Strategy Officer Dubai and New York, June 9, 2026 - VEON Ltd. (NASDAQ: VEON), a global digital oper...
09/06/2026
Tuesday 9 June 2026
Katie Price: Nothing to Hide, a candid and unfiltered accou...