
NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM).
The open model combined with NVIDIA accelerated computing equips developers, researchers and businesses to innovate responsibly across a wide variety of applications.
Trained on NVIDIA AI Meta engineers trained Llama 3 on computer clusters packing 24,576 NVIDIA H100 Tensor Core GPUs, linked with RoCE and NVIDIA Quantum-2 InfiniBand networks.
To further advance the state of the art in generative AI, Meta recently described plans to scale its infrastructure to 350,000 H100 GPUs.
Putting Llama 3 to Work Versions of Llama 3, accelerated on NVIDIA GPUs, are available today for use in the cloud, data center, edge and PC.
From a browser, developers can try Llama 3 at ai.nvidia.com. It's packaged as an NVIDIA NIM microservice with a standard application programming interface that can be deployed anywhere.
Businesses can fine-tune Llama 3 with their data using NVIDIA NeMo, an open-source framework for LLMs that's part of the secure, supported NVIDIA AI Enterprise platform. Custom models can be optimized for inference with NVIDIA TensorRT-LLM and deployed with NVIDIA Triton Inference Server.
Taking Llama 3 to Devices and PCs Llama 3 also runs on NVIDIA Jetson Orin for robotics and edge computing devices, creating interactive agents like those in the Jetson AI Lab.
What's more, NVIDIA RTX and GeForce RTX GPUs for workstations and PCs speed inference on Llama 3. These systems give developers a target of more than 100 million NVIDIA-accelerated systems worldwide.
Get Optimal Performance with Llama 3 Best practices in deploying an LLM for a chatbot involves a balance of low latency, good reading speed and optimal GPU use to reduce costs.
Such a service needs to deliver tokens - the rough equivalent of words to an LLM - at about twice a user's reading speed which is about 10 tokens/second.
Applying these metrics, a single NVIDIA H200 Tensor Core GPU generated about 3,000 tokens/second - enough to serve about 300 simultaneous users - in an initial test using the version of Llama 3 with 70 billion parameters.
That means a single NVIDIA HGX server with eight H200 GPUs could deliver 24,000 tokens/second, further optimizing costs by supporting more than 2,400 users at the same time.
For edge devices, the version of Llama 3 with eight billion parameters generated up to 40 tokens/second on Jetson AGX Orin and 15 tokens/second on Jetson Orin Nano.
Advancing Community Models An active open-source contributor, NVIDIA is committed to optimizing community software that helps users address their toughest challenges. Open-source models also promote AI transparency and let users broadly share work on AI safety and resilience.
Learn more about how NVIDIA's AI inference platform, including how NIM, TensorRT-LLM and Triton use state-of-the-art techniques such as low-rank adaptation to accelerate the latest LLMs.
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
04/08/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
04/07/2026
April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...
03/06/2026
SES and Viva, Mexico's ultra low-cost airline, have launched multi-orbit satellite inflight connectivity on Viva's Airbus aircraft. A total of 60 A320s ...
03/06/2026
The College Football Playoff, ESPN, and TNT Sports have announced kick times and...
03/06/2026
RED Digital Cinema will exhibit at Cine Gear Expo 2026 (Booth 33, June 5-6, Universal Studios Lot), hosting three panels and hands-on product demonstrations.
P...
03/06/2026
Roku has announced the Soccer Zone, a dedicated hub for FIFA World Cup 2026 content available across the United States, Canada, Mexico, Brazil, Colombia, Argent...
03/06/2026
Liverpool FC and Wasabi Technologies have announced a multi-year extension of their partnership, with Wasabi continuing as the club's official cloud storage...
03/06/2026
Grass Valley has announced that Australian News Channel (ANC), operator of Sky News Australia, has deployed Grass Valley AMPP as part of the relocation of its n...
03/06/2026
The Riedel Group has announced the appointment of Gudrun Scharler as CEO of Riedel Networks. She succeeds Michael Martens, who has led Riedel Networks since 201...
03/06/2026
Clear-Com has announced the deployment of its intercom solutions at Itaka Arena ...
03/06/2026
Scott Coker has announced six senior executive appointments for his new global MMA promotion, which launched earlier this year with $60 million in financing. Ad...
03/06/2026
Telemundo, the exclusive Spanish-language home of the FIFA World Cup 2026, has announced its digital and social media programming for the tournament, running Ju...
03/06/2026
The Bundesliga has announced the launch of Captain, an AI assistant built into t...
03/06/2026
Anthony James Partners (AJP) is serving as technology consultant for a moderniza...
03/06/2026
Audio-Technica has announced the recipients of its annual sales rep firm awards for the 2025-26 fiscal year. The awards were presented by Jim Schanz, Executive ...
03/06/2026
iodyne will exhibit at Cine Gear Expo 2026 alongside RED Digital Cinema and Adob...
03/06/2026
The broadcaster will deploy 39 cameras for game coverage, six cameras for studio...
03/06/2026
The 2026 FIFA World Cup is providing a great opportunity for not only Fox Sports...
03/06/2026
Sony Electronics is introducing the SRG-AS10, a 4K 60p compatible PTZ Auto Framing camera that uses Sony's proprietary AI to automatically recognize and tra...
03/06/2026
Game Creek Video's Flagship A, B, C, and D unit will wind up its first year ...
03/06/2026
According to Caretta Research, the sports-rights market may be hitting the brake...
03/06/2026
Retrieval still courtesy of Tribeca.
By Jessica Herndon
This year, the Tribeca...
03/06/2026
Step sequencer-style panning tool revealed
Alongside their flagship self-titled sound-design platform, Sound Particles offer an array of creative effects an...
03/06/2026
Now features full H90 algorithm library
Eventide have announced the upcoming launch of the H9 Harmonizer Gen 2, a new and improved version of their hugely p...
03/06/2026
Significant discount available until 1 October 2026
Aim Audio have just announced a promotion that sees a significant discount applied to their Essence micr...
03/06/2026
Clarification: SBS's position on definitions of antisemitism
3 June, 2026
Media releases
Statement by Mandi Wicks, SBS Director of News and Current Aff...
03/06/2026
Rohde & Schwarz to supply CERTIUM advanced communications system to Memmingen Ai...
03/06/2026
eds3_5_jq(document).ready(function($) { $(#eds_sliderM519).chameleonSlider_2_1({...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
For more than three decades, Re-recording Mixer Andrew Wilson, AMPS, CAS, has helped bring the natural world to the screen with exceptional audio enjoyed by mil...
03/06/2026
Telestream, a global leader in media workflow technologies, will showcase its latest innovations for modern AV production environments at InfoComm 2026 (Booth N...
03/06/2026
DPA Microphones will present a comprehensive portfolio of integrated audio solutions designed to meet the evolving needs of today's professional AV environm...
03/06/2026
Lightware announces the GVN-HC-TX220AP, a new transmitter in the Gemini GVN 1G AV-over-IP family that introduces full-featured USB-C for professional 1Gb AV-ove...
03/06/2026
Evergent, the customer management and monetization leader for streaming and digital subscription businesses, and Minno, the global leader in faith-based content...
03/06/2026
Alfalite, Europe's only LED display manufacturer, has completed a new broadcast installation with the deployment of two UHD Finepix 1.5 MATIX AlfaCOB LED di...
03/06/2026
Broadcast Solutions, a leading systems integrator and provider of innovative solutions for the broadcast media industry, is completing a contract to build eight...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
03/06/2026
Creamsource, known for its tried-and-true Vortex Series of cinematic lighting, has announced the Vortex2 (V2) and Vortex2 Soft (V2S), two compact additions to t...
03/06/2026
Do not read this media release! Andy Lee's smash-hit show returns to ABC thi...
03/06/2026
Faster, More Flexible AI Matting Fuels Boris FX Silhouette
Jessie Electa Petrov June 2, 2026
0 Comments
The 2026 release helps artists tackle complex ...
03/06/2026
La T l 's National League Hockey Expansion Powered by Blackmagic Design
Brie Clayton June 2, 2026
0 Comments
Blackmagic Videohub 120 120 12G provi...