
NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM).
The open model combined with NVIDIA accelerated computing equips developers, researchers and businesses to innovate responsibly across a wide variety of applications.
Trained on NVIDIA AI Meta engineers trained Llama 3 on computer clusters packing 24,576 NVIDIA H100 Tensor Core GPUs, linked with RoCE and NVIDIA Quantum-2 InfiniBand networks.
To further advance the state of the art in generative AI, Meta recently described plans to scale its infrastructure to 350,000 H100 GPUs.
Putting Llama 3 to Work Versions of Llama 3, accelerated on NVIDIA GPUs, are available today for use in the cloud, data center, edge and PC.
From a browser, developers can try Llama 3 at ai.nvidia.com. It's packaged as an NVIDIA NIM microservice with a standard application programming interface that can be deployed anywhere.
Businesses can fine-tune Llama 3 with their data using NVIDIA NeMo, an open-source framework for LLMs that's part of the secure, supported NVIDIA AI Enterprise platform. Custom models can be optimized for inference with NVIDIA TensorRT-LLM and deployed with NVIDIA Triton Inference Server.
Taking Llama 3 to Devices and PCs Llama 3 also runs on NVIDIA Jetson Orin for robotics and edge computing devices, creating interactive agents like those in the Jetson AI Lab.
What's more, NVIDIA RTX and GeForce RTX GPUs for workstations and PCs speed inference on Llama 3. These systems give developers a target of more than 100 million NVIDIA-accelerated systems worldwide.
Get Optimal Performance with Llama 3 Best practices in deploying an LLM for a chatbot involves a balance of low latency, good reading speed and optimal GPU use to reduce costs.
Such a service needs to deliver tokens - the rough equivalent of words to an LLM - at about twice a user's reading speed which is about 10 tokens/second.
Applying these metrics, a single NVIDIA H200 Tensor Core GPU generated about 3,000 tokens/second - enough to serve about 300 simultaneous users - in an initial test using the version of Llama 3 with 70 billion parameters.
That means a single NVIDIA HGX server with eight H200 GPUs could deliver 24,000 tokens/second, further optimizing costs by supporting more than 2,400 users at the same time.
For edge devices, the version of Llama 3 with eight billion parameters generated up to 40 tokens/second on Jetson AGX Orin and 15 tokens/second on Jetson Orin Nano.
Advancing Community Models An active open-source contributor, NVIDIA is committed to optimizing community software that helps users address their toughest challenges. Open-source models also promote AI transparency and let users broadly share work on AI safety and resilience.
Learn more about how NVIDIA's AI inference platform, including how NIM, TensorRT-LLM and Triton use state-of-the-art techniques such as low-rank adaptation to accelerate the latest LLMs.
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
04/08/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
04/07/2026
April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...
01/06/2026
January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026
Throughout the week, Dolby brings to life the latest innovatio...
02/05/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
01/05/2026
January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Evergent introduces its Agentic Revenue Orchestration Platform, transforming how subscription businesses across direct-to-consumer streaming, pay-TV, telecommun...
15/04/2026
Harmonic's XOS Media Processor Delivers Exceptional Video Quality to More than Half of U.S. Public Media Viewership
Harmonic (NASDAQ: HLIT) today announce...
15/04/2026
LONGMONT, COLORADO, APRIL 15, 2026 DPA Microphones N Series Digital Wireless System users in North America can now take full advantage of the system's exc...
15/04/2026
Cobalt Iron, a leading provider of SaaS-based enterprise data protection, today announced the launch of Compass Tape Gateway (CTG), a transformative enhancemen...
15/04/2026
Disguise to Showcase Cutting-Edge Experience Tech for Sports, Broadcast and More...
15/04/2026
Arooj Aftab Makes the Music She Wants to Hear The singular artist explores the juxtaposition of grief and joy, dark and light, in her distinctive sound.
Apri...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Interra Systems, a provider of end-to-end quality assurance solutions for the digital media industry, is proud to announce its central role in the digital trans...
15/04/2026
VO Integrates Client-Side and Server-Side Ad Technologies with its Secure Video Player and Segmentation Tools, Leveraging Broadpeak's SSAI and BPK SmartLib ...
15/04/2026
TMT Insights, a leader in professional services and software development for the media and entertainment (M&E) industry, today announced the launch of TMT AI at...
15/04/2026
Encompass Digital Media, a global leader in managed broadcast and cloud-based media services, today announced an expanded partnership with Oracle Cloud Infrastr...
15/04/2026
Middleman Software, a leading provider of signaling and synchronization solutions for broadcast and streaming workflows, today announced it will unveil PhaseLoc...
15/04/2026
Deity Microphones today announced the PR-4, a compact six-track field recorder featuring four inputs, 32-bit float recording, advanced routing, and a workflow-f...
15/04/2026
Blackmagic Design Announces Fairlight Live
Brie Clayton April 15, 2026
0 Comments
Powerful software-based live audio mixer with support for SMPTE-2110...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
15/04/2026
Strategic Alignment With the Riedel Group for Innovation and Growth
Thomas Riedel, founder and owner of Riedel Communications and the Riedel Group, has acqui...
15/04/2026
SDVI, the leading platform provider for cloud-native media supply chains, today announced that the company's Rally platform has been deployed by Iyuno Media...
15/04/2026
Frequency, the engine behind many of the worlds leading streaming television channels, has become the platform of choice for leading U.S. broadcasters bringing ...
15/04/2026
New Service Simplifies Integration of AI Applications with Exceptional Reliability and Security
SAN JOSE, Calif. April 14, 2026 Harmonic (NASDAQ: HLIT) t...
15/04/2026
During the 2025 off-season, VIDI completed a major upgrade of the broadcast infrastructure across all 36 stadiums in Germany's 1st and 2nd Bundesliga. The p...
15/04/2026
Clear-Com announced significant updates to its award-winning Arcadia Central Station as well as enhancements to the Eclipse HX Digital Matrix System now runn...
15/04/2026
Shotoku Introduces the World to Aura P2 PTZ Prompter Panner at NAB SHOW 2026
Brie Clayton April 15, 2026
0 Comments
New system removes PTZ pan restric...
15/04/2026
LiveU Announces Expanded Collaboration with Sony at NAB Show, Adding Direct File...
15/04/2026
This 24-Hour Performance Art Piece Needs Your Participation Keefer Glenshaws latest project is about collaborating with the Berklee community and beyond.
Apr...
15/04/2026
Traditional data centers only stored, retrieved and processed data. In the generative and agentic AI era, these facilities have evolved into AI token factories....
15/04/2026
Two Cities Television and Keeper Pictures to produce six-part series in association with Screen Ireland for RT and ITV Studios
RT has commissioned Two Citie...
15/04/2026
RT 100
RT unveils The Signal' a specially commissioned 60-second film
...
15/04/2026
The NAB Show 2026 trade show, running April 18-22 in Las Vegas, is set to showcase a wave of new features and optimizations for top video editing applications. ...
14/04/2026
ToolsOnAir and Omnistream Bring Mobile SRT Contribution into Professional 24/7 P...
14/04/2026
Haivision has announced the Falkon X4, a 5G mobile video transmitter for live broadcast and remote production. The device will be showcased at NAB Show 2026.
...
14/04/2026
Founded by veterans of the production-truck world, Stripe TV was born out of a d...
14/04/2026
Grass Valley has expanded its partnership with Studio Berlin, supplying 12 LDX 135 UHD/HDR camera systems and 12 LDX 180 Super 35mm cinematic cameras, including...
14/04/2026
LTN has announced two senior appointments ahead of NAB Show 2026: Mark Romano as Vice President, Multichannel Platforms, and Edward Cox as Vice President, Sales...
14/04/2026
Sennheiser and Italian rental company Agor have begun preparations for their Eu...
14/04/2026
DAZN and ADI Predictstreet have announced a partnership to integrate ADI Predict...
14/04/2026
Deltatre has announced the completion of a direct-to-fan website and app for Leg...