Sony Pixel Power calrec Sony

How Scaling Laws Drive Smarter, More Powerful AI

12/02/2025

Just as there are widely understood empirical laws of nature - for example, what goes up must come down, or every action has an equal and opposite reaction - the field of AI was long defined by a single idea: that more compute, more training data and more parameters makes a better AI model.

However, AI has since grown to need three distinct laws that describe how applying compute resources in different ways impacts model performance. Together, these AI scaling laws - pretraining scaling, post-training scaling and test-time scaling, also called long thinking - reflect how the field has evolved with techniques to use additional compute in a wide variety of increasingly complex AI use cases.

The recent rise of test-time scaling - applying more compute at inference time to improve accuracy - has enabled AI reasoning models, a new class of large language models (LLMs) that perform multiple inference passes to work through complex problems, while describing the steps required to solve a task. Test-time scaling requires intensive amounts of computational resources to support AI reasoning, which will drive further demand for accelerated computing.

What Is Pretraining Scaling? Pretraining scaling is the original law of AI development. It demonstrated that by increasing training dataset size, model parameter count and computational resources, developers could expect predictable improvements in model intelligence and accuracy.

Each of these three elements - data, model size, compute - is interrelated. Per the pretraining scaling law, outlined in this research paper, when larger models are fed with more data, the overall performance of the models improves. To make this feasible, developers must scale up their compute - creating the need for powerful accelerated computing resources to run those larger training workloads.

This principle of pretraining scaling led to large models that achieved groundbreaking capabilities. It also spurred major innovations in model architecture, including the rise of billion- and trillion-parameter transformer models, mixture of experts models and new distributed training techniques - all demanding significant compute.

And the relevance of the pretraining scaling law continues - as humans continue to produce growing amounts of multimodal data, this trove of text, images, audio, video and sensor information will be used to train powerful future AI models.

Pretraining scaling is the foundational principle of AI development, linking the size of models, datasets and compute to AI gains. Mixture of experts, depicted above, is a popular model architecture for AI training. What Is Post-Training Scaling? Pretraining a large foundation model isn't for everyone - it takes significant investment, skilled experts and datasets. But once an organization pretrains and releases a model, they lower the barrier to AI adoption by enabling others to use their pretrained model as a foundation to adapt for their own applications.

This post-training process drives additional cumulative demand for accelerated computing across enterprises and the broader developer community. Popular open-source models can have hundreds or thousands of derivative models, trained across numerous domains.

Developing this ecosystem of derivative models for a variety of use cases could take around 30x more compute than pretraining the original foundation model.

Developing this ecosystem of derivative models for a variety of use cases could take around 30x more compute than pretraining the original foundation model.

Post-training techniques can further improve a model's specificity and relevance for an organization's desired use case. While pretraining is like sending an AI model to school to learn foundational skills, post-training enhances the model with skills applicable to its intended job. An LLM, for example, could be post-trained to tackle a task like sentiment analysis or translation - or understand the jargon of a specific domain, like healthcare or law.

The post-training scaling law posits that a pretrained model's performance can further improve - in computational efficiency, accuracy or domain specificity - using techniques including fine-tuning, pruning, quantization, distillation, reinforcement learning and synthetic data augmentation.

Fine-tuning uses additional training data to tailor an AI model for specific domains and applications. This can be done using an organization's internal datasets, or with pairs of sample model input and outputs.

Distillation requires a pair of AI models: a large, complex teacher model and a lightweight student model. In the most common distillation technique, called offline distillation, the student model learns to mimic the outputs of a pretrained teacher model.

Reinforcement learning, or RL, is a machine learning technique that uses a reward model to train an agent to make decisions that align with a specific use case. The agent aims to make decisions that maximize cumulative rewards over time as it interacts with an environment - for example, a chatbot LLM that is positively reinforced by thumbs up reactions from users. This technique is known as reinforcement learning from human feedback (RLHF). Another, newer technique, reinforcement learning from AI feedback (RLAIF), instead uses feedback from AI models to guide the learning process, streamlining post-training efforts.

Best-of-n sampling generates multiple outputs from a language model and selects the one with the highest reward score based on a reward model. It's often used to improve an AI's outputs without modifying model parameters, offering an alternative to fine-tuning with reinforcement learning.

Search methods explore a range of potential decision paths before selecting a final output. This post-training technique can iteratively improve the model's responses
LINK: https://blogs.nvidia.com/blog/ai-scaling-laws/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

04/08/2026

Dalet Announces Commercial Availability of Dalia, Bringing Media-Aware Agentic AI to Enterprise Productions

Dalet, a leading technology and service provider for media-rich organizations, t...

04/07/2026

Detective Conan: Fallen Angel of the Highway Opens in Dolby Cinemas Across Japan, Presented in Dolby Atmos and Dolby ...

April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...

08/06/2026

SVG Sit-Down: FIFA's Oscar Sanchez and HBS's Paul King on Deploying 16 Production Teams in 16 World Cup Venues

The two execs discuss the goals and the challenges of building a huge multi-cult...

08/06/2026

SVG All-Stars: Judi Weiss, Senior Manager, Remote Operations (Studio), ESPN

This industry legend has been a road warrior since the early '90s and is, today, a lynchpin on College GameDay The sports-production industry continues to ...

08/06/2026

DAZN Extends Combat Sports Portfolio with EFC Partnership

DAZN has agreed a multi-year rights partnership with Extreme Fighting Championship (EFC), expanding its global combat sports portfolio. DAZN will broadcast EFC...

08/06/2026

Whitepaper: TVU Networks Tackles the Future of Sports Production

As fans move across broadcast, streaming, social, mobile and creator-led platforms, and they expect every angle, in real time, TVU Networks has released a new e...

08/06/2026

Infront Strengthens Diamond League Global Broadcast Network Through Spate of New Agreements

Infront has further strengthened the 2026 Diamond League's international bro...

08/06/2026

Grass Valley Appoints Sam Craig as Vice President, Global Pre-Sales

Grass Valley has announced the appointment of Sam Craig as Vice President, Global Pre-Sales. Based in the United States, Craig will lead Grass Valley's glob...

08/06/2026

SMPTE Announces Summer 2026 IP Networking and ST 2110 Course Lineup

SMPTE has announced its summer 2026 education lineup covering IP networking and SMPTE ST 2110, including two standalone courses, a 13-week boot camp, and a hand...

08/06/2026

Sportradar and Kalshi Announce Multi-Year Data and Infrastructure Partnership for Prediction Markets

Sportradar Group AG has announced a multi-year global agreement with Kalshi, a p...

08/06/2026

FIFA and Gamefam Launch FIFA World Cup 2026 Event on Roblox

FIFA and Gamefam have announced the launch of a FIFA World Cup 2026 event on Roblox, running June 5 through July 31, 2026. The event centers on FIFA Super Socce...

08/06/2026

Stanley Pup 2026: Fun Is the Goal When the Rescue-Dog Competition Returns for Year 3

A collaboration of NHL Productions, Monumental Sports, and Michael Levitt Produc...

08/06/2026

Bryce Adair, CBS Production Assistant, Dies in Car Crash at 31

CBS Sports production staffer Bryce Adair has died at the age of 31 after sustaining injuries during a car accident on Wednesday while in Ohio working the the P...

08/06/2026

Behind the Mic: Russell Wilson and Kyle Long Join CBS Sports The NFL Today

Behind The Mic provides a roundup of recent news regarding on-air talent, including new deals, departures, and assignments compiled from press releases and repo...

08/06/2026

WMT Digital Launches Sports Customer Data Platform Built to Turn Fan Intelligence into Revenue

WMT Digital today announced the launch of WMT Fan Intelligence, designed to unif...

08/06/2026

Miami Heat Ink New Deal With WPLG Local 10, Join NBAs Growing OTA Crowd

The Heat are the seventh NBA team to air their games over-the-air next season...

08/06/2026

Paramount and UFC Expand Media Rights Partnership to Canada Beginning in 2027

Paramount and UFC have announced an expansion of their media rights partnership making Paramount the exclusive home of UFC Numbered Event main cards in Canada ...

08/06/2026

Qatar Football Association Deploys Virtual Advertising Technology for First Time with Sponix

The Qatar Football Association (QFA), in cooperation with Sponix and broadcaster...

08/06/2026

Project B Appoints HBS as Official Host Broadcaster for Global Basketball Circuit

Project B has announced the appointment of Host Broadcast Services (HBS) as Offi...

08/06/2026

Ohio State's Marco Fragale on the Growth of On-Campus B1G+ Productions and StudentU

The Buckeyes now produce and stream upwards of 200 sports events annually with a...

08/06/2026

Meet the New GLOW Ambassadors Leading Spotify's Global Pride Celebration

At Spotify, our commitment to the LGBTQIA community is year-round. Through GLOW, our global music program, we celebrate and amplify the contributions of queer ...

08/06/2026

GearExpo UK: Keyboard & Synth Update

Get Hands-on With Keyboard & Synth Brands GearExpo UK wouldn't be complete without some synth action, and we've got some of the industry's most ...

08/06/2026

IK Multimedia's ARC On-Ear gains IEM support

50 popular in-ear monitoring system profiles added The latest update for IK Multimedia's headphone-correction system has just arrived, and introduces ca...

08/06/2026

Audeze announce the MM-520

Manny Marroquin signature cans upgraded with SLAM Technology The flagship model in Audeze's Manny Marroquin Signature Series has just been treated to an...

08/06/2026

Air domain supremacy redefined - New counter UAS, space solutions and directional communications from Rohde & Schwarz debut at ILA

Air domain supremacy redefined - New counter UAS, space solutions and directiona...

08/06/2026

BBC announces Hercule, starring Edward Bluemel as Agatha Christie's legendary detective Hercule Poirot

he BBC and BritBox have announced that Edward Bluemel (We Might Regret This, My ...

08/06/2026

Sam Craig Joins Grass Valley as VP, Global Presales

Share Copy link Facebook X Linkedin Bluesky Email...

08/06/2026

SMPTE Sets Course Lineup With ST 2110, IP Networking Focus

Share Copy link Facebook X Linkedin Bluesky Email...

08/06/2026

Glensound puts practical audio connectivity centre stage...

Dante interfaces, intercom and control systems for modern AV environments Glensound will showcase a focused range of networked audio and intercom solutions fo...

08/06/2026

Bminty and T18 celebrates the channel s 1st anniversary -...

Bminty is proud to celebrate the first anniversary of T18, the new free national French DTT channel launched on June 6, 2025. This symbolic milestone also marks...

08/06/2026

DaVinci Resolve Used For Global Collaboration on Tony Foster Documentary

DaVinci Resolve Used For Global Collaboration on Tony Foster Documentary Brie Clayton June 8, 2026 0 Comments DaVinci Resolve Studio and Blackmagic Cl...

08/06/2026

New York Cracks Down on AI Bots

Share Copy link Facebook X Linkedin Bluesky Email...

08/06/2026

Sky introduces Real Time feature on Sky Glass and Sky Stream to bring fans closer to the World Cup action

Plus, 20% off TVs ahead of kick-offMonday 8 June 2026 Sky introduces Real Time...

08/06/2026

FOX Secures Live NFL Game Package in Mexico Starting in Fall 2026

FOX Secures Live NFL Game Package in Mexico Starting in Fall 2026 Agreement Features Thursday Night Football, Sunday Games Package, Thanksgiving Day Games, al...

08/06/2026

FOX One, FOX Sports and Indeed Name Austin Franklin and Kevin Akoto as FOX One Chief World Cup Watchers

FOX One, FOX Sports and Indeed Name Austin Franklin and Kevin Akoto as FOX One C...

08/06/2026

Hamburg Open. Hamburg. 14-15 January 2026

Meet us on the show floor Stand #286310 Discover Nara, the media management tool used by major facilities including Harbor and Molinare. Nara v2 introduces re...

08/06/2026

Micro Salon. Paris. 5-6 February 2026

L'art de l'Image dans Reflet dans un diamant mort' Une conversation avec le directeur de la photographie Manuel Dacosse, SBC et l' talonneur Pe...

08/06/2026

HPA Tech Retreat. Rancho Mirage. 15-19 February 2026

Embracing today's modern media workflows: FilmLight presents Nara 2.0, with FilmLight API Designed to support the growing demands of today's production ...

08/06/2026

Beyond the prompt: Colour grading in the age of AI. Berlin. 18 February 2026

Moderated by Andy Minuth, FilmLight's Colour Workflow Specialist Wednesday 18 February 6:00pm / Doors open 7:00pm / Presentation in German 8:00pm / Drin...

08/06/2026

Modern workflow simplified: FilmLight presents Nara and Daylight with FilmLight API. London. 14 April 2026

Join us on April 14 at 10:00am for a technical roundtable with the Filmlight dev...

08/06/2026

The Fundamentals of Coding and Machine-Assisted Development. London. Various

You're invited to FLAPI Classroom The fundamentals of coding and machine-assisted development These sessions will help you build the skills needed to cre...

08/06/2026

Colour Masterclass at MELS. Montreal. 9 May 2026

With Sylvain Canaux (St Louis, Paris) and J r me Cloutier (MELS, Montreal) Wednesday 6 May Pick your time: 1:00PM / 5:00PM Note: The presentation will be hel...

08/06/2026

Simplify your workflows with FLAPI. Los Angeles. 9 June 2026

FilmLight, 1107 N El Centro Ave, Los Angeles Doors open at 3:30pm Join the FilmLight team on June 9th at 4pm to learn how FilmLight products and APIs can stre...

08/06/2026

The Creators List launched to Help Brands Connect With Top Creators In Cannes

The Creators List launched to Help Brands Connect With Top Creators In CannesThe curated directory launched by Tubefilter, Comscore, Whalar Group and Gospel Sta...

08/06/2026

It's almost kick off time! RT KIDS show Total Football returns for a second season with a brand new co-host

Irish YouTube star DavidMC joins Aisling O'Reilly to tackle all things socce...

07/06/2026

Decksaver's Sping 2026 Drop

Company introduce 21 new protective covers Decksaver have just announced their Sping 2026 Drop, which sees a total of 21 new models added to their ever-grow...

07/06/2026

How to Recreate the Star Trek TNG jump to warp in After Effects

How to Recreate the Star Trek TNG jump to warp in After Effects Graham Quince June 7, 2026 0 Comments In 1987, John Knoll at ILM had to use labor-inte...

07/06/2026

Filming Coffee Communities with Blackmagic PYXIS

Filming Coffee Communities with Blackmagic PYXIS Brie Clayton June 7, 2026 0 Comments Remote travel, fast setups and hard sun shaped this documentary ...

07/06/2026

NVIDIA and Doosan Group Collaborate to Advance Physical AI and AI Factory Infrastructure

NVIDIA and Doosan Group are expanding their collaboration to advance new opportu...