Sony Pixel Power calrec Sony

Exploring the Revenue-Generating Potential of AI Factories

15/05/2025

AI is creating value for everyone - from researchers in drug discovery to quantitative analysts navigating financial market changes.

The faster an AI system can produce tokens, a unit of data used to string together outputs, the greater its impact. That's why AI factories are key, providing the most efficient path from time to first token to time to first value.

AI factories are redefining the economics of modern infrastructure. They produce intelligence by transforming data into valuable outputs - whether tokens, predictions, images, proteins or other forms - at massive scale.

They help enhance three key aspects of the AI journey - data ingestion, model training and high-volume inference. AI factories are being built to generate tokens faster and more accurately, using three critical technology stacks: AI models, accelerated computing infrastructure and enterprise-grade software.

Read on to learn how AI factories are helping enterprises and organizations around the world convert the most valuable digital commodity - data - into revenue potential.

From Inference Economics to Value Creation Before building an AI factory, it's important to understand the economics of inference - how to balance costs, energy efficiency and an increasing demand for AI.

Throughput refers to the volume of tokens that a model can produce. Latency is the amount of tokens that the model can output in a specific amount of time, which is often measured in time to first token - how long it takes before the first output appears - and time per output token, or how fast each additional token comes out. Goodput is a newer metric, measuring how much useful output a system can deliver while hitting key latency targets.

User experience is key for any software application, and the same goes for AI factories. High throughput means smarter AI, and lower latency ensures timely responses. When both of these measures are balanced properly, AI factories can provide engaging user experiences by quickly delivering helpful outputs.

For example, an AI-powered customer service agent that responds in half a second is far more engaging and valuable than one that responds in five seconds, even if both ultimately generate the same number of tokens in the answer.

Companies can take the opportunity to place competitive prices on their inference output, resulting in more revenue potential per token.

Measuring and visualizing this balance can be difficult - which is where the concept of a Pareto frontier comes in.

AI Factory Output: The Value of Efficient Tokens The Pareto frontier, represented in the figure below, helps visualize the most optimal ways to balance trade-offs between competing goals - like faster responses vs. serving more users simultaneously - when deploying AI at scale.

The vertical axis represents throughput efficiency, measured in tokens per second (TPS), for a given amount of energy used. The higher this number, the more requests an AI factory can handle concurrently.

The horizontal axis represents the TPS for a single user, representing how long it takes for a model to give a user the first answer to a prompt. The higher the value, the better the expected user experience. Lower latency and faster response times are generally desirable for interactive applications like chatbots and real-time analysis tools.

The Pareto frontier's maximum value - shown as the top value of the curve - represents the best output for given sets of operating configurations. The goal is to find the optimal balance between throughput and user experience for different AI workloads and applications.

The best AI factories use accelerated computing to increase tokens per watt - optimizing AI performance while dramatically increasing energy efficiency across AI factories and applications.

The animation above compares user experience when running on NVIDIA H100 GPUs configured to run at 32 tokens per second per user, versus NVIDIA B300 GPUs running at 344 tokens per second per user. At the configured user experience, Blackwell Ultra delivers over a 10x better experience and almost 5x higher throughput, enabling up to 50x higher revenue potential.

How an AI Factory Works in Practice An AI factory is a system of components that come together to turn data into intelligence. It doesn't necessarily take the form of a high-end, on-premises data center, but could be an AI-dedicated cloud or hybrid model running on accelerated compute infrastructure. Or it could be a telecom infrastructure that can both optimize the network and perform inference at the edge.

Any dedicated accelerated computing infrastructure paired with software turning data into intelligence through AI is, in practice, an AI factory.

The components include accelerated computing, networking, software, storage, systems, and tools and services.

When a person prompts an AI system, the full stack of the AI factory goes to work. The factory tokenizes the prompt, turning data into small units of meaning - like fragments of images, sounds and words.

Each token is put through a GPU-powered AI model, which performs compute-intensive reasoning on the AI model to generate the best response. Each GPU performs parallel processing - enabled by high-speed networking and interconnects - to crunch data simultaneously.

An AI factory will run this process for different prompts from users across the globe. This is real-time inference, producing intelligence at industrial scale.

Because AI factories unify the full AI lifecycle, this system is continuously improving: inference is logged, edge cases are flagged for retraining and optimization loops tighten over time - all without manual intervention, an example of goodput in action.

Leading global security technology company Lockheed Martin has built its own AI factory to support diverse uses across its business. Through its
LINK: https://blogs.nvidia.com/blog/revenue-potential-ai-factories/...
See more stories from nvidia

More from Nvidia

24/06/2025

NVIDIA and Partners Highlight Next-Generation Robotics, Automation and AI Technologies at Automatica

From the heart of Germany's automotive sector to manufacturing hubs across F...

19/06/2025

Step Inside the Vault: The Borderland' Series Arrives on GeForce NOW

GeForce NOW is throwing open the vault doors to welcome the legendary Borderland series to the cloud. Whether a seasoned Vault Hunter or new to the mayhem of P...

18/06/2025

Plug and Play: Build a G-Assist Plug-In Today

Project G-Assist - available through the NVIDIA App - is an experimental AI assistant that helps tune, control and optimize NVIDIA GeForce RTX systems. NVIDIA&...

17/06/2025

Hexagon Taps NVIDIA Robotics and AI Software to Build and Deploy AEON, a New Humanoid

As a global labor shortage leaves 50 million positions unfilled across industrie...

13/06/2025

NVIDIA and Deutsche Telekom Partner to Advance Germany's Sovereign AI

Industrial AI isn't slowing down. Germany is ready. Following London Tech Week and GTC Paris at VivaTech, NVIDIA founder and CEO Jensen Huang's Europea...

12/06/2025

NVIDIA TensorRT Boosts Stable Diffusion 3.5 Performance on NVIDIA GeForce RTX and RTX PRO GPUs

Generative AI has reshaped how people create, imagine and interact with digital ...

12/06/2025

Turn RTX ON With 40% Off Performance Day Passes

Level up GeForce NOW experiences this summer with 40% off Performance Day Passes. Enjoy 24 hours of premium cloud gaming with RTX ON, delivering low latency and...

11/06/2025

NVIDIA DRIVE Full-Stack Autonomous Vehicle Software Rolls Out

NVIDIA is launching a comprehensive, industry-defining autonomous vehicle (AV) software platform to accelerate large-scale deployment of safe, intelligent trans...

11/06/2025

NVIDIA Research Casts New Light on Scenes With AI-Powered Rendering for Physical AI Development

NVIDIA Research has developed an AI light switch for videos that can turn daytim...

11/06/2025

European Researchers Develop AI-Native Wireless Networks With NVIDIA 6G Research Portfolio

Using NVIDIA platforms, tools and libraries, European telecommunications institu...

11/06/2025

NVIDIA Scores Consecutive Win for End-to-End Autonomous Driving Grand Challenge at CVPR

NVIDIA was today named an Autonomous Grand Challenge winner at the Computer Visi...

11/06/2025

European Robot Makers Adopt NVIDIA Isaac, Omniverse and Halos to Develop Safe, Physical AI-Driven Robot Fleets

In the face of growing labor shortages and need for sustainability, European man...

11/06/2025

Retail Reboot: Major Global Brands Transform End-to-End Operations With NVIDIA

AI is packing and shipping efficiency for the retail and consumer packaged goods (CPG) industries, with a majority of surveyed companies in the space reporting ...

11/06/2025

NVIDIA Brings Physical AI to European Cities With New Blueprint for Smart City AI

Urban populations are expected to double by 2050, which means around 2.5 billion...

11/06/2025

Calling on LLMs: New NVIDIA AI Blueprint Helps Automate Telco Network Configuration

Telecom companies last year spent nearly $295 billion in capital expenditures an...

11/06/2025

European Broadcasting Union and NVIDIA Partner on Sovereign AI to Support Public Broadcasters

In a new effort to advance sovereign AI for European public service media, NVIDI...

11/06/2025

NVIDIA CEO Drops the Blueprint for Europe's AI Boom

At GTC Paris - held alongside VivaTech, Europe's largest tech event - NVIDIA founder and CEO Jensen Huang delivered a clear message: Europe isn't just a...

10/06/2025

The Blue Lion Supercomputer Will Run on NVIDIA Vera Rubin - Here's Why That Matters

Germany's Leibniz Supercomputing Centre, LRZ, is gaining a new supercomputer...

10/06/2025

Clear Skies Ahead: New NVIDIA Earth-2 Generative AI Foundation Model Simulates Global Climate at Kilometer-Scale Resolution

With a more detailed simulation of the Earth's climate, scientists and resea...

10/06/2025

Cisco and NVIDIA Advance Security for Enterprise AI Factories

Cisco and NVIDIA are helping set a new standard for secure, scalable and high-performance enterprise AI. Announced today at the Cisco Live conference in San Di...

09/06/2025

UK Prime Minister, NVIDIA CEO Set the Stage as AI Lights Up Europe

AI isn't waiting. And this week, neither is Europe. At London's Olympia, under a ceiling of steel beams and enveloped by the thrum of startup pitches, ...

08/06/2025

AI Maker, Not an AI Taker': UK Builds Its Vision With NVIDIA Infrastructure

U.K. Prime Minister Keir Starmer's ambition for Britain to be an AI maker, not an AI taker, is becoming a reality at London Tech Week. With NVIDIA's ...

05/06/2025

GeForce NOW Kicks Off a Summer of Gaming With 25 New Titles This June

GeForce NOW is a gamer's ticket to an unforgettable summer of gaming. With 25 titles coming this month and endless ways to play, the summer is going to be e...

04/06/2025

NVIDIA Blackwell Delivers Breakthrough Performance in Latest MLPerf Training Results

NVIDIA is working with companies worldwide to build out AI factories - speeding ...

04/06/2025

How 1X Technologies' Robots Are Learning to Lend a Helping Hand

Humans learn the norms, values and behaviors of society from each other - and Bernt B rnich, founder and CEO of 1X Technologies, thinks robots should learn like...

04/06/2025

NVIDIA RTX Blackwell GPUs Accelerate Professional-Grade Video Editing

4:2:2 cameras - capable of capturing double the color information compared with most standard cameras - are becoming widely available for consumers. At the same...

02/06/2025

Bring Receipts: New NVIDIA AI Blueprint Detects Fraudulent Credit Card Transactions With Precision

Editor's note: This blog, originally published on October 28, 2024, has been...

02/06/2025

Researchers and Students in Trkiye Build AI, Robotics Tools to Boost Disaster Readiness

Since a 7.8-magnitude earthquake hit Syria and T rkiye two years ago - leaving 5...

29/05/2025

The Supercomputer Designed to Accelerate Nobel-Worthy Science

Ready for a front-row seat to the next scientific revolution? That's the idea behind Doudna - a groundbreaking supercomputer announced today at Lawrence Be...

29/05/2025

Run LLMs on AnythingLLM Faster With NVIDIA RTX AI PCs

Large language models (LLMs), trained on datasets with billions of tokens, can generate high-quality content. They're the backbone for many of the most popu...

29/05/2025

RTX on Deck: The GeForce NOW Native App for Steam Deck Is Here

GeForce NOW is supercharging Valve's Steam Deck with a new native app - delivering the high-quality GeForce RTX-powered gameplay members are used to on a po...

28/05/2025

NVIDIA's Bartley Richardson on How Teams of AI Agents Provide Next-Level Automation

Building effective agentic AI systems requires rethinking how technology interac...

27/05/2025

How Dell Technologies Is Building the Engines of AI Factories With NVIDIA Blackwell

Over a century ago, Henry Ford pioneered the mass production of cars and engines...

27/05/2025

NVIDIA and Google Partnership Gains Momentum With the Latest Blackwell and Gemini Announcements

NVIDIA and Google share a long-standing relationship rooted in advancing AI inno...

22/05/2025

Sale Into Summer With 40% Off GeForce NOW Six-Month Performance Memberships

GeForce NOW is turning up the heat this summer with a hot new deal. For a limited time, save 40% on six-month Performance memberships and enjoy premium GeForce ...

21/05/2025

NVIDIA and SAP Bring AI Agents to the Physical World

As robots increasingly make their way to the largest enterprises' manufacturing plants and warehouses, the need for access to critical business and operatio...

20/05/2025

Siemens Makes Factory Floors Smarter With Industrial AI

Industrial AI is transforming how factories operate, innovate and scale. The convergence of AI, simulation and digital twins is poised to unlock new levels of ...

19/05/2025

NVIDIA and Microsoft Accelerate Agentic AI Innovation, From Cloud to PC

Agentic AI is redefining scientific discovery and unlocking research breakthroughs and innovations across industries. Through deepened collaboration, NVIDIA and...

19/05/2025

NVIDIA Research Breakthroughs Put Advanced Robots in Motion

Across robot training and development, NVIDIA Research is uncovering breakthroughs in areas such as multimodal generative AI and synthetic data generation. The...

19/05/2025

NVIDIA and Microsoft Advance Development on RTX AI PCs

Generative AI is transforming PC software into breakthrough experiences - from digital humans to writing assistants, intelligent agents and creative tools. NVI...

18/05/2025

NVIDIA CEO Envisions AI Infrastructure Industry Worth Trillions of Dollars'

Electricity. The Internet. Now it's time for another major technology, AI, to sweep the globe. NVIDIA founder and CEO Jensen Huang took the stage at a pack...

18/05/2025

NVIDIA Expands Omniverse Blueprint for AI Factory Digital Twins With New Ecosystem Integrations, Development Tools

Empowering engineering teams with more tools for building AI factories, NVIDIA t...

18/05/2025

AI Blueprint for Video Search and Summarization Now Available to Deploy Video Analytics AI Agents Across Industries

The age of video analytics AI agents is here. Video is one of the defining feat...

18/05/2025

Semiconductor Industry Accelerates Design Manufacturing With NVIDIA Blackwell and CUDA-X

TSMC, Cadence, KLA, Siemens and Synopsys are advancing semiconductor manufacturi...

18/05/2025

NVIDIA Grace CPU C1 Gains Broad Support in Edge, Telco and Storage

NVIDIA is highlighting significant momentum for its new Grace CPU C1 this week at the COMPUTEX trade show in Taipei, with a strong showing of support from key o...

18/05/2025

NVIDIA-Powered Supercomputer to Enable Quantum Leap for Taiwan Research

Researchers across Taiwan are tackling complex challenges in AI development, climate science and quantum computing. Their work will soon be boosted by a new sup...

18/05/2025

That's One Smart Hospital! Taiwan Medical Centers Deploy Life-Saving Innovations With NVIDIA System-Builder Partners

Leading healthcare organizations across the globe are using agentic AI, robotics...

18/05/2025

NVIDIA Grows Quantum Computing Ecosystem With Taiwan Manufacturers and Supercomputing

Quantum computing promises to shorten the path to solving some of the world'...

15/05/2025

Into the Omniverse: Computational Fluid Dynamics Simulation Finds Smoothest Flow With AI-Driven Digital Twins

Editor's note: This post is part of Into the Omniverse, a series focused on ...

15/05/2025

Exploring the Revenue-Generating Potential of AI Factories

AI is creating value for everyone - from researchers in drug discovery to quantitative analysts navigating financial market changes. The faster an AI system ca...