Sony Pixel Power calrec Sony

How NVIDIA AI Foundry Lets Enterprises Forge Custom Generative AI Models

23/07/2024

Businesses seeking to harness the power of AI need customized models tailored to their specific industry needs.

NVIDIA AI Foundry is a service that enables enterprises to use data, accelerated computing and software tools to create and deploy custom models that can supercharge their generative AI initiatives.

Just as TSMC manufactures chips designed by other companies, NVIDIA AI Foundry provides the infrastructure and tools for other companies to develop and customize AI models - using DGX Cloud, foundation models, NVIDIA NeMo software, NVIDIA expertise, as well as ecosystem tools and support.

The key difference is the product: TSMC produces physical semiconductor chips, while NVIDIA AI Foundry helps create custom models. Both enable innovation and connect to a vast ecosystem of tools and partners.

Enterprises can use AI Foundry to customize NVIDIA and open community models, including the new Llama 3.1 collection, as well as NVIDIA Nemotron, CodeGemma by Google DeepMind, CodeLlama, Gemma by Google DeepMind, Mistral, Mixtral, Phi-3, StarCoder2 and others.

Industry Pioneers Drive AI Innovation Industry leaders Amdocs, Capital One, Getty Images, KT, Hyundai Motor Company, SAP, ServiceNow and Snowflake are among the first using NVIDIA AI Foundry. These pioneers are setting the stage for a new era of AI-driven innovation in enterprise software, technology, communications and media.

Organizations deploying AI can gain a competitive edge with custom models that incorporate industry and business knowledge, said Jeremy Barnes, vice president of AI Product at ServiceNow. ServiceNow is using NVIDIA AI Foundry to fine-tune and deploy models that can integrate easily within customers' existing workflows.

The Pillars of NVIDIA AI Foundry NVIDIA AI Foundry is supported by the key pillars of foundation models, enterprise software, accelerated computing, expert support and a broad partner ecosystem.

Its software includes AI foundation models from NVIDIA and the AI community as well as the complete NVIDIA NeMo software platform for fast-tracking model development.

The computing muscle of NVIDIA AI Foundry is NVIDIA DGX Cloud, a network of accelerated compute resources co-engineered with the world's leading public clouds - Amazon Web Services, Google Cloud and Oracle Cloud Infrastructure. With DGX Cloud, AI Foundry customers can develop and fine-tune custom generative AI applications with unprecedented ease and efficiency, and scale their AI initiatives as needed without significant upfront investments in hardware. This flexibility is crucial for businesses looking to stay agile in a rapidly changing market.

If an NVIDIA AI Foundry customer needs assistance, NVIDIA AI Enterprise experts are on hand to help. NVIDIA experts can walk customers through each of the steps required to build, fine-tune and deploy their models with proprietary data, ensuring the models tightly align with their business requirements.

NVIDIA AI Foundry customers have access to a global ecosystem of partners that can provide a full range of support. Accenture, Deloitte, Infosys, Tata Consultancy Services and Wipro are among the NVIDIA partners that offer AI Foundry consulting services that encompass design, implementation and management of AI-driven digital transformation projects. Accenture is first to offer its own AI Foundry-based offering for custom model development, the Accenture AI Refinery framework.

Additionally, service delivery partners such as Data Monsters, Quantiphi, Slalom and SoftServe help enterprises navigate the complexities of integrating AI into their existing IT landscapes, ensuring that AI applications are scalable, secure and aligned with business objectives.

Customers can develop NVIDIA AI Foundry models for production using AIOps and MLOps platforms from NVIDIA partners, including ActiveFence, AutoAlign, Cleanlab, DataDog, Dataiku, Dataloop, DataRobot, Deepchecks, Domino Data Lab, Fiddler AI, Giskard, New Relic, Scale, Tumeryk and Weights & Biases.

Customers can output their AI Foundry models as NVIDIA NIM inference microservices - which include the custom model, optimized engines and a standard API - to run on their preferred accelerated infrastructure.

Inferencing solutions like NVIDIA TensorRT-LLM deliver improved efficiency for Llama 3.1 models to minimize latency and maximize throughput. This enables enterprises to generate tokens faster while reducing total cost of running the models in production. Enterprise-grade support and security is provided by the NVIDIA AI Enterprise software suite.

NVIDIA NIM and TensorRT-LLM minimize inference latency and maximize throughput for Llama 3.1 models to generate tokens faster. The broad range of deployment options includes NVIDIA-Certified Systems from global server manufacturing partners including Cisco, Dell Technologies, Hewlett Packard Enterprise, Lenovo and Supermicro, as well as cloud instances from Amazon Web Services, Google Cloud and Oracle Cloud Infrastructure.

Additionally, Together AI, a leading AI acceleration cloud, today announced it will enable its ecosystem of over 100,000 developers and enterprises to use its NVIDIA GPU-accelerated inference stack to deploy Llama 3.1 endpoints and other open models on DGX Cloud.

Every enterprise running generative AI applications wants a faster user experience, with greater efficiency and lower cost, said Vipul Ved Prakash, founder and CEO of Together AI. Now, developers and enterprises using the Together Inference Engine can maximize performance, scalability and security on NVIDIA DGX Cloud.

NVIDIA NeMo Speeds and Simplifies Custom Model Development With NVIDIA NeMo integrated into AI Foundry, developers have at their fingertips the tools needed to curate data, customize foundation models and evaluate performance. NeMo technologies include:

NeMo Curator is a GPU-accelerated data
LINK: https://blogs.nvidia.com/blog/ai-foundry-enterprise-generative-ai/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

04/08/2026

Dalet Announces Commercial Availability of Dalia, Bringing Media-Aware Agentic AI to Enterprise Productions

Dalet, a leading technology and service provider for media-rich organizations, t...

04/07/2026

Detective Conan: Fallen Angel of the Highway Opens in Dolby Cinemas Across Japan, Presented in Dolby Atmos and Dolby ...

April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

23/05/2026

FOX Sports, IMS Productions Scale Up Indy 500 Production With New In-Car Cameras, AR Graphics, Cinematic Sets

In its second year as rightsholder, FOX Sports goes bigger across the board for ...

23/05/2026

Inside Apple TVs MLS iPhone Production with Royce Dickerson, Apple Live Sports, Executive Producer

Tonight's MLS matchup between the LA Galaxy and the Houston Dynamo FC will m...

23/05/2026

IK Multimedia reveal ReSing Voices Japanese Pack

AI-powered vocal tool gains first new language expansion IK Multimedia's AI-powered voice-creation software has seen a number of updates since it launch...

23/05/2026

Building a better future: Nielsen celebrates Global Volunteer Month and Earth Day 2026 with record participation

Nielsen Global Leadership Network graduates celebrate Earth Day 2026 Nielsen vo...

23/05/2026

Gray Media Names New Station General Managers

Share Copy link Facebook X Linkedin Bluesky Email...

23/05/2026

New CIMM Paper Urges Industry to Rethink How Media Is Evaluated

Share Copy link Facebook X Linkedin Bluesky Email...

23/05/2026

Lawo to Showcase Edge One, Efficient IP Workflows at InfoComm 2026

Share Copy link Facebook X Linkedin Bluesky Email...

23/05/2026

Spectrum Launches Ultra-Low Latency Internet

Share Copy link Facebook X Linkedin Bluesky Email...

22/05/2026

Germanys Magenta TV Selects DMC to Provide FIFA World Cup Technical Support for Studios in Munich, New York City

Germany's Magenta TV, which will have 44 exclusive FIFA World Cup match broa...

22/05/2026

DAZN Grabs IFAF Flag Football Global Rights

DAZN, the world's leading sports entertainment platform, has acquired global broadcast rights to the International Federation of American Football's ( I...

22/05/2026

ATHLOS 2026 Season Set for October Debut in London; Aurora Media Worldwide Named Host Broadcast Partner

ATHLOS, the all-women's professional track and field league, has announced i...

22/05/2026

NATAS to Stream Sports, News, and Documentary Emmy Awards Live on YouTube

The National Academy of Television Arts & Sciences (NATAS) today announced that the 47th Annual Sports Emmy Awards and the 47th Annual News & Documentary Emmy A...

22/05/2026

Wooden Camera Rolls Out New Blackmagic URSA Accessories

Wooden Camera today announced the release of new accessories for the Blackmagic URSA Cine Immersive. The new lineup includes a redesigned Top Plate and Side Rai...

22/05/2026

YES Network, OTT Advisors Extend Streaming Partnership for Sixth Season

YES Network and OTT Advisors have announced a sixth consecutive season of their streaming partnership, continuing their collaboration on the Gotham app. OTT Adv...

22/05/2026

NESN Monster Week' Returns With Full Red Sox Broadcast From Atop Green Monster

NESN, New England's premier sports network, will again turn its camera to Fe...

22/05/2026

Dale Pro Audio RF Over Fiber Webinar Set for May 28

Dale Pro Audio is hosting an RF over Fiber Livestream Webinar on May 28 from 1-2:30 pm EST. With major sporting events and large-scale productions putting incre...

22/05/2026

Audio-Technica Appoints Humrichouser, Schanz to New Roles

Audio-Technica has announced key leadership appointments designed to further strengthen its sales organization and drive continued growth across the Americas. M...

22/05/2026

Scott Coker Launches Global MMA League With $60 Million in Backing

After nearly four decades shaping the global combat sports landscape, Scott Coker has announced a powerful return as he looks to build a new international mixed...

22/05/2026

Skyline Launches xOps Vanguard Runway for Autonomous Era

Skyline Communications, the company behind the globally deployed DataMiner xOps platform, today announced the launch of xOps Vanguard Runway, a strategic accele...

22/05/2026

The American Rodeo Takes Over Globe Life Field for Championship Weekend

For the fully onsite production, 30 cameras - including a SkyCam and Megalodon - will capture the action in Texas One of the world's biggest rodeo producti...

22/05/2026

Argentinas Torneos Taps Imagine Versio for Playout Operations Upgrade

Leading Argentina-based sports media company Torneos y Competencias S.A. has modernized its playout operations, implementing a fully redundant, multichannel env...

22/05/2026

Owl AI and Major League Pickleball Go Live with First-Ever AI Officiating System Powered by Broadcast Cameras and the Cloud

As the 2026 Major League Pickleball season kicks off this weekend in Dallas, it ...

22/05/2026

Shure, Edge Sound Research Look to Innovate via Partnership

Shure has become a minority investor in Edge Sound Research, a start-up company that is developing new experiential audio technologies that redefine how many au...

22/05/2026

SVG Rewind: MLBs UmpCam AR System Puts Fans Inside the Strike Zone Like Never Before

In advance of this year's Sports Emmy Awards, SVG is taking a deep dive into...

22/05/2026

Jelly Roll Offers Up 2026 Stanley Cup Playoff Theme Song for NHL, Amazon Music

The National Hockey League (NHL) and Amazon Music announced that GRAMMY Award-winning superstar Jelly Roll will provide the official theme song of the 2026 Stan...

22/05/2026

David Pogue, Andy Beach Keynotes Highlight Silicon Valley Video Summer Camp, July 14 at De Anza College

David Pogue will keynote SVV Summer Camp and discuss Apple at 50: How the World...

22/05/2026

FOX Sports, IMS Productions Scale Up Indy 500 Production in Year Two With New In-Car Cameras, AR Graphics, and Cinematic Sets

In its second year as rightsholder, FOX Sports goes bigger across the board for ...

22/05/2026

FOX Sports' Indy 500 Director Mitch Riggin on the Tech and Storytelling for the Greatest Spectacle in Racing

The broadcaster is drawing on lessons learned in its first year of covering the ...

22/05/2026

How Spotify's Rebuilt Ad Platform Is Delivering New Value for Brands

At our 2026 Investor Day, we shared an inside look at the rebuild of our advertising business. This pivot to our own purpose-built platform is already driving s...

22/05/2026

Spotify Levels Up Our Podcast Experience With New Features for Fans and Creators

Podcasting on Spotify continues to grow, and so do the ways listeners engage with it. At Investor Day 2026, we shared how we're building the next chapter of...

22/05/2026

CEDAR Audio introduce Icons Bundles

Limited-time collections now available Restoration experts CEDAR Audio have recently launched a new line of Icons plug-ins that make their powerful processo...

22/05/2026

Boss expand PS-1 Plugout Pedal

Three new classics join Model Pass line-up Boss' PX-1 Plugout Pedal offers an innovative approach to guitar pedals, providing users with a hardware stom...

22/05/2026

SGL Carbon commissions photovoltaic system and lays the foundation for a new nitrogen plant at its Meitingen site

At its Meitingen site, SGL Carbon has implemented two key projects to further de...

22/05/2026

Statement regarding 2026 National NAIDOC Lifetime Achievement Award for the late Rhoda Roberts AO

Statement regarding 2026 National NAIDOC Lifetime Achievement Award for the late...

22/05/2026

Polsat Reclaims Second Place and ByteDance Enters Top 10 as Polish Viewing Moves Beyond the Living Room in April

Latest data reveals steady distributor rankings, a seasonal shift toward digital...

22/05/2026

FCC Votes to Update Disaster Information Reporting System

Share Copy link Facebook X Linkedin Bluesky Email...

22/05/2026

Amagi delivers 30 per cent revenue growth in FY26 Adjuste...

Amagi Media Labs Limited (NSE: AMAGI, BSE: 544679), a cloud-native SaaS platform providing AI-enabled solutions to global media and entertainment companies, tod...

22/05/2026

Annima Post Relies on Cintel to Revive Classic Mexican Films

An nima Post Relies on Cintel to Revive Classic Mexican Films Brie Clayton May 22, 2026 0 Comments Film scanner and DaVinci Resolve Studio help manage...

22/05/2026

Boris FX Sapphire Adds Optical Beauty and Hypnotic Textures

Boris FX Sapphire Adds Optical Beauty and Hypnotic Textures Jessie Electa Petrov May 22, 2026 0 Comments The 2026.5 release introduces advanced defocu...

22/05/2026

Deployment Preserves Trusted Workflows While Enabling a P...

Deployment Preserves Trusted Workflows While Enabling a Path to UHD and SMPTE ST 2110 Leading Argentina-based sports media company Torneos y Competencias S.A....

22/05/2026

Study: AI Labeling Does Not Hurt Video Ad Performance

Share Copy link Facebook X Linkedin Bluesky Email...

22/05/2026

NAB Show Makes 200+ Sessions Available on Demand

Share Copy link Facebook X Linkedin Bluesky Email...

22/05/2026

Torneos Upgrades Multichannel Playout with Imagine's Versio

Share Copy link Facebook X Linkedin Bluesky Email...

22/05/2026

Ex-Husband, Current Husband, One Wild Rescue: Korean Action Comedy Husbands in Action' Premieres June 19

Back to All News Ex-Husband, Current Husband, One Wild Rescue: Korean Action Co...

22/05/2026

Beta Da Silva hosts live performances from 20 new Irish artists in Sessions from Oblivion on 2FM's New Music Show

Catch the latest in Irish music live from venues such as Whelan's, R is n Du...

21/05/2026

CBS Sports Expands WNBA Tip-Off Show To Cover Half of 20-Game, Regular-Season Package

Game Creek Video Columbia and Celtic, NEP Supershooter 8 will house onsite produ...