Sony Pixel Power calrec Sony

How NVIDIA AI Foundry Lets Enterprises Forge Custom Generative AI Models

23/07/2024

Businesses seeking to harness the power of AI need customized models tailored to their specific industry needs.

NVIDIA AI Foundry is a service that enables enterprises to use data, accelerated computing and software tools to create and deploy custom models that can supercharge their generative AI initiatives.

Just as TSMC manufactures chips designed by other companies, NVIDIA AI Foundry provides the infrastructure and tools for other companies to develop and customize AI models - using DGX Cloud, foundation models, NVIDIA NeMo software, NVIDIA expertise, as well as ecosystem tools and support.

The key difference is the product: TSMC produces physical semiconductor chips, while NVIDIA AI Foundry helps create custom models. Both enable innovation and connect to a vast ecosystem of tools and partners.

Enterprises can use AI Foundry to customize NVIDIA and open community models, including the new Llama 3.1 collection, as well as NVIDIA Nemotron, CodeGemma by Google DeepMind, CodeLlama, Gemma by Google DeepMind, Mistral, Mixtral, Phi-3, StarCoder2 and others.

Industry Pioneers Drive AI Innovation Industry leaders Amdocs, Capital One, Getty Images, KT, Hyundai Motor Company, SAP, ServiceNow and Snowflake are among the first using NVIDIA AI Foundry. These pioneers are setting the stage for a new era of AI-driven innovation in enterprise software, technology, communications and media.

Organizations deploying AI can gain a competitive edge with custom models that incorporate industry and business knowledge, said Jeremy Barnes, vice president of AI Product at ServiceNow. ServiceNow is using NVIDIA AI Foundry to fine-tune and deploy models that can integrate easily within customers' existing workflows.

The Pillars of NVIDIA AI Foundry NVIDIA AI Foundry is supported by the key pillars of foundation models, enterprise software, accelerated computing, expert support and a broad partner ecosystem.

Its software includes AI foundation models from NVIDIA and the AI community as well as the complete NVIDIA NeMo software platform for fast-tracking model development.

The computing muscle of NVIDIA AI Foundry is NVIDIA DGX Cloud, a network of accelerated compute resources co-engineered with the world's leading public clouds - Amazon Web Services, Google Cloud and Oracle Cloud Infrastructure. With DGX Cloud, AI Foundry customers can develop and fine-tune custom generative AI applications with unprecedented ease and efficiency, and scale their AI initiatives as needed without significant upfront investments in hardware. This flexibility is crucial for businesses looking to stay agile in a rapidly changing market.

If an NVIDIA AI Foundry customer needs assistance, NVIDIA AI Enterprise experts are on hand to help. NVIDIA experts can walk customers through each of the steps required to build, fine-tune and deploy their models with proprietary data, ensuring the models tightly align with their business requirements.

NVIDIA AI Foundry customers have access to a global ecosystem of partners that can provide a full range of support. Accenture, Deloitte, Infosys, Tata Consultancy Services and Wipro are among the NVIDIA partners that offer AI Foundry consulting services that encompass design, implementation and management of AI-driven digital transformation projects. Accenture is first to offer its own AI Foundry-based offering for custom model development, the Accenture AI Refinery framework.

Additionally, service delivery partners such as Data Monsters, Quantiphi, Slalom and SoftServe help enterprises navigate the complexities of integrating AI into their existing IT landscapes, ensuring that AI applications are scalable, secure and aligned with business objectives.

Customers can develop NVIDIA AI Foundry models for production using AIOps and MLOps platforms from NVIDIA partners, including ActiveFence, AutoAlign, Cleanlab, DataDog, Dataiku, Dataloop, DataRobot, Deepchecks, Domino Data Lab, Fiddler AI, Giskard, New Relic, Scale, Tumeryk and Weights & Biases.

Customers can output their AI Foundry models as NVIDIA NIM inference microservices - which include the custom model, optimized engines and a standard API - to run on their preferred accelerated infrastructure.

Inferencing solutions like NVIDIA TensorRT-LLM deliver improved efficiency for Llama 3.1 models to minimize latency and maximize throughput. This enables enterprises to generate tokens faster while reducing total cost of running the models in production. Enterprise-grade support and security is provided by the NVIDIA AI Enterprise software suite.

NVIDIA NIM and TensorRT-LLM minimize inference latency and maximize throughput for Llama 3.1 models to generate tokens faster. The broad range of deployment options includes NVIDIA-Certified Systems from global server manufacturing partners including Cisco, Dell Technologies, Hewlett Packard Enterprise, Lenovo and Supermicro, as well as cloud instances from Amazon Web Services, Google Cloud and Oracle Cloud Infrastructure.

Additionally, Together AI, a leading AI acceleration cloud, today announced it will enable its ecosystem of over 100,000 developers and enterprises to use its NVIDIA GPU-accelerated inference stack to deploy Llama 3.1 endpoints and other open models on DGX Cloud.

Every enterprise running generative AI applications wants a faster user experience, with greater efficiency and lower cost, said Vipul Ved Prakash, founder and CEO of Together AI. Now, developers and enterprises using the Together Inference Engine can maximize performance, scalability and security on NVIDIA DGX Cloud.

NVIDIA NeMo Speeds and Simplifies Custom Model Development With NVIDIA NeMo integrated into AI Foundry, developers have at their fingertips the tools needed to curate data, customize foundation models and evaluate performance. NeMo technologies include:

NeMo Curator is a GPU-accelerated data
LINK: https://blogs.nvidia.com/blog/ai-foundry-enterprise-generative-ai/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

14/02/2026

Cineverse Acquires TV Monetization Platform IndiCue

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

ESPN's Audiences for College Basketball On Track for Major Growth

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

TCL Display Technologies Deployed at Winter Olympics

Share Copy link Facebook X Linkedin Bluesky Email...

14/02/2026

Boston Conservatory Orchestra Helps Peter and Leonardo Dugan Complete Their Dream Piece

Boston Conservatory Orchestra Helps Peter and Leonardo Dugan Complete Their Dre...

13/02/2026

OBS Accelerates Shift to Cloud

Olympic Broadcasting Services (OBS) has provided an update on its adoption of the cloud as it continues on its journey to fully migrate to IT-based systems by 2...

13/02/2026

France Tlvisions Launches France 2 UHD with Dolby Vision and Dolby Atmos to Max out AC-4 Experiences for Winter Olympics Fans

France T l visions has successfully launched France 2 UHD featuring Dolby Vision...

13/02/2026

OBS Expands Athlete Moment,' Family Reunions to Capture Human Side of Winter Games

Partnering with Worldwide Olympic Partner TCL, OBS deploys connected Athlete Mom...

13/02/2026

Men's Figure Skating Photo Gallery

The men's figure skating long-form program is tonight, and it promises to be an exciting night for fans in the stands, fans at home, and even the production...

13/02/2026

Entertainment Takes the NBA Court in a Big Way

With new partnership between the league and NBC, workflows distinguish more between live, broadcast sound There'll be a lot new for the 75th NBA All-Star W...

13/02/2026

SVG GameDay, Episode 3: Sean Tabler - Producing Hockey in the City of Angels

In-venue and creative video staffers at the professional and collegiate level have one major thing in common: the intensity and attention to detail ramps up dur...

13/02/2026

Teradek Introduces RF-X, Revolutionizing Mission-Critical Signal Redundancy

Teradek announces the launch of RF-X Auto Switcher, a revolutionary appliance designed to deliver flawless, uncompromised signal integrity for the world's m...

13/02/2026

Synamedia & Globecast Selected for FA Cup Cloud Distribution

Globecast and Synamedia announces that Pitch International (Pitch), the leading London-based sports marketing agency, has gone live with cloud-based distributi...

13/02/2026

Ratings Roundup: NBC Sports' Legendary February Hits Record Viewership Levels

Ratings Roundup is a rundown of recent rating news and is derived from press rel...

13/02/2026

NBC Olympics' Amy Rosenfeld on the Drone Craze, Friends & Family Moments, Stamford's Role for Milano Cortina

Far from the action in the snow and on the ice, the team controls the production...

13/02/2026

2026 Daytona 500: FOX Sports' Mike Davies, George Grill on Working Within an IP-Based Compound, Solving the Ops Puzzle of the Super Bowl of Racing

The Daytona 500 is called The Super Bowl of Racing for a reason. Whether it's the culmination to five days of action on the track, the sheer size and scop...

13/02/2026

OBS Expands AI-Powered Content Workflows

For the Milano Cortina Games, Olympic Broadcasting Services (OBS) is delivering more than 6,500 hours of content, with more than 900 hours of live action, sprea...

13/02/2026

NBC Sports Director Pierre Moossa Previews NBC's First NBA All-Star Production in 24 Years

After 24-year absence, NBC Sports returns to NBA All-Star Weekend with unique ca...

13/02/2026

Film Festival Watch: 18 Sundance Institute-Supported Projects To Watch at the 2026 Berlin International Film Festival

By Jessica Herndon We may have just wrapped an unforgettable 2026 Sundance Film...

13/02/2026

Give Me the Backstory: Get to Know Amanda Kramer, the Writer-Director Behind By Design

By Jessica Herndon One of the most exciting things about the Sundance Film Fest...

13/02/2026

Women in Podcasting Craft New Connections at Spotify's Galentine's Day Celebration in LA

This Wednesday in Los Angeles, Spotify brought together a group of podcast creat...

13/02/2026

Spotify and LoveShackFancy Bring Galentine's Glam to NYC, Featuring Special Performance by Joshua Basset

Yesterday, Spotify and LoveShackFancy hosted a Galentine's and Gents Lunch a...

13/02/2026

L3Harris Successfully Completes First Phase of P25 Transition for Florida SLERS

The upgrade to a Project 25 network provides state agencies communicating on the Statewide Law Enforcement Radio System flexibility to tailor the network to the...

13/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24...

Riedel Communications has officially opened a new office in Kuala Lumpur, Malaysia, marking a strategic expansion of its global Customer Success and IT software...

13/02/2026

ES Broadcast Hire duo celebrate 10-year anniversary with...

Two of ES Broadcast Hire's longest-serving employees recently celebrated a decade working for the company. Annie Breislin, Operations Manager, and Charles ...

13/02/2026

Disguise Opens Experience Center and Office in Atlanta

Disguise, the award-winning technology company powering global experiences, today unveils a new 8,000-square-foot office and Experience Center in Atlanta, creat...

13/02/2026

Mavis Expands External Camera Support with Accsoon SeeMo...

At BSC Expo 2026, Mavis announced full support for the Accsoon SeeMo series of iOS camera adapters across Mavis Camera and Mavis Monitor apps. This new integrat...

13/02/2026

Butcher Bird Studios Keeps Signals Flowing Seamlessly Acr...

Executing technically ambitious live streams, virtual productions, and immersive media today requires talent, creativity, and the right supporting technology. L...

13/02/2026

LTN makes key appointments and introduces new Technology...

Michal Miskin-Amir, Jonathan Stanton and Bobby Bond to lead technical advances amid surge in demand for LTN's IP video transport services as satellite capac...

13/02/2026

NATO Upgrades Broadcast Studio with Grass Valley Cameras

Grass Valley, the pioneering media and entertainment technology innovator, has won a competitive NATO-wide tender to provide the new camera system for NATO'...

13/02/2026

Digital Azul strengthens remote production strategy with...

Wireless IP intercom underpins agile, multi-location live production workflows Digital Azul, the independent production powerhouse specialising in complex liv...

13/02/2026

Actus Digital Sets a New Standard for QA Monitoring and C...

Actus Digital, a LiveU company, will unveil major new enhancements to its Actus X Intelligent Monitoring Platform at NAB Show (LiveU booth N1740), reinforcing i...

13/02/2026

FA Cup goes IP with Pitch International plus Synamedia an...

Globecast, a worldwide leader in broadcast services, and leading video software provider, Synamedia, today announced that Pitch International (Pitch), the leadi...

13/02/2026

Rai Selects Imagine Selenio Network Processor for IP Migration

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

CIMM Details Research Plans for 2026 and New Board Appointments

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Teradek Unveils RF-A Auto Switcher

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Spectrum Launches 'Invincible Wifi'

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Actus Digital to Introduce Actus X Platform Enhancements At NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Sennheiser Wireless Spectera Solution Tackles Super Bowl LX With Ease

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

Nate Bargatze to Receive 2026 NAB Television Chairman's Award

Share Copy link Facebook X Linkedin Bluesky Email...

13/02/2026

UKTV Highlights Saturday February 14th - Friday February 20th 2026

What can I watch on UKTV this week?What can I stream on U this week? This guide highlights romantic dramas for Valentine's Day, alternative relationship t...

13/02/2026

New RT series tells stranger-than-fiction stories of Irish con artists

New RT series tells stranger-than-fiction stories of Irish con artists Swindlers airs Wednesday 18 February, 9.35pm on RT One and RT Player Swindlers, a...

12/02/2026

Chyron Merges Live Web Content and CG Graphics with PRIME 5.3

Chyron unveils PRIME 5.3, the latest software release of the company's powerful engine for live production graphics. PRIME 5.3 delivers the first official i...

12/02/2026

SVG New Sponsor Spotlight: Interra Systems' Anupama Anantharaman on Protecting Live Sports Quality Across IP and OTT Workflows

The vendor's VP of Product Management explains how quality assurance, monito...

12/02/2026

LTN Makes Key Appointments and Introduces New Technology Organization

LTN announces the appointment of three experienced executives to lead its new Technology organization: Michal Miskin-Amir as EVP and Head of Technology, Jonatha...

12/02/2026

Riedel Opens Kuala Lumpur Office to Strengthen Global 24/7 Software and IT Support

Riedel Communications has officially opened a new office in Kuala Lumpur, Malays...