Sony Pixel Power calrec Sony

NVIDIA Expands Collaboration With Microsoft to Help Developers Build, Deploy AI Applications Faster

21/05/2024

If optimized AI workflows are like a perfectly tuned orchestra - where each component, from hardware infrastructure to software libraries, hits exactly the right note - then the long-standing harmony between NVIDIA and Microsoft is music to developers' ears.

The latest AI models developed by Microsoft, including the Phi-3 family of small language models, are being optimized to run on NVIDIA GPUs and made available as NVIDIA NIM inference microservices. Other microservices developed by NVIDIA, such as the cuOpt route optimization AI, are regularly added to Microsoft Azure Marketplace as part of the NVIDIA AI Enterprise software platform.

In addition to these AI technologies, NVIDIA and Microsoft are delivering a growing set of optimizations and integrations for developers creating high-performance AI apps for PCs powered by NVIDIA GeForce RTX and NVIDIA RTX GPUs.

Building on the progress shared at NVIDIA GTC, the two companies are furthering this ongoing collaboration at Microsoft Build, an annual developer event, taking place this year in Seattle through May 23.

Accelerating Microsoft's Phi-3 Models Microsoft is expanding its family of Phi-3 open small language models, adding small (7-billion-parameter) and medium (14-billion-parameter) models similar to its Phi-3-mini, which has 3.8 billion parameters. It's also introducing a new 4.2-billion-parameter multimodal model, Phi-3-vision, that supports images and text.

All of these models are GPU-optimized with NVIDIA TensorRT-LLM and available as NVIDIA NIMs, which are accelerated inference microservices with a standard application programming interface (API) that can be deployed anywhere.

APIs for the NIM-powered Phi-3 models are available at ai.nvidia.com and through NVIDIA AI Enterprise on the Azure Marketplace.

NVIDIA cuOpt Now Available on Azure Marketplace NVIDIA cuOpt, a GPU-accelerated AI microservice for route optimization, is now available in Azure Marketplace via NVIDIA AI Enterprise. cuOpt features massively parallel algorithms that enable real-time logistics management for shipping services, railway systems, warehouses and factories.

The model has set two dozen world records on major routing benchmarks, demonstrating the best accuracy and fastest times. It could save billions of dollars for the logistics and supply chain industries by optimizing vehicle routes, saving travel time and minimizing idle periods.

Through Azure Marketplace, developers can easily integrate the cuOpt microservice with Azure Maps to support teal-time logistics management and other cloud-based workflows, backed by enterprise-grade management tools and security.

Optimizing AI Performance on PCs With NVIDIA RTX The NVIDIA accelerated computing platform is the backbone of modern AI - helping developers build solutions for over 100 million Windows GeForce RTX-powered PCs and NVIDIA RTX-powered workstations worldwide.

NVIDIA and Microsoft are delivering new optimizations and integrations to Windows developers to accelerate AI in next-generation PC and workstation applications. These include:

Faster inference performance for large language models via the NVIDIA DirectX driver, the Generative AI ONNX Runtime extension and DirectML. These optimizations, available now in the GeForce Game Ready, NVIDIA Studio and NVIDIA RTX Enterprise Drivers, deliver up to 3x faster performance on NVIDIA and GeForce RTX GPUs.

Optimized performance on RTX GPUs for AI models like Stable Diffusion and Whisper via WebNN, an API that enables developers to accelerate AI models in web applications using on-device hardware.

With Windows set to support PyTorch through DirectML, thousands of Hugging Face models will work in Windows natively. NVIDIA and Microsoft are collaborating to scale performance on more than 100 million RTX GPUs.

Join NVIDIA at Microsoft Build Conference attendees can visit NVIDIA booth FP28 to meet developer experts and experience live demos of NVIDIA NIM, NVIDIA cuOpt, NVIDIA Omniverse and the NVIDIA RTX AI platform. The booth also highlights the NVIDIA MONAI platform for medical imaging workflows and NVIDIA BioNeMo generative AI platform for drug discovery - both available on Azure as part of NVIDIA AI Enterprise.

Attend sessions with NVIDIA speakers to dive into the capabilities of the NVIDIA RTX AI platform on Windows PCs and discover how to deploy generative AI and digital twin tools on Microsoft Azure.

And sign up for the Developer Showcase, taking place Wednesday, to discover how developers are building innovative generative AI using NVIDIA AI software on Azure.
LINK: https://blogs.nvidia.com/blog/microsoft-build-optimized-ai-developers/...
See more stories from nvidia

More from Nvidia

15/07/2025

Deadline Extended - Create a Project G-Assist Plug-In for a Chance to Win an NVIDIA GeForce RTX GPU and Laptop

Submissions for NVIDIA's Plug and Play: Project G-Assist Plug-In Hackathon a...

14/07/2025

NVIDIA CEO Jensen Huang Promotes AI in Washington, DC and China

This month, NVIDIA founder and CEO Jensen Huang promoted AI in both Washington, D.C. and Beijing - emphasizing the benefits that AI will bring to business and s...

11/07/2025

A Gaming GPU Helps Crack the Code on a Thousand-Year Cultural Conversation

Ceramics - the humble mix of earth, fire and artistry - have been part of a global conversation for millennia. From Tang Dynasty trade routes to Renaissance pa...

10/07/2025

From Terabytes to Turnkey: AI-Powered Climate Models Go Mainstream

In the race to understand our planet's changing climate, speed and accuracy are everything. But today's most widely used climate simulators often strugg...

10/07/2025

Indonesia on Track to Achieve Sovereign AI Goals With NVIDIA, Cisco and IOH

As one of the world's largest emerging markets, Indonesia is making strides toward its Golden 2045 Vision - an initiative tapping digital technologies and...

10/07/2025

Reach the PEAK' on GeForce NOW

Grab a friend and climb toward the clouds - PEAK is now available on GeForce NOW, enabling members to try the hugely popular indie hit on virtually any device. ...

10/07/2025

How to Run Coding Assistants for Free on RTX AI PCs and Workstations

Coding assistants or copilots - AI-powered assistants that can suggest, explain and debug code - are fundamentally changing how software is developed for both e...

08/07/2025

Asking an Encyclopedia-Sized Question: How To Make the World Smarter with Multi-Million Token Real-Time Inference

Modern AI applications increasingly rely on models that combine huge parameter c...

03/07/2025

GeForce NOW's 20 July Games Bring the Heat to the Cloud

The forecast this month is showing a 100% chance of epic gaming. Catch the scorching lineup of 20 titles coming to the cloud, which gamers can play whether indo...

02/07/2025

NVIDIA RTX AI Accelerates FLUX.1 Kontext - Now Available for Download

Black Forest Labs, one of the world's leading AI research labs, just changed the game for image generation. The lab's FLUX.1 image models have earned g...

01/07/2025

How AI Factories Can Help Relieve Grid Stress

In many parts of the world, including major technology hubs in the U.S., there's a yearslong wait for AI factories to come online, pending the buildout of n...

26/06/2025

Run Google DeepMind's Gemma 3n on NVIDIA Jetson and RTX

As of today, NVIDIA now supports the general availability of Gemma 3n on NVIDIA RTX and Jetson. Gemma, previewed by Google DeepMind at Google I/O last month, in...

26/06/2025

Into the Omniverse: World Foundation Models Advance Autonomous Vehicle Simulation and Safety

Editor's note: This blog is a part of Into the Omniverse, a series focused o...

26/06/2025

Startup Uses NVIDIA RTX-Powered Generative AI to Make Coolers, Cooler

Mark Theriault founded the startup FITY envisioning a line of clever cooling products: cold drink holders that come with freezable pucks to keep beverages cold ...

26/06/2025

Game On With GeForce NOW, the Membership That Keeps on Delivering

This GFN Thursday rolls out a new reward and games for GeForce NOW members. Whether hunting for hot new releases or rediscovering timeless classics, members can...

24/06/2025

Introducing NVFP4 for Efficient and Accurate Low-Precision Inference

To get the most out of AI, optimizations are critical. When developers think about optimizing AI models for inference, model compression techniques-such as quan...

24/06/2025

HPE and NVIDIA Debut AI Factory Stack to Power Next Industrial Shift

To speed up AI adoption across industries, HPE and NVIDIA today launched new AI factory offerings at HPE Discover in Las Vegas. The new lineup includes everyth...

24/06/2025

NVIDIA and Partners Highlight Next-Generation Robotics, Automation and AI Technologies at Automatica

From the heart of Germany's automotive sector to manufacturing hubs across F...

19/06/2025

Step Inside the Vault: The Borderland' Series Arrives on GeForce NOW

GeForce NOW is throwing open the vault doors to welcome the legendary Borderland series to the cloud. Whether a seasoned Vault Hunter or new to the mayhem of P...

18/06/2025

Plug and Play: Build a G-Assist Plug-In Today

Project G-Assist - available through the NVIDIA App - is an experimental AI assistant that helps tune, control and optimize NVIDIA GeForce RTX systems. NVIDIA&...

17/06/2025

Hexagon Taps NVIDIA Robotics and AI Software to Build and Deploy AEON, a New Humanoid

As a global labor shortage leaves 50 million positions unfilled across industrie...

13/06/2025

NVIDIA and Deutsche Telekom Partner to Advance Germany's Sovereign AI

Industrial AI isn't slowing down. Germany is ready. Following London Tech Week and GTC Paris at VivaTech, NVIDIA founder and CEO Jensen Huang's Europea...

12/06/2025

NVIDIA TensorRT Boosts Stable Diffusion 3.5 Performance on NVIDIA GeForce RTX and RTX PRO GPUs

Generative AI has reshaped how people create, imagine and interact with digital ...

12/06/2025

Turn RTX ON With 40% Off Performance Day Passes

Level up GeForce NOW experiences this summer with 40% off Performance Day Passes. Enjoy 24 hours of premium cloud gaming with RTX ON, delivering low latency and...

11/06/2025

NVIDIA DRIVE Full-Stack Autonomous Vehicle Software Rolls Out

NVIDIA is launching a comprehensive, industry-defining autonomous vehicle (AV) software platform to accelerate large-scale deployment of safe, intelligent trans...

11/06/2025

NVIDIA Research Casts New Light on Scenes With AI-Powered Rendering for Physical AI Development

NVIDIA Research has developed an AI light switch for videos that can turn daytim...

11/06/2025

European Researchers Develop AI-Native Wireless Networks With NVIDIA 6G Research Portfolio

Using NVIDIA platforms, tools and libraries, European telecommunications institu...

11/06/2025

NVIDIA Scores Consecutive Win for End-to-End Autonomous Driving Grand Challenge at CVPR

NVIDIA was today named an Autonomous Grand Challenge winner at the Computer Visi...

11/06/2025

European Robot Makers Adopt NVIDIA Isaac, Omniverse and Halos to Develop Safe, Physical AI-Driven Robot Fleets

In the face of growing labor shortages and need for sustainability, European man...

11/06/2025

Retail Reboot: Major Global Brands Transform End-to-End Operations With NVIDIA

AI is packing and shipping efficiency for the retail and consumer packaged goods (CPG) industries, with a majority of surveyed companies in the space reporting ...

11/06/2025

NVIDIA Brings Physical AI to European Cities With New Blueprint for Smart City AI

Urban populations are expected to double by 2050, which means around 2.5 billion...

11/06/2025

Calling on LLMs: New NVIDIA AI Blueprint Helps Automate Telco Network Configuration

Telecom companies last year spent nearly $295 billion in capital expenditures an...

11/06/2025

European Broadcasting Union and NVIDIA Partner on Sovereign AI to Support Public Broadcasters

In a new effort to advance sovereign AI for European public service media, NVIDI...

11/06/2025

NVIDIA CEO Drops the Blueprint for Europe's AI Boom

At GTC Paris - held alongside VivaTech, Europe's largest tech event - NVIDIA founder and CEO Jensen Huang delivered a clear message: Europe isn't just a...

10/06/2025

The Blue Lion Supercomputer Will Run on NVIDIA Vera Rubin - Here's Why That Matters

Germany's Leibniz Supercomputing Centre, LRZ, is gaining a new supercomputer...

10/06/2025

Clear Skies Ahead: New NVIDIA Earth-2 Generative AI Foundation Model Simulates Global Climate at Kilometer-Scale Resolution

With a more detailed simulation of the Earth's climate, scientists and resea...

10/06/2025

Cisco and NVIDIA Advance Security for Enterprise AI Factories

Cisco and NVIDIA are helping set a new standard for secure, scalable and high-performance enterprise AI. Announced today at the Cisco Live conference in San Di...

09/06/2025

UK Prime Minister, NVIDIA CEO Set the Stage as AI Lights Up Europe

AI isn't waiting. And this week, neither is Europe. At London's Olympia, under a ceiling of steel beams and enveloped by the thrum of startup pitches, ...

08/06/2025

AI Maker, Not an AI Taker': UK Builds Its Vision With NVIDIA Infrastructure

U.K. Prime Minister Keir Starmer's ambition for Britain to be an AI maker, not an AI taker, is becoming a reality at London Tech Week. With NVIDIA's ...

05/06/2025

GeForce NOW Kicks Off a Summer of Gaming With 25 New Titles This June

GeForce NOW is a gamer's ticket to an unforgettable summer of gaming. With 25 titles coming this month and endless ways to play, the summer is going to be e...

04/06/2025

NVIDIA Blackwell Delivers Breakthrough Performance in Latest MLPerf Training Results

NVIDIA is working with companies worldwide to build out AI factories - speeding ...

04/06/2025

How 1X Technologies' Robots Are Learning to Lend a Helping Hand

Humans learn the norms, values and behaviors of society from each other - and Bernt B rnich, founder and CEO of 1X Technologies, thinks robots should learn like...

04/06/2025

NVIDIA RTX Blackwell GPUs Accelerate Professional-Grade Video Editing

4:2:2 cameras - capable of capturing double the color information compared with most standard cameras - are becoming widely available for consumers. At the same...

02/06/2025

Bring Receipts: New NVIDIA AI Blueprint Detects Fraudulent Credit Card Transactions With Precision

Editor's note: This blog, originally published on October 28, 2024, has been...

02/06/2025

Researchers and Students in Trkiye Build AI, Robotics Tools to Boost Disaster Readiness

Since a 7.8-magnitude earthquake hit Syria and T rkiye two years ago - leaving 5...

29/05/2025

The Supercomputer Designed to Accelerate Nobel-Worthy Science

Ready for a front-row seat to the next scientific revolution? That's the idea behind Doudna - a groundbreaking supercomputer announced today at Lawrence Be...

29/05/2025

Run LLMs on AnythingLLM Faster With NVIDIA RTX AI PCs

Large language models (LLMs), trained on datasets with billions of tokens, can generate high-quality content. They're the backbone for many of the most popu...

29/05/2025

RTX on Deck: The GeForce NOW Native App for Steam Deck Is Here

GeForce NOW is supercharging Valve's Steam Deck with a new native app - delivering the high-quality GeForce RTX-powered gameplay members are used to on a po...

28/05/2025

NVIDIA's Bartley Richardson on How Teams of AI Agents Provide Next-Level Automation

Building effective agentic AI systems requires rethinking how technology interac...

27/05/2025

How Dell Technologies Is Building the Engines of AI Factories With NVIDIA Blackwell

Over a century ago, Henry Ford pioneered the mass production of cars and engines...