
NVIDIA today announced at Microsoft Build new AI performance optimizations and integrations for Windows that help deliver maximum performance on NVIDIA GeForce RTX AI PCs and NVIDIA RTX workstations.
Large language models (LLMs) power some of the most exciting new use cases in generative AI and now run up to 3x faster with ONNX Runtime (ORT) and DirectML using the new NVIDIA R555 Game Ready Driver. ORT and DirectML are high-performance tools used to run AI models locally on Windows PCs.
WebNN, an application programming interface for web developers to deploy AI models, is now accelerated with RTX via DirectML, enabling web apps to incorporate fast, AI-powered capabilities. And PyTorch will support DirectML execution backends, enabling Windows developers to train and infer complex AI models on Windows natively. NVIDIA and Microsoft are collaborating to scale performance on RTX GPUs.
These advancements build on NVIDIA's world-leading AI platform, which accelerates more than 500 applications and games on over 100 million RTX AI PCs and workstations worldwide.
RTX AI PCs - Enhanced AI for Gamers, Creators and Developers NVIDIA introduced the first PC GPUs with dedicated AI acceleration, the GeForce RTX 20 Series with Tensor Cores, along with the first widely adopted AI model to run on Windows, NVIDIA DLSS, in 2018. Its latest GPUs offer up to 1,300 trillion operations per second of dedicated AI performance.
In the coming months, Copilot+ PCs equipped with new power-efficient systems-on-a-chip and RTX GPUs will be released, giving gamers, creators, enthusiasts and developers increased performance to tackle demanding local AI workloads, along with Microsoft's new Copilot+ features.
For gamers on RTX AI PCs, NVIDIA DLSS boosts frame rates by up to 4x, while NVIDIA ACE brings game characters to life with AI-driven dialogue, animation and speech.
For content creators, RTX powers AI-assisted production workflows in apps like Adobe Premiere, Blackmagic Design DaVinci Resolve and Blender to automate tedious tasks and streamline workflows. From 3D denoising and accelerated rendering to text-to-image and video generation, these tools empower artists to bring their visions to life.
For game modders, NVIDIA RTX Remix, built on the NVIDIA Omniverse platform, provides AI-accelerated tools to create RTX remasters of classic PC games. It makes it easier than ever to capture game assets, enhance materials with generative AI tools and incorporate full ray tracing.
For livestreamers, the NVIDIA Broadcast application delivers high-quality AI-powered background subtraction and noise removal, while NVIDIA RTX Video provides AI-powered upscaling and auto-high-dynamic range to enhance streamed video quality.
Enhancing productivity, LLMs powered by RTX GPUs execute AI assistants and copilots faster, and can process multiple requests simultaneously.
And RTX AI PCs allow developers to build and fine-tune AI models directly on their devices using NVIDIA's AI developer tools, which include NVIDIA AI Workbench, NVIDIA cuDNN and CUDA on Windows Subsystem for Linux. Developers also have access to RTX-accelerated AI frameworks and software development kits like NVIDIA TensorRT, NVIDIA Maxine and RTX Video.
The combination of AI capabilities and performance deliver enhanced experiences for gamers, creators and developers.
Faster LLMs and New Capabilities for Web Developers Microsoft recently released the generative AI extension for ORT, a cross-platform library for AI inference. The extension adds support for optimization techniques like quantization for LLMs like Phi-3, Llama 3, Gemma and Mistral. ORT supports different execution providers for inferencing via various software and hardware stacks, including DirectML.
ORT with the DirectML backend offers Windows AI developers a quick path to develop AI capabilities, with stability and production-grade support for the broad Windows PC ecosystem. NVIDIA optimizations for the generative AI extension for ORT, available now in R555 Game Ready, Studio and NVIDIA RTX Enterprise Drivers, help developers get up to 3x faster performance on RTX compared to previous drivers.
Inference performance for three LLMs using ONNX Runtime and the DirectML execution provider with the latest R555 GeForce driver compared to the previous R550 driver. INSEQ=2000 representative of document summarization workloads. All data captured with GeForce RTX 4090 GPU using batch size 1. The generative AI extension support for int4 quantization, plus the NVIDIA optimizations, result in up to 3x faster performance for LLMs. Developers can unlock the full capabilities of RTX hardware with the new R555 driver, bringing better AI experiences to consumers, faster. It includes:
Support for DQ-GEMM metacommand to handle INT4 weight-only quantization for LLMs
New RMSNorm normalization methods for Llama 2, Llama 3, Mistral and Phi-3 models
Group and multi-query attention mechanisms, and sliding window attention to support Mistral
In-place KV updates to improve attention performance
Support for GEMM of non-multiple-of-8 tensors to improve context phase performance
Additionally, NVIDIA has optimized AI workflows within WebNN to deliver the powerful performance of RTX GPUs directly within browsers. The WebNN standard helps web app developers accelerate deep learning models with on-device AI accelerators, like Tensor Cores.
Now available in developer preview, WebNN uses DirectML and ORT Web, a Javascript library for in-browser model execution, to make AI applications more accessible across multiple platforms. With this acceleration, popular models like Stable Diffusion, SD Turbo and Whisper run up to 4x faster on WebNN compared to WebGPU and are now available for developers to use. Microsoft Build attendees can learn more about developing on RTX in the Accelerating development on Windows PCs w
North America Stories
18/06/2026
Ratings Roundup is a rundown of recent rating news and is derived from press rel...
18/06/2026
PTZOptics has unveiled new Visual Reasoning demonstrations at InfoComm 2026 (Boo...
18/06/2026
IBC2026 will take place at the RAI Amsterdam from September 11-14, bringing toge...
18/06/2026
InfoComm 2026 held its first-ever Media Day on June 17, providing journalists an...
18/06/2026
FOR-A America has announced a Trade Agreements Act (TAA)-compliant LED display solution combining Alfalite's Litepix LED displays and Brompton Technology...
18/06/2026
In-venue and creative video staffers at the professional and collegiate level have one major thing in common: the intensity and attention to detail ramps up dur...
18/06/2026
The International Federation of American Football (IFAF) and TMRW Sports have an...
18/06/2026
AJA Video Systems has unveiled Io Xpand, a Thunderbolt 5-enabled PCIe expansion ...
18/06/2026
ESPN has announced its coverage plans for the 30th anniversary of the WNBA's...
18/06/2026
FOX Sports' Big Noon Kickoff will broadcast live from Wembley Stadium in Lon...
18/06/2026
InfoComm 2026 opened on Wednesday at the Las Vegas Convention Center, bringing t...
18/06/2026
As media companies look to deliver more live, VOD, and snackable sports content ...
18/06/2026
Top L-R: Take Me Home, The Lake
Bottom L-R: TheyDream, Union County
Free Summer Screening Series Announced
Screenings for the Local Utah Community at...
18/06/2026
The average daily TV screen time in May was 3 hours and 36 minutes, marking a clear decrease of 15 minutes compared to April. This trend proved to be much stron...
18/06/2026
Integration embeds Gracenote content intelligence, including contextual segments...
18/06/2026
AJA Video Systems unveiled Io Xpand, a high-performance Thunderbolt 5-enabled PCIe expansion chassis for AJA KONA and Corvid video and audio I/O cards. As dema...
18/06/2026
New NVRT-driven workflow enables on-demand traffic mirroring and guided troubleshooting for AV teams, integrators, and support organizations ahead of InfoComm 2...
18/06/2026
At InfoComm 2026, Mavis announced a major update to Mavis Studio, its live production app for the iPad, with new features designed to make professional AV produ...
18/06/2026
Transaction Positions Harmonic as a Pure-Play Broadband Company
Harmonic Inc. (NASDAQ: HLIT), the worldwide leader in virtualized broadband solutions, today a...
18/06/2026
ACT Entertainment and NETGEAR have announced a strategic partnership that establishes verified interoperability between NETGEAR Switches and key technologies fr...
18/06/2026
IBC2026 is set to bring the global media, entertainment and technology community together at the RAI Amsterdam from 11 14 September, enabling industry players f...
18/06/2026
Pliant Technologies is proud to introduce CrewCom Flex , the world's first frameless matrix intercom system and the only complete matrix IP-based intercom s...
18/06/2026
The XR Sports Alliance (XRSA) has announced that a new cohort of members has joined the strategic initiative: ActionStreamer, Antigravity, Creative Artists Agen...
18/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
18/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
18/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
18/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
18/06/2026
Linda May Han Oh Receives Guggenheim Fellowship The Berklee professor, bassist, and composer will use the fellowship to debut Dreams of Knowing, an interdisci...
18/06/2026
Play favorite titles from popular game libraries, keep progress synced and jump ...
18/06/2026
The digital era gave the advertising and marketing industry speed; the AI era is giving it autonomous operations.
For companies building next-generation techn...
17/06/2026
EVS has announced it has received the EcoVadis Gold Medal for sustainability performance, ranking among the top 5% of companies globally in the Technology/Mid-S...
17/06/2026
Chyron has released Weather 2.4, an update to its weather suite for broadcasters and meteorologists. The release focuses on enhancements to the DataFlow module,...
17/06/2026
VSiN, The Sports Betting Network, has announced the launch of Best Bets TV, a free ad-supported streaming TV (FAST) channel. The 24/7 channel is currently avail...
17/06/2026
DAZN's Team Whistle and Snap Inc. have announced a creator program centered ...
17/06/2026
LiveU is providing video transmission technology for broadcasters, production companies, and public safety agencies across North America's busy Summer of So...
17/06/2026
The National Academy of Television Arts and Sciences (NATAS) has announced that Laurens Grant and Jacob Ullman have joined its Board of Directors. Chief of Staf...
17/06/2026
Akta, the AI-First SaaS video platform for modern broadcast and streaming operations, today announced that its video platform is now generally available on Orac...
17/06/2026
This recent graduate from Houston found inspiration in technical directing and now eyes a future career in sports production...
17/06/2026
Audio-Technica (booth C7959) arrives at InfoComm 2026 in Las Vegas with a slate ...
17/06/2026
SNS has published a guide addressing growing demand for AI-powered video indexing, transcription, facial recognition, and searchable metadata across media libra...
17/06/2026
Providius has announced Providius Direct, a workflow for investigating network i...
17/06/2026
NEP Group has announced the commercial availability of NEP Platform, a software ...
17/06/2026
A new white paper examining Secure Reliable Transport (SRT) and Reliable Interne...
17/06/2026
New research from subscription bundling platform Bango finds that younger sports fans are increasingly consuming sport through highlights, clips, and social med...
17/06/2026
SMPTE has announced that its complete Standards catalog is now freely available to the global media technology community, including all published SMPTE Standard...
17/06/2026
Harmonic has completed the sale of its Video Business to MediaKind for $145 mill...
17/06/2026
Omaha Productions will produce the 2026 World Series of Poker (WSOP) in Las Vega...
17/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
17/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
17/06/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...