Sony Pixel Power calrec Sony

NVIDIA Advances Physical AI at CVPR With Largest Indoor Synthetic Dataset

17/06/2024

NVIDIA contributed the largest ever indoor synthetic dataset to the Computer Vision and Pattern Recognition (CVPR) conference's annual AI City Challenge - helping researchers and developers advance the development of solutions for smart cities and industrial automation.

The challenge, garnering over 700 teams from nearly 50 countries, tasks participants to develop AI models to enhance operational efficiency in physical settings, such as retail and warehouse environments, and intelligent traffic systems.

Teams tested their models on the datasets that were generated using NVIDIA Omniverse, a platform of application programming interfaces (APIs), software development kits (SDKs) and services that enable developers to build Universal Scene Description (OpenUSD)-based applications and workflows.

Creating and Simulating Digital Twins for Large Spaces In large indoor spaces like factories and warehouses, daily activities involve a steady stream of people, small vehicles and future autonomous robots. Developers need solutions that can observe and measure activities, optimize operational efficiency, and prioritize human safety in complex, large-scale settings.

Researchers are addressing that need with computer vision models that can perceive and understand the physical world. It can be used in applications like multi-camera tracking, in which a model tracks multiple entities within a given environment.

To ensure their accuracy, the models must be trained on large, ground-truth datasets for a variety of real-world scenarios. But collecting that data can be a challenging, time-consuming and costly process.

AI researchers are turning to physically based simulations - such as digital twins of the physical world - to enhance AI simulation and training. These virtual environments can help generate synthetic data used to train AI models. Simulation also provides a way to run a multitude of what-if scenarios in a safe environment while addressing privacy and AI bias issues.

Creating synthetic data is important for AI training because it offers a large, scalable, and expandable amount of data. Teams can generate a diverse set of training data by changing many parameters including lighting, object locations, textures and colors.

Building Synthetic Datasets for the AI City Challenge This year's AI City Challenge consists of five computer vision challenge tracks that span traffic management to worker safety.

NVIDIA contributed datasets for the first track, Multi-Camera Person Tracking, which saw the highest participation, with over 400 teams. The challenge used a benchmark and the largest synthetic dataset of its kind - comprising 212 hours of 1080p videos at 30 frames per second spanning 90 scenes across six virtual environments, including a warehouse, retail store and hospital.

Created in Omniverse, these scenes simulated nearly 1,000 cameras and featured around 2,500 digital human characters. It also provided a way for the researchers to generate data of the right size and fidelity to achieve the desired outcomes.

The benchmarks were created using Omniverse Replicator in NVIDIA Isaac Sim, a reference application that enables developers to design, simulate and train AI for robots, smart spaces or autonomous machines in physically based virtual environments built on NVIDIA Omniverse.

Omniverse Replicator, an SDK for building synthetic data generation pipelines, automated many manual tasks involved in generating quality synthetic data, including domain randomization, camera placement and calibration, character movement, and semantic labeling of data and ground-truth for benchmarking.

Ten institutions and organizations are collaborating with NVIDIA for the AI City Challenge:

Australian National University, Australia

Emirates Center for Mobility Research, UAE

Indian Institute of Technology Kanpur, India

Iowa State University, U.S.

Johns Hopkins University, U.S.

National Yung-Ming Chiao-Tung University, Taiwan

Santa Clara University, U.S.

The United Arab Emirates University, UAE

University at Albany - SUNY, U.S.

Woven by Toyota, Japan

Driving the Future of Generative Physical AI Researchers and companies around the world are developing infrastructure automation and robots powered by physical AI - which are models that can understand instructions and autonomously perform complex tasks in the real world.

Generative physical AI uses reinforcement learning in simulated environments, where it perceives the world using accurately simulated sensors, performs actions grounded by laws of physics, and receives feedback to reason about the next set of actions.

Developers can tap into developer SDKs and APIs, such as the NVIDIA Metropolis developer stack - which includes a multi-camera tracking reference workflow - to add enhanced perception capabilities for factories, warehouses and retail operations. And with the latest release of NVIDIA Isaac Sim, developers can supercharge robotics workflows by simulating and training AI-based robots in physically based virtual spaces before real-world deployment.

Researchers and developers are also combining high-fidelity, physics-based simulation with advanced AI to bridge the gap between simulated training and real-world application. This helps ensure that synthetic training environments closely mimic real-world conditions for more seamless robot deployment.

NVIDIA is taking the accuracy and scale of simulations further with the recently announced NVIDIA Omniverse Cloud Sensor RTX, a set of microservices that enable physically accurate sensor simulation to accelerate the development of fully autonomous machines.

This technology will allow autonomous systems, whether a factory, vehicle or robot, to gather essential data to effectively perceive, navigate and interact with the real world. Using these microservices, developers can run large-scale te
LINK: https://blogs.nvidia.com/blog/ai-city-challenge-omniverse-cvpr/...
See more stories from nvidia

North America Stories

20/04/2026

Google Cloud Embraces the Rise of Agentic Production

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Creators Go All in on AI, Niche Content

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

NBC Sports' Jon Miller: Broadcast Is Having a Moment'

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Beyond the Lift and Shift': Cloud Migration's New Mandate

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Virtual Production Finds Its Footing

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Corporate Creators: All Companies Are Media Companies Now

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

IABM Rebrands as the International Association of MediaTech

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

CBS Detroit Debuts New AR/VR Technology-Driven Studio

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Fox Sports Taps Appear X Platform for Remote Production

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

CueScript and Lighting Design Group Expand Customer Oppor...

CueScript and Lighting Design Group Expand Customer Opportunities Through New Partnership Find both companies at 2026 NAB Show in CueScript Booth # C 4720 ...

20/04/2026

Layercake Deepens Bitmovin Integration to Power End-to-En...

[Sydney, NSW, 20 April 2026] - Layercake, the company behind the intelligent media orchestration platform Streamcake, today announced the formalisation of its i...

20/04/2026

FOX Sports selects Appear X Platform for next-generation...

Deployment spans FOX Sports' REMI infrastructure, IP production for a major global soccer event, and its Jewel Events production systems Appear, a global l...

20/04/2026

Pro Sound Effects Launches the Industry's First and Only Native Sound Effects Integration for Avid Media Composer at NAB 2026

Pro Sound Effects Launches the Industry's First and Only Native Sound Effect...

20/04/2026

SBE Elevates Fred Willard to SBE Fellow

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Blackmagic Design Announces Blackmagic Camera for iOS 3.3 Update

Blackmagic Design Announces Blackmagic Camera for iOS 3.3 Update Brie Clayton April 20, 2026 0 Comments New update adds camera control and monitoring ...

20/04/2026

Maxon Announces Free Tools and Mobile Expansion of ZBrush and Cinema 4D

Maxon Announces Free Tools and Mobile Expansion of ZBrush and Cinema 4D Brie Clayton April 20, 2026 0 Comments Cinema 4D brings professional 3D workfl...

20/04/2026

Vizrt AI Keyer kills the green screen and creates virtual scenes in any environment

Vizrt AI Keyer kills the green screen and creates virtual scenes in any environm...

20/04/2026

Ikegami Announces VFE-P07D Monocular OLED Viewfinder

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

EVS Launches Choreon Robotic Control Solution

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Ross Video Showcases End-To-End Production Ecosystem at 2026 Nab Show

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

Heidi Steffen to Become President of TitanTV

Share Copy link Facebook X Linkedin Bluesky Email...

20/04/2026

NVIDIA and Partners Showcase the Future of AI-Driven Manufacturing at Hannover Messe 2026

Manufacturing is at an inflection point. Across every major industrial economy, ...

20/04/2026

Autonomous AI at Scale: Adobe Agents Unlock Breakthrough Creative Intelligence With NVIDIA and WPP

AI agents are transforming how work gets done across all industries, acceleratin...

19/04/2026

NAB Show 2026 Is Here! Follow All of our Live Coverage!

Blackmagic Design has announced the ATEM 4 M/E Constellation IP and ATEM 4 M/E Constellation IP Plus, two SMPTE 2110-native live production switchers. The ATEM ...

19/04/2026

Live From NAB 2026: Grass Valley CEO Jon Wilson on AMPPs Explosive Growth, Hybrid Workflows, and Whats New at the Show

Grass Valley is finding the right balance between its hardware heritage with an ...

19/04/2026

Live From NAB 2026: Oracles Kip Schauer on Why OCI Is Doubling Down on Media, Sports, and Broadcast

Oracle's strategy rests on the foundational strengths of Oracle Cloud Infras...

19/04/2026

Live From NAB 2026: Program Productions Jess Kowatch on Whats New with ProCrewz and the Impact of AI on Crewing

Program Productions, the live sports production industry's leading crewer, i...

19/04/2026

Live From NAB 2026: Aggrekos Joe Scionti on Powering the Super Bowl, PGA Championship, and the Road to the FIFA World Cup

At the 2026 NAB Show in Las Vegas, SVG sat down with Joe Scionti, Account Manage...

19/04/2026

NAB 2026: Evertz to highlight evertz.io XChange for live event management and market switching

Evertz (Booth N817) is set to present new services within its evertz.io platform...

19/04/2026

NAB 2026: Evertz to showcase IPMX-certified NUCLEUS and MMA platforms for AV and ST 2110 integration

Evertz (Booth N817) will showcase its IPMX-certified NUCLEUS platform alongside ...

19/04/2026

NAB 2026: Evertz to showcase ENX media core for hybrid SDI and IP facilities

Evertz (Booth N817) is set to showcase ENX at NAB 2026, a media core platform designed to support hybrid SDI and IP infrastructures in production facilities and...

19/04/2026

NAB 2026: Evertz introduces Studer VistaVUE Touch for broadcast control

Evertz (Booth N817) will introduce Studer VistaVUE Touch at NAB 2026, a control surface designed to integrate audio, video and control workflows within a custom...

19/04/2026

NAB 2026: Evertz highlights X-CALIBER high-density encoding platform for media transport

Evertz (Booth N817) will highlight X-CALIBER at NAB 2026, an encoding and decodi...

19/04/2026

NAB 2026: Cobalt Digital introduces blueCORE standalone processors for SDI and ST 2110 workflows

Cobalt Digital (Booth N1340) will introduce the blueCORE family of standalone si...

19/04/2026

NAB 2026: Chyron and Asport to demonstrate AI-driven end-to-end sports production and distribution workflows

Chyron and Asport (Booth N2441) will demonstrate an integrated sports video work...

19/04/2026

NAB 2026: MediaKind outlines growth of Multiview deployments as Charter rollout expands in North America

MediaKind (Booth W1743) provided an update on its Multiview deployments at NAB S...

19/04/2026

NAB 2026: Calrec and Grass Valley announce partnership to integrate ImPulseV with AMPP platform

Calrec (Booth C6907) and Grass Valley (Booth C2408) announced a long-term broadc...

19/04/2026

NAB 2026: Oracle and partners to demo MoQ-based streaming ecosystem

Oracle is bringing a multi-partner demonstration of Media over QUIC (MoQ)-based live streaming to NAB Show 2026, showcasing how independent systems from multipl...

19/04/2026

NAB 2026: Encompass Digital Media and Oracle Cloud Infrastructure expand partnership for cloud-native broadcast operations

Encompass Digital Media announced an expanded partnership with Oracle Cloud Infr...

19/04/2026

SportsTechBuzz at NAB 2026, Day 1: Live Reports From the Show Floor in Vegas

The NAB Show is in full swing, and the SVG and SVG Europe editorial teams are chasing down the hottest stories from all over the Las Vegas Convention Center. He...

19/04/2026

NAB 2026: Blackmagic Design Announces ATEM 4 M/E Constellation IP Switchers

Blackmagic Design has announced the ATEM 4 M/E Constellation IP and ATEM 4 M/E Constellation IP Plus, two SMPTE 2110-native live production switchers. The ATEM ...

19/04/2026

TV and Radio HQ Moves Closer to the Conversation

Share Copy link Facebook X Linkedin Bluesky Email...

19/04/2026

Expanded Creator Lab Seeds Digital Synergies

Share Copy link Facebook X Linkedin Bluesky Email...

19/04/2026

Why Streamers Are Seizing the Now

Share Copy link Facebook X Linkedin Bluesky Email...

19/04/2026

Make Every Dollar Count on Set

Share Copy link Facebook X Linkedin Bluesky Email...

19/04/2026

Tech Transforms the Live Sports Playbook

Share Copy link Facebook X Linkedin Bluesky Email...

19/04/2026

Amagi Managed Services Modernizes Broadcasting Operations...

Amagi, the agentic industry cloud platform for unified broadcast, streaming, and monetization, today announced that AccuWeather , the most trusted source of wea...

19/04/2026

Calrec and Grass Valley unlock exceptional choice and fle...

Calrec (Booth:C6907) and Grass Valley (Booth: C2408) are today announcing a long-term broadcast audio technology partnership at NAB Show 2026. The companies are...

19/04/2026

Ikegami Announces VFE-P07D Monocular OLED Viewfinder with...

Ikegami announces a further expansion to its range of on-camera viewfinders. Scheduled for introduction on Ikegamis Central Hall booth C3819 at the April 19th -...