Sony Pixel Power calrec Sony

NVIDIA Advances Robot Learning and Humanoid Development With New AI and Simulation Tools

06/11/2024

www.1x.tech

Robotics developers can greatly accelerate their work on AI-enabled robots, including humanoids, using new AI and simulation tools and workflows that NVIDIA revealed this week at the Conference for Robot Learning (CoRL) in Munich, Germany.

The lineup includes the general availability of the NVIDIA Isaac Lab robot learning framework; six new humanoid robot learning workflows for Project GR00T, an initiative to accelerate humanoid robot development; and new world-model development tools for video data curation and processing, including the NVIDIA Cosmos tokenizer and NVIDIA NeMo Curator for video processing.

The open-source Cosmos tokenizer provides robotics developers superior visual tokenization by breaking down images and videos into high-quality tokens with exceptionally high compression rates. It runs up to 12x faster than current tokenizers, while NeMo Curator provides video processing curation up to 7x faster than unoptimized pipelines.

Also timed with CoRL, NVIDIA presented 23 papers and nine workshops related to robot learning and released training and workflow guides for developers. Further, Hugging Face and NVIDIA announced they're collaborating to accelerate open-source robotics research with LeRobot, NVIDIA Isaac Lab and NVIDIA Jetson for the developer community.

Accelerating Robot Development With Isaac Lab NVIDIA Isaac Lab is an open-source, robot learning framework built on NVIDIA Omniverse, a platform for developing OpenUSD applications for industrial digitalization and physical AI simulation.

Developers can use Isaac Lab to train robot policies at scale. This open-source unified robot learning framework applies to any embodiment - from humanoids to quadrupeds to collaborative robots - to handle increasingly complex movements and interactions.

Leading commercial robot makers, robotics application developers and robotics research entities around the world are adopting Isaac Lab, including 1X, Agility Robotics, The AI Institute, Berkeley Humanoid, Boston Dynamics, Field AI, Fourier, Galbot, Mentee Robotics, Skild AI, Swiss-Mile, Unitree Robotics and XPENG Robotics.

Project GR00T: Foundations for General-Purpose Humanoid Robots Building advanced humanoids is extremely difficult, demanding multilayer technological and interdisciplinary approaches to make the robots perceive, move and learn skills effectively for human-robot and robot-environment interactions.

Project GR00T is an initiative to develop accelerated libraries, foundation models and data pipelines to accelerate the global humanoid robot developer ecosystem.

Six new Project GR00T workflows provide humanoid developers with blueprints to realize the most challenging humanoid robot capabilities. They include:

GR00T-Gen for building generative AI-powered, OpenUSD-based 3D environments

GR00T-Mimic for robot motion and trajectory generation

GR00T-Dexterity for robot dexterous manipulation

GR00T-Control for whole-body control

GR00T-Mobility for robot locomotion and navigation

GR00T-Perception for multimodal sensing

Humanoid robots are the next wave of embodied AI, said Jim Fan, senior research manager of embodied AI at NVIDIA. NVIDIA research and engineering teams are collaborating across the company and our developer ecosystem to build Project GR00T to help advance the progress and development of global humanoid robot developers.

New Development Tools for World Model Builders Today, robot developers are building world models - AI representations of the world that can predict how objects and environments respond to a robot's actions. Building these world models is incredibly compute- and data-intensive, with models requiring thousands of hours of real-world, curated image or video data.

NVIDIA Cosmos tokenizers provide efficient, high-quality encoding and decoding to simplify the development of these world models. They set a new standard of minimal distortion and temporal instability, enabling high-quality video and image reconstructions.

Providing high-quality compression and up to 12x faster visual reconstruction, the Cosmos tokenizer paves the path for scalable, robust and efficient development of generative applications across a broad spectrum of visual domains.

1X, a humanoid robot company, has updated the 1X World Model Challenge dataset to use the Cosmos tokenizer.

NVIDIA Cosmos tokenizer achieves really high temporal and spatial compression of our data while still retaining visual fidelity, said Eric Jang, vice president of AI at 1X Technologies. This allows us to train world models with long horizon video generation in an even more compute-efficient manner.

Other humanoid and general-purpose robot developers, including XPENG Robotics and Hillbot, are developing with the NVIDIA Cosmos tokenizer to manage high-resolution images and videos.

NeMo Curator now includes a video processing pipeline. This enables robot developers to improve their world-model accuracy by processing large-scale text, image and video data.

Curating video data poses challenges due to its massive size, requiring scalable pipelines and efficient orchestration for load balancing across GPUs. Additionally, models for filtering, captioning and embedding need optimization to maximize throughput.

NeMo Curator overcomes these challenges by streamlining data curation with automatic pipeline orchestration, reducing processing time significantly. It supports linear scaling across multi-node, multi-GPU systems, efficiently handling over 100 petabytes of data. This simplifies AI development, reduces costs and accelerates time to market.

Advancing the Robot Learning Community at CoRL The nearly two dozen research papers the NVIDIA robotics team released with CoRL cover breakthroughs in integrating vision language models for improved environmental understanding and task execution, temporal robot navigati
LINK: https://blogs.nvidia.com/blog/robot-learning-humanoid-development/...
See more stories from nvidia

North America Stories

14/04/2026

LTN Appoints Mark Romano, Edward Cox to Leadership Positions

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Grass Valley Adds Telestream Vantage, Pulse and UP to AMPP

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

XenData Announces Backup, Archive and Cloud-Connect for LucidLink

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Thomas Riedel Acquires ARRI

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Shotoku Introduces the World to Aura P2 PTZ Prompter Pann...

Shotoku Introduces the World to Aura P2 PTZ Prompter Panner at NAB 2026 New system removes PTZ pan restrictions for teleprompter-based productions Shotoku US...

14/04/2026

Telestream and Grass Valley Connect Live and File-Based W...

Integration of Vantage, Pulse, and Telestream UP with Grass Valley AMPP Ecosystem enables scalable, interoperable workflows spanning live production and file-ba...

14/04/2026

Dalet Appoints Brian Doheny as President and Chief Revenu...

Enterprise growth leader to scale Dalet's next phase of innovation and global expansion New York, NY April 14, 2026 Dalet, a leading technology and ser...

14/04/2026

Berklee Celebrates Prince's Legacy in Two-Night Signature Series Event

Berklee Celebrates Prince's Legacy in Two-Night Signature Series Event Directed by Tia Fuller, the Prince Project (April 16-17) brings together more than ...

14/04/2026

Arooj Aftab Is Anything but Predictable

Arooj Aftab Is Anything but Predictable The singular artist explores the juxtaposition of grief and joy, dark and light, in her distinctive sound. April 14, ...

14/04/2026

Appear Expands X Platform from Core to Edge at NAB Show 2...

Appear launches include XM estate management and new X Platform processing enhancements to add density for next-generation hybrid & IP workflows, X5 is also now...

14/04/2026

Synamedia turns OTT content into TikTok-style feeds with...

Addressing the needs of a new generation's viewing habits, Synamedia launches GO Shorts. The AI-powered module turns existing catalogues into TikTok-style ...

14/04/2026

NAB 2026 - Vubiquity and Eluvio Showcase Streaming Soluti...

Vubiquity, an Amdocs company and global leader in technology-led media services, will be showcasing a new end-to-end streaming solution in collaboration with El...

14/04/2026

LiveU Announces Expanded Collaboration with Sony at NAB S...

LiveU today announced a significant expansion of its collaboration with Sony Corporation, introducing integrated support for Sony's file-based workflow solu...

14/04/2026

BBC World Service TV selects Open Broadcast Systems for I...

Open Broadcast Systems (https://www.obe.tv/) has announced that BBC World Service has selected its decoders for IP Television distribution. The high-quality, lo...

14/04/2026

Blackmagic Design Announces DaVinci Resolve 21

Blackmagic Design Announces DaVinci Resolve 21 Brie Clayton April 14, 2026 0 Comments Major update adds new Photo page bringing Hollywood's most a...

14/04/2026

Sinclair's WTOV Taps Brightline for Lighting Upgrade

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Grass Valley Showcases Alliance Ecosystem at 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

FCC Selects New Lead Administrator for U.S. Cyber Trust Mark Program

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Gray Media Names Jim Hays GM of WTHI

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

NAB Blasts CTA in FCC Sports Probe Comments

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Wowza to Showcase AI-Powered Video Workflows and Emerging...

Wowza will return to NAB Show 2026 with a set of live demonstrations focused on how video infrastructure is evolving for a new generation of AI-powered and oper...

14/04/2026

Stegawave Debuts Real-Time Forensic Watermarking to Tackle Piracy in Live Sports Streaming

Stegawave Debuts Real-Time Forensic Watermarking to Tackle Piracy in Live Sports...

14/04/2026

Living in Boston: A Guide for Incoming Boston Conservatory Students

Living in Boston: A Guide for Incoming Boston Conservatory Students From navigating the T to balancing school with professional gigs, a current student shar...

14/04/2026

Just What Is Genre These Days, Anyway?

Just What Is Genre These Days, Anyway? Understanding the business and art of genre-bending in 2026. April 13, 2026 By Bryan Parys Illustration by Jack Fla...

14/04/2026

Lenora Helm Hammonds Is Turning Passion Into Plan A

Lenora Helm Hammonds Is Turning Passion Into Plan A The dean of the Professional Education Division has seen the industry from all sides. Now shes bringing it...

14/04/2026

How Michelle Zalabak Found Her Dream Career in Music and Finance

How Michelle Zalabak Found Her Dream Career in Music and Finance The Warner Music Group deal analysis manager helps determine what artist catalogs are worth a...

13/04/2026

Jnger Audio Joins EBU ADM Implementers Group as Founding Member

Telos Alliance has announced that J nger Audio has joined the EBU ADM Implementers Group (ADM-IG) as a founding member. The group is focused on advancing ADM an...

13/04/2026

NAB 2026: Grass Valley to Showcase Alliance Partner Ecosystem

Grass Valley will demonstrate its Alliance Partner ecosystem at NAB Show 2026 (Booth C2408, Central Hall, April 19-22), showing AMPP integrations across live pr...

13/04/2026

NAB 2026: Media Links to Demonstrate IP Transport Solutions

Media Links will exhibit at NAB Show 2026 (Booth W2033), demonstrating IP transport solutions for live production including hitless protection technology, Xscen...

13/04/2026

NBC Sports Partners with Overtime for OT7 Football League and Navy All-American Bowl

NBC Sports has announced a programming, distribution, and sales partnership with...

13/04/2026

FloSports Promotes Jayar Donlan from COO to President

FloSports has promoted Chief Operating Officer Jayar Donlan to President, effective immediately. In his new role, Donlan will lead the company's commercial,...

13/04/2026

MASV Case Study: PanCam Pictures Uses MASV for Remote Post-Production at Senior Bowl 2026

PanCam Pictures, the documentary production company founded by Paul Camarata, us...

13/04/2026

NAB 2026: Mimir to Showcase Cloud Production Platform

Mimir will exhibit at NAB Show 2026 (North Hall, Booth N2850), demonstrating its cloud-native media production platform with new capabilities including Mimir Cu...

13/04/2026

NAB 2026: BBright Adds RIST Protocol Support to IP Gateway

BBright has announced that its IP Gateway now supports the Reliable Internet Stream Transport (RIST) protocol. The addition will be introduced at NAB Show 2026 ...

13/04/2026

Net Insight Awarded ESA NAVISP Development Project for PNT Technology

Net Insight has been awarded a development project through the European Space Agency's Navigation Innovation and Support Program (NAVISP), with co-funding f...

13/04/2026

NAB 2026: intoPIX to Showcase JPEG XS, IPMX, and SMPTE 2110 Solutions

intoPIX will exhibit at NAB Show 2026, marking the company's 20th anniversary. The company will demonstrate its JPEG XS compression portfolio and IPMX-appro...

13/04/2026

Inside the Launch of BravesVision: How Braves, Raycom Sports Pulled Off One of the Most Ambitious Efforts in Regional-Sports-Media History

Starting from scratch, the team built an in-house content platform comprising ga...

13/04/2026

NAB 2026: AI Will Make Its Presence Felt in Audio Offerings, Presentations

Here's a look at some of the new products and updates, along with audio-centric conferences, that attendees will find next week at the show When the 2026 N...

13/04/2026

NAB 2026: Avid to Demonstrate Integrated Newsroom Capabilities

Avid will launch new integrated newsroom capabilities for Avid for News at NAB Show 2026 (Booth N2226, April 18-22), demonstrating how Avid Content Core connect...

13/04/2026

NAB 2026: Synamedia Launches Cloud-Controlled Edge Playout Version of Quortex PowerVu

Synamedia has announced a new version of Quortex PowerVu, an IP-native, software...

13/04/2026

NAB 2026: Mediaproxy Adds AI Brand and Advertisement Tracking to LogServer

Mediaproxy has developed a suite of AI-powered tools for brand and advertisement tracking, integrated into its LogServer compliance logging and analysis platfor...

13/04/2026

NAB 2026: Disguise to Demonstrate Media Server and Software Integrations

Disguise will demonstrate its media servers and software at NAB Show 2026, appearing across five partner booths in Central Hall: MRMC, B&H, Planar, CarbonBlack,...

13/04/2026

NAB 2026: OpenDrives Introduces Edge Hybrid Cloud-Edge Performance Accelerator

OpenDrives is introducing OpenDrives Edge at NAB Show 2026, a hybrid cloud-edge performance accelerator for distributed video and rich media workflows. The prod...

13/04/2026

ESPN Returns to The Shed for 2026 WNBA Draft, Expanding Camera Arsenal and Deepening Fan Coverage

The show will deploy 18 cameras across two sets and the draft floor, including a...

13/04/2026

When Missiles Move at 5X the Speed of Sound, Timing Is Everything

L3Harris is accelerating the development of infrared payloads for Space Development Agency's Tranche 2 Tracking Layer, to help meet urgent national defense ...

13/04/2026

US Army Selects L3Harris for Next-Generation Night-Vision System

By leveraging cutting-edge unfilmed Gen III image intensifier technology, NOVA delivers unmatched clarity, range, and reliability in low-light environments - en...

13/04/2026

Harvey Arnold Represents the Best of Broadcast Engineering

Share Copy link Facebook X Linkedin Bluesky Email...

13/04/2026

Ross Video and HighField AI to Deliver AI-Assisted Graphics Creation

Share Copy link Facebook X Linkedin Bluesky Email...

13/04/2026

Disguise to Showcase Cutting-Edge Experience Tech for Bro...

Explore new Disguise plugins, including Sony's VP integration; Listen to panels across partner booths at Sony and B&H Disguise, the company powering everyt...

13/04/2026

TAG Video Systems Joins MXL Interoperability Initiative t...

TAG Video Systems, the leading IP-native Realtime Media Platform, has announced its participation in the Media Exchange Layer (MXL) interop initiative. TAG has ...