Sony Pixel Power calrec Sony

NVIDIA Advances Robot Learning and Humanoid Development With New AI and Simulation Tools

06/11/2024

www.1x.tech

Robotics developers can greatly accelerate their work on AI-enabled robots, including humanoids, using new AI and simulation tools and workflows that NVIDIA revealed this week at the Conference for Robot Learning (CoRL) in Munich, Germany.

The lineup includes the general availability of the NVIDIA Isaac Lab robot learning framework; six new humanoid robot learning workflows for Project GR00T, an initiative to accelerate humanoid robot development; and new world-model development tools for video data curation and processing, including the NVIDIA Cosmos tokenizer and NVIDIA NeMo Curator for video processing.

The open-source Cosmos tokenizer provides robotics developers superior visual tokenization by breaking down images and videos into high-quality tokens with exceptionally high compression rates. It runs up to 12x faster than current tokenizers, while NeMo Curator provides video processing curation up to 7x faster than unoptimized pipelines.

Also timed with CoRL, NVIDIA presented 23 papers and nine workshops related to robot learning and released training and workflow guides for developers. Further, Hugging Face and NVIDIA announced they're collaborating to accelerate open-source robotics research with LeRobot, NVIDIA Isaac Lab and NVIDIA Jetson for the developer community.

Accelerating Robot Development With Isaac Lab NVIDIA Isaac Lab is an open-source, robot learning framework built on NVIDIA Omniverse, a platform for developing OpenUSD applications for industrial digitalization and physical AI simulation.

Developers can use Isaac Lab to train robot policies at scale. This open-source unified robot learning framework applies to any embodiment - from humanoids to quadrupeds to collaborative robots - to handle increasingly complex movements and interactions.

Leading commercial robot makers, robotics application developers and robotics research entities around the world are adopting Isaac Lab, including 1X, Agility Robotics, The AI Institute, Berkeley Humanoid, Boston Dynamics, Field AI, Fourier, Galbot, Mentee Robotics, Skild AI, Swiss-Mile, Unitree Robotics and XPENG Robotics.

Project GR00T: Foundations for General-Purpose Humanoid Robots Building advanced humanoids is extremely difficult, demanding multilayer technological and interdisciplinary approaches to make the robots perceive, move and learn skills effectively for human-robot and robot-environment interactions.

Project GR00T is an initiative to develop accelerated libraries, foundation models and data pipelines to accelerate the global humanoid robot developer ecosystem.

Six new Project GR00T workflows provide humanoid developers with blueprints to realize the most challenging humanoid robot capabilities. They include:

GR00T-Gen for building generative AI-powered, OpenUSD-based 3D environments

GR00T-Mimic for robot motion and trajectory generation

GR00T-Dexterity for robot dexterous manipulation

GR00T-Control for whole-body control

GR00T-Mobility for robot locomotion and navigation

GR00T-Perception for multimodal sensing

Humanoid robots are the next wave of embodied AI, said Jim Fan, senior research manager of embodied AI at NVIDIA. NVIDIA research and engineering teams are collaborating across the company and our developer ecosystem to build Project GR00T to help advance the progress and development of global humanoid robot developers.

New Development Tools for World Model Builders Today, robot developers are building world models - AI representations of the world that can predict how objects and environments respond to a robot's actions. Building these world models is incredibly compute- and data-intensive, with models requiring thousands of hours of real-world, curated image or video data.

NVIDIA Cosmos tokenizers provide efficient, high-quality encoding and decoding to simplify the development of these world models. They set a new standard of minimal distortion and temporal instability, enabling high-quality video and image reconstructions.

Providing high-quality compression and up to 12x faster visual reconstruction, the Cosmos tokenizer paves the path for scalable, robust and efficient development of generative applications across a broad spectrum of visual domains.

1X, a humanoid robot company, has updated the 1X World Model Challenge dataset to use the Cosmos tokenizer.

NVIDIA Cosmos tokenizer achieves really high temporal and spatial compression of our data while still retaining visual fidelity, said Eric Jang, vice president of AI at 1X Technologies. This allows us to train world models with long horizon video generation in an even more compute-efficient manner.

Other humanoid and general-purpose robot developers, including XPENG Robotics and Hillbot, are developing with the NVIDIA Cosmos tokenizer to manage high-resolution images and videos.

NeMo Curator now includes a video processing pipeline. This enables robot developers to improve their world-model accuracy by processing large-scale text, image and video data.

Curating video data poses challenges due to its massive size, requiring scalable pipelines and efficient orchestration for load balancing across GPUs. Additionally, models for filtering, captioning and embedding need optimization to maximize throughput.

NeMo Curator overcomes these challenges by streamlining data curation with automatic pipeline orchestration, reducing processing time significantly. It supports linear scaling across multi-node, multi-GPU systems, efficiently handling over 100 petabytes of data. This simplifies AI development, reduces costs and accelerates time to market.

Advancing the Robot Learning Community at CoRL The nearly two dozen research papers the NVIDIA robotics team released with CoRL cover breakthroughs in integrating vision language models for improved environmental understanding and task execution, temporal robot navigati
LINK: https://blogs.nvidia.com/blog/robot-learning-humanoid-development/...
See more stories from nvidia

North America Stories

01/07/2026

Adder Technology Names Neil Hillier as CEO

Share Copy link Facebook X Linkedin Bluesky Email...

01/07/2026

IBCAP Opens New Anti-Piracy Lab in Denver

Share Copy link Facebook X Linkedin Bluesky Email...

01/07/2026

FCC Plans to Auction 160 MHZ of Mid-Band Spectrum

Share Copy link Facebook X Linkedin Bluesky Email...

01/07/2026

CBS Miami Launches 'Hope 4 Venezuela' Relief Effort

Share Copy link Facebook X Linkedin Bluesky Email...

01/07/2026

Chyron Launches the All-New Chyron Academy: A Reimagined, Hands-On Learning Experience for Live Broadcast Production

Chyron Launches the All-New Chyron Academy: A Reimagined, Hands-On Learning Expe...

01/07/2026

Amplium Captures Kawasaki Brave Thunders Game with Blackmagic URSA Cine Immersive

Amplium Captures Kawasaki Brave Thunders Game with Blackmagic URSA Cine Immersiv...

01/07/2026

Boris FX Optics Expands Plugin Support to Apple Photos, Capture One, and Affinity Photo

Boris FX Optics Expands Plugin Support to Apple Photos, Capture One, and Affinit...

30/06/2026

CazTVs 12 ENG Teams Across North America Keep Brazilian Fans on Top of World Cup

As Brazil's only way for fans to see all 104 matches, YouTube channel proves the power of digital...

30/06/2026

Telemundo, Peacock See More Record Setting World Cup Audiences

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

FOR-A America Adds Two Execs to U.S. Sales Team

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

A3SA Disputes Weigel Assertions that NextGen TV Threatens EAS

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

MainStreaming Selected by ITV to Support Delivery of ITVX...

MainStreaming, the award-winning and innovative Edge Video Delivery Network, today announced that it has been selected by ITV to support the delivery of ITVX, I...

30/06/2026

Chyron Launches New Chyron Academy

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

Clear-Com Upgrades Communication Systems for Jeopardy and...

When Wheel of Fortune and Jeopardy! needed to upgrade their wireless communications system, they turned to Clear-Com FreeSpeak wireless for their iconic televi...

30/06/2026

Supreme Court Gives Trump Tight Control over Independent Regulators

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

Rocket Lab to Acquire Iridium in $8 Billion Deal

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

Kyocera AVX Releases New Web-Based Antenna Integration Tool

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

YouTube Shorts Get a Makeover

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

Rise Announces 2026 Worldwide Mentoring Cohorts

Share Copy link Facebook X Linkedin Bluesky Email...

30/06/2026

Other World Computing Launches New Atlas Core Line with 256GB CFExpress 4.0 Type B Memory Card

Other World Computing Launches New Atlas Core Line with 256GB CFExpress 4.0 Type...

30/06/2026

DaVinci Resolve Studio Used for Taketoshi Sado's Perfume Cold Sleep -25 years Document-

DaVinci Resolve Studio Used for Taketoshi Sado's Perfume Cold Sleep -25 year...

30/06/2026

NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science

Life sciences has entered an era of computational scale, and for more than a dec...

30/06/2026

FOR-A America Expands U.S. Sales Team to Accelerate Growth of Software-Defined Solutions

Fernando Cruz and Jaz Wray Join as Regional Sales Managers; Bringing Extensive S...

30/06/2026

How NVIDIA's Inference Software Stack Powers the Lowest Token Cost

As organizations move from AI pilots to production AI factories, infrastructure decisions have shifted from peak chip specifications to cost per token: how many...

30/06/2026

Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning

Editor's note: This post is part of Into the Omniverse, a series focused on ...

30/06/2026

June 29, 2026

Scripps Research scientists demonstrate a faster, cheaper route to making critical drugs using common table sugar New method illustrates how to build a tough ch...

29/06/2026

Op-Ed: Why the 2026 World Cup Is Redefining the Economics of Live Sports Production

By Andy Rayner, CTO, Appear The 2026 FIFA World Cup is the largest football tou...

29/06/2026

Study: Esports Plays Major Role in Gen Z Media Habits, Purchasing Behavior

A new multi-country study from ESL FACEIT Group, Hero Esports, and Niko Partners estimates that 400 million Gen Z consumers regularly engage with esports, under...

29/06/2026

ESPN Sets Multiplatform Plans for America 250 Celebration

ESPN will mark America's 250th anniversary with a series of content initiatives across its linear, digital, and streaming platforms, including a special edi...

29/06/2026

OBSBOT Named Official Camera and Webcam Partner of Esports World Cup 2026

The Esports Foundation has named OBSBOT the Official Camera and Webcam Partner for the Esports World Cup 2026, bringing the company's AI-powered imaging tec...

29/06/2026

Insight Productions Launches Insight Storm, 53-Foot Esports Broadcast Truck

Insight Productions has launched Insight Storm, a 53-foot mobile broadcast unit designed specifically for esports production, live entertainment, and digital-fi...

29/06/2026

Gravity Media Delivers Global Distribution, Streaming Services for World Economic Forum in Dalian

Gravity Media once again provided broadcast, streaming, and content-distribution...

29/06/2026

Wimbledon Introduces AI-Powered Fan Features, Modernized Digital Platforms for 2026 Championships

The All England Lawn Tennis Club and IBM have introduced new and enhanced digita...

29/06/2026

Comcast to Spin Off NBCUniversal, Sky

Share Copy link Facebook X Linkedin Bluesky Email...

29/06/2026

Supreme Court Rules Trump Can Fire FTC Commissioner Without Cause

Share Copy link Facebook X Linkedin Bluesky Email...

29/06/2026

CVP Launches Warranty plus to Protect Professional Equipm...

CVP, one of Europes leading suppliers of professional video and broadcast solutions, has announced the launch of CVP Warranty , a new extended warranty programm...

29/06/2026

TitanTV Hires Mark Hadley as Technology Specialist

Share Copy link Facebook X Linkedin Bluesky Email...

29/06/2026

Fox Sports Delivers Another FIFA Mens World Cup Audience Record

Share Copy link Facebook X Linkedin Bluesky Email...

29/06/2026

Claude Meets Blackwell Ultra: Anthropic's Models Now Run on NVIDIA GB300 in Azure

Anthropic's Claude models in Microsoft Foundry - hosted on Microsoft Azure a...

29/06/2026

Open Models, Closed Environments: Palantir Brings Secure AI to US Agencies With NVIDIA Nemotron

Showcasing the importance of open source innovation in American AI, Palantir'...

28/06/2026

Freedman Labs Releases PrepMyMedia and ViewMyAttic for Post Production Professionals - 50% off for COWs!

Freedman Labs Releases PrepMyMedia and ViewMyAttic for Post Production Professio...

27/06/2026

Through Their Lens: What Cinematographer Amy Vincent Saw at the 2026 Directors Lab

There's no doubt that you've seen the world through Amy Vincent's ey...

27/06/2026

Apogee CRAS Symphony Mkii Education Feature Blog

Why CRAS Upgraded to Symphony I/O MK II When an audio school runs studios all day, every day, gear doesn't just need to sound good , it needs to survive rea...

27/06/2026

MultiDyne Acquires the Assets of MRMC

Share Copy link Facebook X Linkedin Bluesky Email...

27/06/2026

Spectrum Intelligence Ventures Launches Latis

Share Copy link Facebook X Linkedin Bluesky Email...

27/06/2026

Krotos Video to Sound Plugin Now Available for Adobe Premiere Pro

Krotos Video to Sound Plugin Now Available for Adobe Premiere Pro Brie Clayton June 26, 2026 0 Comments Editors can analyze footage, generate synchron...

27/06/2026

Mirai Media Elevates Digital and Broadcast Productions with Blackmagic Design

Mirai Media Elevates Digital and Broadcast Productions with Blackmagic Design Brie Clayton June 26, 2026 0 Comments Studio uses Ultimatte 12 HD and Po...

27/06/2026

Lutra Cafe & Bakery Opens At American Tobacco Campus

DURHAM, N.C. - JUNE 26, 2026 - Lutra Cafe & Bakery has opened its first brick-and-mortar location at American Tobacco Campus after owner Chris McLaurin operated...

26/06/2026

SVG GameDay, Ep. 21: Minnesota Vikings Allan Wertheimer - Large-Scale Shows in Minny

In-venue and creative video staffers at the professional and collegiate level ha...

26/06/2026

Strike Fighter League Announces Second Online Tournament, Set for July 25 in Las Vegas

Strike Fighter League (SFL), a professional air combat digital sport combining f...