Sony Pixel Power calrec Sony

NVIDIA Advances Robot Learning and Humanoid Development With New AI and Simulation Tools

06/11/2024

www.1x.tech

Robotics developers can greatly accelerate their work on AI-enabled robots, including humanoids, using new AI and simulation tools and workflows that NVIDIA revealed this week at the Conference for Robot Learning (CoRL) in Munich, Germany.

The lineup includes the general availability of the NVIDIA Isaac Lab robot learning framework; six new humanoid robot learning workflows for Project GR00T, an initiative to accelerate humanoid robot development; and new world-model development tools for video data curation and processing, including the NVIDIA Cosmos tokenizer and NVIDIA NeMo Curator for video processing.

The open-source Cosmos tokenizer provides robotics developers superior visual tokenization by breaking down images and videos into high-quality tokens with exceptionally high compression rates. It runs up to 12x faster than current tokenizers, while NeMo Curator provides video processing curation up to 7x faster than unoptimized pipelines.

Also timed with CoRL, NVIDIA presented 23 papers and nine workshops related to robot learning and released training and workflow guides for developers. Further, Hugging Face and NVIDIA announced they're collaborating to accelerate open-source robotics research with LeRobot, NVIDIA Isaac Lab and NVIDIA Jetson for the developer community.

Accelerating Robot Development With Isaac Lab NVIDIA Isaac Lab is an open-source, robot learning framework built on NVIDIA Omniverse, a platform for developing OpenUSD applications for industrial digitalization and physical AI simulation.

Developers can use Isaac Lab to train robot policies at scale. This open-source unified robot learning framework applies to any embodiment - from humanoids to quadrupeds to collaborative robots - to handle increasingly complex movements and interactions.

Leading commercial robot makers, robotics application developers and robotics research entities around the world are adopting Isaac Lab, including 1X, Agility Robotics, The AI Institute, Berkeley Humanoid, Boston Dynamics, Field AI, Fourier, Galbot, Mentee Robotics, Skild AI, Swiss-Mile, Unitree Robotics and XPENG Robotics.

Project GR00T: Foundations for General-Purpose Humanoid Robots Building advanced humanoids is extremely difficult, demanding multilayer technological and interdisciplinary approaches to make the robots perceive, move and learn skills effectively for human-robot and robot-environment interactions.

Project GR00T is an initiative to develop accelerated libraries, foundation models and data pipelines to accelerate the global humanoid robot developer ecosystem.

Six new Project GR00T workflows provide humanoid developers with blueprints to realize the most challenging humanoid robot capabilities. They include:

GR00T-Gen for building generative AI-powered, OpenUSD-based 3D environments

GR00T-Mimic for robot motion and trajectory generation

GR00T-Dexterity for robot dexterous manipulation

GR00T-Control for whole-body control

GR00T-Mobility for robot locomotion and navigation

GR00T-Perception for multimodal sensing

Humanoid robots are the next wave of embodied AI, said Jim Fan, senior research manager of embodied AI at NVIDIA. NVIDIA research and engineering teams are collaborating across the company and our developer ecosystem to build Project GR00T to help advance the progress and development of global humanoid robot developers.

New Development Tools for World Model Builders Today, robot developers are building world models - AI representations of the world that can predict how objects and environments respond to a robot's actions. Building these world models is incredibly compute- and data-intensive, with models requiring thousands of hours of real-world, curated image or video data.

NVIDIA Cosmos tokenizers provide efficient, high-quality encoding and decoding to simplify the development of these world models. They set a new standard of minimal distortion and temporal instability, enabling high-quality video and image reconstructions.

Providing high-quality compression and up to 12x faster visual reconstruction, the Cosmos tokenizer paves the path for scalable, robust and efficient development of generative applications across a broad spectrum of visual domains.

1X, a humanoid robot company, has updated the 1X World Model Challenge dataset to use the Cosmos tokenizer.

NVIDIA Cosmos tokenizer achieves really high temporal and spatial compression of our data while still retaining visual fidelity, said Eric Jang, vice president of AI at 1X Technologies. This allows us to train world models with long horizon video generation in an even more compute-efficient manner.

Other humanoid and general-purpose robot developers, including XPENG Robotics and Hillbot, are developing with the NVIDIA Cosmos tokenizer to manage high-resolution images and videos.

NeMo Curator now includes a video processing pipeline. This enables robot developers to improve their world-model accuracy by processing large-scale text, image and video data.

Curating video data poses challenges due to its massive size, requiring scalable pipelines and efficient orchestration for load balancing across GPUs. Additionally, models for filtering, captioning and embedding need optimization to maximize throughput.

NeMo Curator overcomes these challenges by streamlining data curation with automatic pipeline orchestration, reducing processing time significantly. It supports linear scaling across multi-node, multi-GPU systems, efficiently handling over 100 petabytes of data. This simplifies AI development, reduces costs and accelerates time to market.

Advancing the Robot Learning Community at CoRL The nearly two dozen research papers the NVIDIA robotics team released with CoRL cover breakthroughs in integrating vision language models for improved environmental understanding and task execution, temporal robot navigati
LINK: https://blogs.nvidia.com/blog/robot-learning-humanoid-development/...
See more stories from nvidia

North America Stories

28/04/2026

Paramount Skydance Will be 49.5% Foreign Owned After WBD Merger

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

KEET PBS Deploys PMVG TechBundle Services To Modernize Operations

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

TAG Video Systems Lens Wins Industry Awards for Cutting T...

TAG Video Systems, the leading IP-native Realtime Media Platform, today announced that Lens, its visual service health interface for broadcast operations, recei...

28/04/2026

AIMS Wins NAB Show 2026 Product of the Year Award for IPM...

Open AV-over-IP Standard Recognized in IT Networking/Infrastructure and Security Category The Alliance for IP Media Solutions (AIMS) today announced that the ...

28/04/2026

VFX History: Slit Scan

VFX History: Slit Scan Graham Quince April 28, 2026 0 Comments How did 2001: A Space Odyssey, Star Wars, Doctor Who and Star Trek: The Next Generation...

28/04/2026

These DaVinci Resolve Effects Will Make You a More Creative Colorist

These DaVinci Resolve Effects Will Make You a More Creative Colorist Kasia Jarco April 28, 2026 0 Comments Creativity in color grading is not about ha...

28/04/2026

A Simple Introduction to Cavalry: Indexed Circle

A Simple Introduction to Cavalry: Indexed Circle Simon Ubsdell April 28, 2026 0 Comments In this new introductory tutorial for Cavalry we're going...

28/04/2026

Rise Upskill - Applications Now Open for Free Global Trai...

Rise, the award-winning advocacy group for gender diversity in the broadcast and media technology sector, is pleased to announce a new global training programme...

28/04/2026

Clear Com Announces New Roles for Brian Grahn and Ben Tur...

Clear-Com has appointed Brian Grahn as Market Outreach Manager of the Americas and Ben Turnwell as Business Development Manager for EMEA live, expanding their ...

28/04/2026

LiveU Steps into the Future at MPTS 2026 with the Introdu...

LiveU is inviting MPTS visitors to step into the companys new Q Era on Stand D32, at The Grand Hall, Olympia, London (May 13-14). The company will showcase its ...

28/04/2026

IBC launches 2026 Innovation Awards to spotlight real-wor...

IBC today announces the launch of the IBC2026 Innovation Awards, with nominations now open for projects, programmes and initiatives that exemplify breakthrough ...

28/04/2026

WNBA to Stream All Preseason Games for Free

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

Nexstar Media Charitable Foundation Sets 30 Days of Giving'

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

Sinclair's Chief Compliance Officer Jeff Lewis to Retire

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

Nielsen Introduces Predictive Sales Lift' Tool

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

Sencore's VB440 Monitoring, Analysis Tool Debuts at NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

Pinterest Makes a Major Push into CTV Advertising

Share Copy link Facebook X Linkedin Bluesky Email...

28/04/2026

Introducing Nx 3-Strip v2 - A Physics-Based Technicolor Reconstruction for DaVinci Resolve

Introducing Nx 3-Strip v2 - A Physics-Based Technicolor Reconstruction for DaVin...

28/04/2026

Into the Omniverse: Manufacturing's Simulation-First Era Has Arrived

Editor's note: This post is part of Into the Omniverse, a series focused on how developers, 3D practitioners, and enterprises can transform their workflows ...

28/04/2026

NVIDIA Launches Nemotron 3 Nano Omni Model, Unifying Vision, Audio and Language for up to 9x More Efficient AI Agents

AI agent systems today juggle separate models for vision, speech and language - ...

27/04/2026

CES Power Acquires Three Ireland-Based Businesses

CES Power, a provider of infrastructure for live events, has announced the acquisition of three Ireland-based businesses: GH Energy Rental Ltd, Event Power, and...

27/04/2026

Fubo to Launch Multiview on Select LG TVs Ahead of 2026 Football Season

FuboTV Inc. has announced it is developing its Multiview feature for the Fubo streaming service on select LG TVs, including 2024, 2025, and newer 4K and 8K mode...

27/04/2026

Shade Raises $14 Million in Funding Round Led by Khosla Ventures

Shade, a file management platform for creative teams, has announced a $14 million funding round led by Khosla Ventures, Construct Capital, and Bling Capital, br...

27/04/2026

AES to Present Immersive Audio Academy 12th Edition on April 30

The Audio Engineering Society (AES) will present the Immersive Audio Academy 12th Edition - Immersive Audio in All Flavors - on April 30, 2026, at 12:00 p.m. ...

27/04/2026

DAZN Launches DAZN48 Creator Program for FIFA World Cup 2026

DAZN has announced DAZN48, a creator program for the FIFA World Cup 2026 that will recruit 48 creators - one representing each of the 48 qualified nations - to ...

27/04/2026

FloSports Acquires Streaming Rights to Four CrossFit Events

FloSports has announced exclusive streaming rights to four CrossFit competitions: Legends Del Mar: CrossFit Semi-Finals, Magic City Games, NorCal Classic, and t...

27/04/2026

NAB 2026: Telestream Pulse Named NAB Show 2026 Product of the Year for Monitoring and Measuring

Telestream has announced that Pulse, its software-defined test and measurement p...

27/04/2026

NAB 2026: Shade Wins 2026 NAB Show Product of the Year Award

Shade has announced it is a Cloud Computing and Storage winner in the 2026 NAB Show Product of the Year Awards. Winners were selected by a panel of industry exp...

27/04/2026

SVG All-Stars: Kelsey Kjeldsen, Senior Director and Video Ads Platform Lead, DTC Products, Tech, and Operations, NBA

Leading the NBA's video-ads platform, this Penn State grad is at the forefro...

27/04/2026

DAZN Extends T100 Triathlon World Tour Rights to Africa

DAZN has expanded its international broadcast rights for the T100 Triathlon World Tour to include Africa. All races from the T100 calendar will be available for...

27/04/2026

NAB 2026: NAB Show Announces 2026 Project and Product of the Year Award Winners

NAB Show has announced the recipients of its 2026 Project of the Year and Product of the Year Awards at a ceremony at the Las Vegas Convention Center. Each wi...

27/04/2026

DIRECTV Launches on Meta Quest Headsets as Sports Season Heats Up

DIRECTV has launched on Meta Quest headsets, becoming the first MVPD to offer live TV through the platform. The timing coincides with the stretch run of the MLB...

27/04/2026

TBL Team Boxing League and MSG Networks Announce Broadcast Partnership

TBL Team Boxing League has announced a broadcast agreement with MSG Networks to air all remaining Season 4 fights live across MSG's television and digital p...

27/04/2026

Audio Exhibitors Showcase New Platforms, Innovative Solutions for Complex Issues

IP integration, interoperability, growth of intercommunications were key concerns for vendors and visitors alike Attendees at the recently concluded 2026 NAB S...

27/04/2026

Behind The Mic: Kenny Beecham to Launch NBA Radio Show on SiriusXM; Mike Tomlin to Join NBC Pregame Show

Behind The Mic provides a roundup of recent news regarding on-air talent, includ...

27/04/2026

On the Show Floor, the Microphone Is Still the Place Where Audio Begins

A pro-audio emphasis, spectrum changes, and on-field audio mark the new products and enhancements to existing offerings Microphones remain the primary point of...

27/04/2026

NAB Show 2026 In Review: Our Complete Collection of Video Interviews with Industry Thought Leaders

The Sports Video Group team was all over the NAB Show floor out in Las Vegas las...

27/04/2026

Study: Local TV Political Ad Spend to Top $4 billion in 2026

Share Copy link Facebook X Linkedin Bluesky Email...

27/04/2026

Nippon TV's In-House Proprietary AI Solution AiDi Wins Product of the Year Award at NAB 2026

Nippon TV's In-House Proprietary AI Solution AiDi Wins Product of the Year A...

27/04/2026

Outpost Introduces Unlimited Collaboration Model for Review and Approval Workflows

Outpost Introduces Unlimited Collaboration Model for Review and Approval Workflo...

27/04/2026

Ikegami Announces VFE-P07D Monocular OLED Viewfinder with Tiltable 3.5-inch LCD Monitor

Ikegami Announces VFE-P07D Monocular OLED Viewfinder with Tiltable 3.5-inch LCD ...

27/04/2026

Custom Consoles Completes Large Module-R MCR Desk and Med...

Custom Consoles announces the completion of a large Module-R desk and MediaWall monitor display mount for an expanded master control room at Gravity Medias West...

27/04/2026

Other World Computing Launches OWC Express 4M2 Ultra - Thunderbolt 5 Four-Slot NVMe M.2 SSD Enclosure

Other World Computing Launches OWC Express 4M2 Ultra - Thunderbolt 5 Four-Slot N...

27/04/2026

Netflix announces El sobrino, the new film by Damin Szifron starring Leonardo Sbaraglia

Back to All News Netflix announces El sobrino, the new film by Dami n Szifron s...

26/04/2026

Director Yoon Jong-bin Returns with The Generals' (WT), An Incisive Chronicle of a Second-in-Command

Back to All News Director Yoon Jong-bin Returns with The Generals' (WT), A...

26/04/2026

'Nine Queens,' Starring Alvaro Morte and Patrick Criado, Starts Production

Back to All News Nine Queens, Starring Alvaro Morte and Patrick Criado, Starts ...

25/04/2026

Mediagenix Sweeps 2026 NAB Awards With Wins for Product of the Year and Best of Show for Scheduling Optimization

Mediagenix Sweeps 2026 NAB Awards With Wins for Product of the Year and Best of ...

25/04/2026

SCHOEPS Microphones Announces Desert Island Boom Set for NAB 2026

SCHOEPS Microphones Announces Desert Island Boom Set for NAB 2026 Brie Clayton April 24, 2026 0 Comments Compact modular set ideal for location sound ...

25/04/2026

Berklee Africana Studies Hosts Gospel Extravaganza 2026

Berklee Africana Studies Hosts Gospel Extravaganza 2026 The Signature Series event will honor three new inductees to the Berklee Gospel Hall of Fame and celeb...