Sony Pixel Power calrec Sony

NVIDIA Advances Robot Learning and Humanoid Development With New AI and Simulation Tools

06/11/2024

www.1x.tech

Robotics developers can greatly accelerate their work on AI-enabled robots, including humanoids, using new AI and simulation tools and workflows that NVIDIA revealed this week at the Conference for Robot Learning (CoRL) in Munich, Germany.

The lineup includes the general availability of the NVIDIA Isaac Lab robot learning framework; six new humanoid robot learning workflows for Project GR00T, an initiative to accelerate humanoid robot development; and new world-model development tools for video data curation and processing, including the NVIDIA Cosmos tokenizer and NVIDIA NeMo Curator for video processing.

The open-source Cosmos tokenizer provides robotics developers superior visual tokenization by breaking down images and videos into high-quality tokens with exceptionally high compression rates. It runs up to 12x faster than current tokenizers, while NeMo Curator provides video processing curation up to 7x faster than unoptimized pipelines.

Also timed with CoRL, NVIDIA presented 23 papers and nine workshops related to robot learning and released training and workflow guides for developers. Further, Hugging Face and NVIDIA announced they're collaborating to accelerate open-source robotics research with LeRobot, NVIDIA Isaac Lab and NVIDIA Jetson for the developer community.

Accelerating Robot Development With Isaac Lab NVIDIA Isaac Lab is an open-source, robot learning framework built on NVIDIA Omniverse, a platform for developing OpenUSD applications for industrial digitalization and physical AI simulation.

Developers can use Isaac Lab to train robot policies at scale. This open-source unified robot learning framework applies to any embodiment - from humanoids to quadrupeds to collaborative robots - to handle increasingly complex movements and interactions.

Leading commercial robot makers, robotics application developers and robotics research entities around the world are adopting Isaac Lab, including 1X, Agility Robotics, The AI Institute, Berkeley Humanoid, Boston Dynamics, Field AI, Fourier, Galbot, Mentee Robotics, Skild AI, Swiss-Mile, Unitree Robotics and XPENG Robotics.

Project GR00T: Foundations for General-Purpose Humanoid Robots Building advanced humanoids is extremely difficult, demanding multilayer technological and interdisciplinary approaches to make the robots perceive, move and learn skills effectively for human-robot and robot-environment interactions.

Project GR00T is an initiative to develop accelerated libraries, foundation models and data pipelines to accelerate the global humanoid robot developer ecosystem.

Six new Project GR00T workflows provide humanoid developers with blueprints to realize the most challenging humanoid robot capabilities. They include:

GR00T-Gen for building generative AI-powered, OpenUSD-based 3D environments

GR00T-Mimic for robot motion and trajectory generation

GR00T-Dexterity for robot dexterous manipulation

GR00T-Control for whole-body control

GR00T-Mobility for robot locomotion and navigation

GR00T-Perception for multimodal sensing

Humanoid robots are the next wave of embodied AI, said Jim Fan, senior research manager of embodied AI at NVIDIA. NVIDIA research and engineering teams are collaborating across the company and our developer ecosystem to build Project GR00T to help advance the progress and development of global humanoid robot developers.

New Development Tools for World Model Builders Today, robot developers are building world models - AI representations of the world that can predict how objects and environments respond to a robot's actions. Building these world models is incredibly compute- and data-intensive, with models requiring thousands of hours of real-world, curated image or video data.

NVIDIA Cosmos tokenizers provide efficient, high-quality encoding and decoding to simplify the development of these world models. They set a new standard of minimal distortion and temporal instability, enabling high-quality video and image reconstructions.

Providing high-quality compression and up to 12x faster visual reconstruction, the Cosmos tokenizer paves the path for scalable, robust and efficient development of generative applications across a broad spectrum of visual domains.

1X, a humanoid robot company, has updated the 1X World Model Challenge dataset to use the Cosmos tokenizer.

NVIDIA Cosmos tokenizer achieves really high temporal and spatial compression of our data while still retaining visual fidelity, said Eric Jang, vice president of AI at 1X Technologies. This allows us to train world models with long horizon video generation in an even more compute-efficient manner.

Other humanoid and general-purpose robot developers, including XPENG Robotics and Hillbot, are developing with the NVIDIA Cosmos tokenizer to manage high-resolution images and videos.

NeMo Curator now includes a video processing pipeline. This enables robot developers to improve their world-model accuracy by processing large-scale text, image and video data.

Curating video data poses challenges due to its massive size, requiring scalable pipelines and efficient orchestration for load balancing across GPUs. Additionally, models for filtering, captioning and embedding need optimization to maximize throughput.

NeMo Curator overcomes these challenges by streamlining data curation with automatic pipeline orchestration, reducing processing time significantly. It supports linear scaling across multi-node, multi-GPU systems, efficiently handling over 100 petabytes of data. This simplifies AI development, reduces costs and accelerates time to market.

Advancing the Robot Learning Community at CoRL The nearly two dozen research papers the NVIDIA robotics team released with CoRL cover breakthroughs in integrating vision language models for improved environmental understanding and task execution, temporal robot navigati
LINK: https://blogs.nvidia.com/blog/robot-learning-humanoid-development/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

02/04/2026

Scripps Completes Sale of WRTV to Circle City Broadcasting

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

GoVertical! AiDi Powers Real-Time 9:16 Autocropping for I...

Already deployed extensively by NBC Sports, FOR-A Corporation will demonstrate GoVertical! AiDi, the real-time 9:16 autocropping feature of viztrick AiDi, durin...

02/04/2026

Elite Media Technologies Selects Interra Systems BATON Fi...

Interra Systems, a provider of end-to-end quality assurance solutions for the digital media industry, announced that Elite Media Technologies has selected its B...

02/04/2026

TDF Expands Broadcast Channel Lineup with Harmonic

Harmonic's Media Processing Solutions Maximize Bandwidth Efficiency for Terrestrial Broadcast Delivery Harmonic (NASDAQ: HLIT) today announced that TDF, a...

02/04/2026

FOR-A's Software-Defined, AI-Powered Development Advances...

NBC Sports Deploys viztrick AiDi to Stream Live Events in 9:16 Mobile-First Formats with Auto Tracking, Development Signals Strategic Shift for FOR-A Long reco...

02/04/2026

Evergent showcases innovations in sports streaming and mo...

Evergent will showcase new innovations in subscriber lifecycle management and monetization at NAB Show 2026 (Las Vegas, April 18 22), including: New advances i...

02/04/2026

Binghamton University Strengthens Student Run Productions...

Riedel Communications is proud to be part of Binghamton University, State University of New York, Athletics' milestone year, celebrating the university'...

02/04/2026

Techex and Encompass Launch Industry-Leading Cloud-Based...

Encompass Digital Media and Techex have today announced new, fully managed, cloud-native Master Control services designed to meet the growing operational demand...

02/04/2026

Winning in the new media economy - Avid debuts fully avai...

Avid today announced it will showcase new innovations designed to help media companies win in the new media economy at NAB Show 2026 (April 18 22, Las Vegas Co...

02/04/2026

PlayBox Neo reinforces MIMO Tech with new Playout capabil...

PlayBox Neo helps AIS PLAY kick-off premier football content direct to fans PlayBox Neo has provided MIMO Tech with a brand-new major installation to extend it...

02/04/2026

Globo transitions primary distribution to SRT over IP wit...

Globo has transitioned its primary content distribution to Secure Reliable Transport over a fully IP-based managed backbone using Synamedia's Quortex PowerV...

02/04/2026

Nexstar Says Pausing Tegna Merger Creates 'Impossible' Challenges

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

FCC Launches Efforts to Strengthen U.S. Drone Ecosystem

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

WAPA+ to Launch on Dish, DishLatino, Sling TV and Sling Freestream

Share Copy link Facebook X Linkedin Bluesky Email...

02/04/2026

Student Spotlight: Al-Fadl Salem

Student Spotlight: Al-Fadl Salem The Danish singer recently performed for the queen of Denmark. April 1, 2026 By Editorial Staff Image by Junia Morrow Wh...

02/04/2026

Taku Hirano's Career Is Defined by Identity

Taku Hirano's Career Is Defined by Identity Whether he's performing, composing, teaching, or developing instruments, the do-it-all percussionist sees ...

02/04/2026

Design Perspective Intelligent Hybrid Software Platforms to Survive the Evolutionary Avalanche

By Lance Maurer, CEO Image generated by AI Engineering is supposed to be fun....

02/04/2026

Continuing to connect with Young Ireland: 2FM Announces Brand-New Daytime Schedule

2FM Breakfast to extend on weekday mornings from 6am to 10am Doireann Garrihy m...

02/04/2026

RT NEWS ANNOUNCES BARRY LENIHAN AS NEW POLITICAL CORRESPONDENT

RT News & Current Affairs is pleased to announce the appointment of RT Radio 1 reporter, Barry Lenihan, as Political Correspondent. Barry has reported across...

02/04/2026

Press Start on April: GeForce NOW Brings 10 Games to the Cloud

No joke - GFN Thursday is skipping the tricks and heading straight into the games. April kicks off with ten new titles, bringing fresh adventures to GeForce NOW...

01/04/2026

SVG New Sponsor Spotlight: Flowstate AI's Sahil Shah on Transforming Video Content with Intelligent AI Agents

As sports media organization continue to seek out new ways to streamline their p...

01/04/2026

SVG GFX Forum 2026: Sessions Now Available to Watch on SVG PLAY

The SVG GFX Forum hit New York City earlier this month for a day packed with sessions focused on the creative strategy and technology behind today's cutting...

01/04/2026

From Buenos Aires to Mexico City, EQUAL Days Bring Latin America Together for Women in Audio

This year, Spotify celebrates the five-year anniversary of EQUAL, our global pro...

01/04/2026

FourFingers announce Tape Splice Pro plug-in

Analogue-style tape splicing in the digital domain In this era of digital recording and multiple layers of Undo, it seems that the fading art of tape splici...

01/04/2026

Zero G introduce Morphology Evolved

Latest release introduces new Orbita Engine Zero G's latest release marks the start of a new series of libraries, as well as introducing an all-new engi...

01/04/2026

Warm Audio introduce the WA-8TRX

Until now, one format has largely been left behind Warm Audio's extensive product range includes modern-day recreations of all manner of sought-after s...

01/04/2026

The Crow Hill Company announce Crystal Pianos

A piano with glass vessels for strings! The Crow Hill Company's recently released Gong Piano offered a refreshing new take on piano libraries, harnessin...

01/04/2026

ESSENCE RS from Aim Audio

Remote Streaming Studio Condenser Aim Audio have just revealed their latest creation, the ESSENCE RS Remote Streaming Studio Condenser, which becomes the wo...

01/04/2026

Call for NFVF funding applications to attend Film Festivals and Markets taking place from 08 - 31 May 2026

The National Film and Video Foundation (NFVF) is pleased to announce that the ca...

01/04/2026

AgileTV powers Liwest's next-generation TV experience with the launch of next IPTV platform in Austria

Bilbao, April 1st, 2026 - AgileTV, a leading provider of end-to-end TV technolog...

01/04/2026

Green Hippo Debuts Hands on Hippotizer Media Server Train...

Green Hippo is excited to announce the launch of its new Hippotizer Media Server training courses at Pixel Academy, a purpose built AV learning hub combining ha...

01/04/2026

TAG Video Systems and Oracle Cloud Infrastructure Partner...

TAG Video Systems, a global leader in IP-native broadcast monitoring, multiviewing, and quality control, today announced a collaboration with Oracle Cloud Infra...

01/04/2026

Professional Wireless Systems PWS Takes on Intercom and R...

Professional Wireless Systems (PWS), a leading provider of wireless audio solutions and RF management, was on site at the Caesars Superdome in New Orleans, wher...

01/04/2026

AgileTV powers Liwest next generation TV experience with...

AgileTV, a leading provider of end-to-end TV technology solutions, has deployed next , the new IPTV platform of the Austrian telco LIWEST, marking the first st...

01/04/2026

LTN and Ateme partner to deliver integrated video process...

LTN, a leader in fully managed IP video transport, and Ateme, a global leader in video compression and delivery solutions, today announced a collaboration integ...

01/04/2026

Adobe Unveils Powerful New Innovations for Creative Pros in Adobe Illustrator

Adobe Unveils Powerful New Innovations for Creative Pros in Adobe Illustrator Deepa Subramaniam April 1, 2026 0 Comments I'm excited to share that...

01/04/2026

Boland Communications Introduces QD4K315HDR10 QD-OLED Series Monitors for Live Production, Film, Post, and Broadcast

Boland Communications Introduces QD4K315HDR10 QD-OLED Series Monitors for Live P...

01/04/2026

2026 NAB Show Exhibitor Insight: Evertz

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Judge Blocks Order Barring NPR and PBS From Funding

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Nikon to Sell Mark Roberts Motion Control

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Mediagenix Showcases Semantic Intelligence-Powered Title Management, Schedule Optimization, and Personalization at NAB 2026

Mediagenix Showcases Semantic Intelligence-Powered Title Management, Schedule Op...

01/04/2026

FCC Approves WJAX-TV License Transfer to Cox

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Scripps Sports Ink Deal for Ion to Air 2026 Teal Rising Cup

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

UK Group Companies Unveil NAB Show Plans

Share Copy link Facebook X Linkedin Bluesky Email...

01/04/2026

Victoria Mont Brings the Multi-Hyphenate Mindset to Career Jam 2026

Victoria Mon t Brings the Multi-Hyphenate Mindset to Career Jam 2026 The Grammy-winning singer, songwriter, and producer shared how versatility and self-inves...

01/04/2026

UKTV announces expanded remit for Jonathan Newman and appoints David Swetman as Director of Content Partnerships & Sales

UKTV today announces that Jonathan Newman has formally stepped into the role of ...