Sony Pixel Power calrec Sony

NVIDIA Advances Robot Learning and Humanoid Development With New AI and Simulation Tools

06/11/2024

www.1x.tech

Robotics developers can greatly accelerate their work on AI-enabled robots, including humanoids, using new AI and simulation tools and workflows that NVIDIA revealed this week at the Conference for Robot Learning (CoRL) in Munich, Germany.

The lineup includes the general availability of the NVIDIA Isaac Lab robot learning framework; six new humanoid robot learning workflows for Project GR00T, an initiative to accelerate humanoid robot development; and new world-model development tools for video data curation and processing, including the NVIDIA Cosmos tokenizer and NVIDIA NeMo Curator for video processing.

The open-source Cosmos tokenizer provides robotics developers superior visual tokenization by breaking down images and videos into high-quality tokens with exceptionally high compression rates. It runs up to 12x faster than current tokenizers, while NeMo Curator provides video processing curation up to 7x faster than unoptimized pipelines.

Also timed with CoRL, NVIDIA presented 23 papers and nine workshops related to robot learning and released training and workflow guides for developers. Further, Hugging Face and NVIDIA announced they're collaborating to accelerate open-source robotics research with LeRobot, NVIDIA Isaac Lab and NVIDIA Jetson for the developer community.

Accelerating Robot Development With Isaac Lab NVIDIA Isaac Lab is an open-source, robot learning framework built on NVIDIA Omniverse, a platform for developing OpenUSD applications for industrial digitalization and physical AI simulation.

Developers can use Isaac Lab to train robot policies at scale. This open-source unified robot learning framework applies to any embodiment - from humanoids to quadrupeds to collaborative robots - to handle increasingly complex movements and interactions.

Leading commercial robot makers, robotics application developers and robotics research entities around the world are adopting Isaac Lab, including 1X, Agility Robotics, The AI Institute, Berkeley Humanoid, Boston Dynamics, Field AI, Fourier, Galbot, Mentee Robotics, Skild AI, Swiss-Mile, Unitree Robotics and XPENG Robotics.

Project GR00T: Foundations for General-Purpose Humanoid Robots Building advanced humanoids is extremely difficult, demanding multilayer technological and interdisciplinary approaches to make the robots perceive, move and learn skills effectively for human-robot and robot-environment interactions.

Project GR00T is an initiative to develop accelerated libraries, foundation models and data pipelines to accelerate the global humanoid robot developer ecosystem.

Six new Project GR00T workflows provide humanoid developers with blueprints to realize the most challenging humanoid robot capabilities. They include:

GR00T-Gen for building generative AI-powered, OpenUSD-based 3D environments

GR00T-Mimic for robot motion and trajectory generation

GR00T-Dexterity for robot dexterous manipulation

GR00T-Control for whole-body control

GR00T-Mobility for robot locomotion and navigation

GR00T-Perception for multimodal sensing

Humanoid robots are the next wave of embodied AI, said Jim Fan, senior research manager of embodied AI at NVIDIA. NVIDIA research and engineering teams are collaborating across the company and our developer ecosystem to build Project GR00T to help advance the progress and development of global humanoid robot developers.

New Development Tools for World Model Builders Today, robot developers are building world models - AI representations of the world that can predict how objects and environments respond to a robot's actions. Building these world models is incredibly compute- and data-intensive, with models requiring thousands of hours of real-world, curated image or video data.

NVIDIA Cosmos tokenizers provide efficient, high-quality encoding and decoding to simplify the development of these world models. They set a new standard of minimal distortion and temporal instability, enabling high-quality video and image reconstructions.

Providing high-quality compression and up to 12x faster visual reconstruction, the Cosmos tokenizer paves the path for scalable, robust and efficient development of generative applications across a broad spectrum of visual domains.

1X, a humanoid robot company, has updated the 1X World Model Challenge dataset to use the Cosmos tokenizer.

NVIDIA Cosmos tokenizer achieves really high temporal and spatial compression of our data while still retaining visual fidelity, said Eric Jang, vice president of AI at 1X Technologies. This allows us to train world models with long horizon video generation in an even more compute-efficient manner.

Other humanoid and general-purpose robot developers, including XPENG Robotics and Hillbot, are developing with the NVIDIA Cosmos tokenizer to manage high-resolution images and videos.

NeMo Curator now includes a video processing pipeline. This enables robot developers to improve their world-model accuracy by processing large-scale text, image and video data.

Curating video data poses challenges due to its massive size, requiring scalable pipelines and efficient orchestration for load balancing across GPUs. Additionally, models for filtering, captioning and embedding need optimization to maximize throughput.

NeMo Curator overcomes these challenges by streamlining data curation with automatic pipeline orchestration, reducing processing time significantly. It supports linear scaling across multi-node, multi-GPU systems, efficiently handling over 100 petabytes of data. This simplifies AI development, reduces costs and accelerates time to market.

Advancing the Robot Learning Community at CoRL The nearly two dozen research papers the NVIDIA robotics team released with CoRL cover breakthroughs in integrating vision language models for improved environmental understanding and task execution, temporal robot navigati
LINK: https://blogs.nvidia.com/blog/robot-learning-humanoid-development/...
See more stories from nvidia

North America Stories

14/05/2026

Sweetwater and Airstream Unveil Mobile Dolby Atmos Recording Studio

Sweetwater and Airstream have announced a custom-built Dolby Atmos mobile recording studio inside an Airstream trailer, set to tour music festivals, schools, tr...

14/05/2026

American Association of Professional Baseball Expands Broadcast Distribution for 2026 Season

The American Association of Professional Baseball (AAPB) has announced a new par...

14/05/2026

ESPN to Establish Week-Long Super Bowl LXI Broadcast Center on Santa Monica Beach

ESPN has announced plans to transform Santa Monica Beach into a broadcast hub du...

14/05/2026

Amagi Announces Major Upgrade to CLOUDPORT Cloud Broadcast Platform

Amagi has announced a significant update to Amagi CLOUDPORT, its cloud-based broadcast playout platform. The update includes 250-plus features shipped in FY25-2...

14/05/2026

Clear-Com to Showcase New Products and Platform Updates at InfoComm 2026

Clear-Com will exhibit at InfoComm 2026 (Booth N7005, June 17-19, Las Vegas Convention Center), introducing a new product that builds on Arcadia Central Station...

14/05/2026

Ikegami to Exhibit at BroadcastAsia 2026 with Two New Viewfinder Premieres

Ikegami will exhibit at BroadcastAsia 2026 (Stand 5D3-1, Singapore Expo, May 20-22), introducing two new viewfinders alongside its existing camera, control, and...

14/05/2026

dB Broadcast Delivers IP-Based OB Trucks for Cloudbass Featuring Grass Valley LDX 100 Cameras

Grass Valley has announced that dB Broadcast has delivered new IP-based outside ...

14/05/2026

NAGRAVISION and WPBSA Launch Play Snooker Digital Platform

NAGRAVISION, a Kudelski Group company, has announced a partnership with the World Professional Billiards and Snooker Association (WPBSA) to launch Play Snooker,...

14/05/2026

Belden To Acquire RUCKUS Networks for Approximately $1.85 Billion

Belden Inc. has announced a definitive agreement to acquire RUCKUS Networks from Vistance Networks for approximately $1.85 billion. The transaction has been app...

14/05/2026

NVIDIA Releases Content Localization Blueprint for AI-Assisted Dubbing and Speech Translation

NVIDIA has released the Content Localization Blueprint, a modular reference arch...

14/05/2026

Disney+ To Stream Banana Bowl Championship Live This October

Disney has announced that Disney will be the exclusive U.S. streaming home of the Banana Bowl, the Banana Ball league season championship, streaming live this ...

14/05/2026

Arkona and Manifold To Exhibit on Magna Systems Stand at BroadcastAsia 2026

Arkona technologies and technology partner manifold will demonstrate their production solutions on the Magna Systems and Engineering stand (Booth 5D1-1) at Broa...

14/05/2026

Haivision Webinar: New Broadcast Contribution Products Featuring Minor League Baseball Case Study

Haivision will host a webinar on Thursday, May 21 at 10 a.m. ET / 4 p.m. CET cov...

14/05/2026

The CW Network and ESPN Announce ACC Sublicense Agreement Through 2030-31 Season

The CW Network and ESPN have announced a sublicense broadcast agreement for The CW to televise ACC football and men's and women's college basketball gam...

14/05/2026

Detroit Pistons Ink Local Media Rights Deal With Scripps Sports

The agreement marks Scripps Sports' first NBA local rights deal...

14/05/2026

Report: Broadcasting Among Hardest-Hit Industries as AI Reshapes the Workforce

A new report from education-technology company Wiingy testing post-ChatGPT predictions against three years of real-world data has identified broadcasting as one...

14/05/2026

Madonna, Shakira, and BTS to Headline First-Ever FIFA World Cup Final Halftime Show

Global Citizen and FIFA have announced that Madonna, Shakira, and BTS will headl...

14/05/2026

Sundance Institute Names 2026 Episodic Lab Fellows

LOS ANGELES, CA, May 14, 2026 - The nonprofit Sundance Institute announced today the cohort selected for the 2026 Episodic Lab program, taking place at Dunaway ...

14/05/2026

The Wraith Shield Advantage: Transforming L3Harris Radios into AI-Enabled Counter-UAS Sensors

Soldiers equipped with Falcon IV radios will soon gain a sense-and-protect capa...

14/05/2026

Getting into the Space Nuclear Power Game with Next-Generation Technology

Artists concept of the L3Harris Next Gen RTG in flight configuration, designed to provide 250 watts of reliable power for decades-long missions in deep space....

14/05/2026

Nielsen data shows NZ vehicle advertisers are shifting gears as fuel pressures make EVs and hybrids an increasingly attractive option

Car ad spend rises sharply in March as more auto buyers turn to electric, hybrid...

14/05/2026

Is This the Year for Agentic AI's Breakout in Broadcast?

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

CBS LA, Los Angeles Rams Ink New TV Deal

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

CueScripts CueiT 4 0 Wins Futures Best of Show Award Pres...

CueScript's CueiT 4.0 Wins Future's Best of Show Award, Presented at 2026 NAB Show by TV Tech CueScript, a leading international developer of professio...

14/05/2026

Expert-Led Education Sessions and Development of Online T...

Expert-Led Education Sessions and Development of Online Training Program Accelerate IPMX Adoption and Deployment The Alliance for IP Media Solutions (AIMS) to...

14/05/2026

Klvr rechargeable battery launches in the USA to cut cost...

Klvr is launching in the United States with a professional-grade rechargeable battery solution that cuts costs and improves performance across live entertainmen...

14/05/2026

Shooting into the depths of Bedlam with URSA Cine 17K 65

Shooting into the depths of Bedlam with URSA Cine 17K 65 Brie Clayton May 14, 2026 0 Comments Indie feature film paired digital 65mm capture with a Bl...

14/05/2026

WeMakeColor expands with Baselight, becoming hybrid color facility

WeMakeColor expands with Baselight, becoming hybrid color facility Caroline Shawley May 14, 2026 0 Comments Boutique Mexican-based studio integrates B...

14/05/2026

Berklee's Summer in the City Returns with Free Concerts Throughout Boston Area

Berklee's Summer in the City Returns with Free Concerts Throughout Boston Ar...

14/05/2026

Chelsey Green Named to Billboard's 2026 Women in Music List

Chelsey Green Named to Billboard's 2026 Women in Music List The Berklee professor and chair of the Recording Academy Board of Trustees joins other high-pr...

14/05/2026

Parks: Tubi, Roku Channel Are Top U.S. FAST Platforms

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

Study: Downstream Fiber Usage Outpaces Cable Broadband

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

NABLF to Honor Kidde With Corporate Leadership Award

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

LABF Awards Four 2026 Preservation Grants

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

Netflix Expands NFL Deal to Five Games

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

Scripps Seals Local Broadcast Deal with Detroit Pistons

Share Copy link Facebook X Linkedin Bluesky Email...

14/05/2026

Glensound marks 60 years of audio innovation at Broadcast...

Six decades of products built around the people who use them...

14/05/2026

Vivid Broadcast builds agile remote production network ar...

Designed to embrace multiple production processes and deliver high-end live broadcast workflows, Vivid Broadcast has combined Calrec True Control 2.0 enabled co...

14/05/2026

Jigsaw24 and EVS Partner to Strengthen Future Proof UK Li...

Today, Jigsaw24, the UK's leading media equipment supplier and systems integrator, announces a new partnership with EVS, a global leader in live video techn...

14/05/2026

Berklee Convenes Leaders in AI, Music for Inaugural AIMS Symposium

Berklee Convenes Leaders in AI, Music for Inaugural AIMS Symposium The three-day event puts musicians at the center of the future of music creation, ethics, a...

14/05/2026

The Last Laugh: NIAJ Fest Wraps Another Successful Los Angeles Takeover

Back to All News The Last Laugh: NIAJ Fest Wraps Another Successful Los Angeles Takeover Entertainment 14 May 2026 GlobalUnited States Link copied to clipb...

14/05/2026

May 13, 2026

Scripps Research establishes endowed chair honoring renowned structural biologist Ian Wilson Keren Lasker to be inaugural chair holder May 13, 2026 LA JOLLA ...

14/05/2026

Sea You in the Cloud: Subnautica 2' Early Access Dives Onto GeForce NOW

Dive masks on - Subnautica 2 is making a splash on GeForce NOW day-and-date with launch, so members can plunge into the title's brand-new alien ocean from a...

13/05/2026

New Adobe Premiere Color Grading Mode Accelerated on NVIDIA GPUs

New Adobe Premiere Color Grading Mode Accelerated on NVIDIA GPUs Joel Pennington May 13, 2026 0 Comments New NVIDIA RTX-accelerated features streamlin...

13/05/2026

dB Broadcast Delivers New IP-based Cloudbass Sports OB Tr...

Grass Valley announced that dB Broadcast has delivered new IP-based outside broadcast (OB) trucks for Cloudbass, featuring Grass Valley LDX 100 Series cameras a...

13/05/2026

Ikegami Announces its Broadcast Asia 2026 Innovations

Ikegami will exhibit the latest additions to its wide range of broadcast production cameras, control units, viewfinders and monitors on stand 5D3-1 at Broadcast...

13/05/2026

XRSA and FISE Partner to Deliver Immersive Action Sports...

FISE, working with the founding members of the XR Sports Alliance (XRSA), Accedo, Qualcomm Technologies, Inc. and HBS, have collaborated to develop an immersive...

13/05/2026

Canon Unveils New EOS R6 V Full-Frame EOS Camera and RF20-50mm F4 L IS USM PZ Built-In Power Zoom Lens

Canon Unveils New EOS R6 V Full-Frame EOS Camera and RF20-50mm F4 L IS USM PZ Bu...

13/05/2026

Boston Conservatory at Berklee Honors Beth Morrison and Moses Pendleton at Commencement Ceremony

Boston Conservatory at Berklee Honors Beth Morrison and Moses Pendleton at Comme...

13/05/2026

CBS Boston to Air CCBL Baseball

Share Copy link Facebook X Linkedin Bluesky Email...