Sony Pixel Power calrec Sony

NVIDIA Advances Robot Learning and Humanoid Development With New AI and Simulation Tools

06/11/2024

www.1x.tech

Robotics developers can greatly accelerate their work on AI-enabled robots, including humanoids, using new AI and simulation tools and workflows that NVIDIA revealed this week at the Conference for Robot Learning (CoRL) in Munich, Germany.

The lineup includes the general availability of the NVIDIA Isaac Lab robot learning framework; six new humanoid robot learning workflows for Project GR00T, an initiative to accelerate humanoid robot development; and new world-model development tools for video data curation and processing, including the NVIDIA Cosmos tokenizer and NVIDIA NeMo Curator for video processing.

The open-source Cosmos tokenizer provides robotics developers superior visual tokenization by breaking down images and videos into high-quality tokens with exceptionally high compression rates. It runs up to 12x faster than current tokenizers, while NeMo Curator provides video processing curation up to 7x faster than unoptimized pipelines.

Also timed with CoRL, NVIDIA presented 23 papers and nine workshops related to robot learning and released training and workflow guides for developers. Further, Hugging Face and NVIDIA announced they're collaborating to accelerate open-source robotics research with LeRobot, NVIDIA Isaac Lab and NVIDIA Jetson for the developer community.

Accelerating Robot Development With Isaac Lab NVIDIA Isaac Lab is an open-source, robot learning framework built on NVIDIA Omniverse, a platform for developing OpenUSD applications for industrial digitalization and physical AI simulation.

Developers can use Isaac Lab to train robot policies at scale. This open-source unified robot learning framework applies to any embodiment - from humanoids to quadrupeds to collaborative robots - to handle increasingly complex movements and interactions.

Leading commercial robot makers, robotics application developers and robotics research entities around the world are adopting Isaac Lab, including 1X, Agility Robotics, The AI Institute, Berkeley Humanoid, Boston Dynamics, Field AI, Fourier, Galbot, Mentee Robotics, Skild AI, Swiss-Mile, Unitree Robotics and XPENG Robotics.

Project GR00T: Foundations for General-Purpose Humanoid Robots Building advanced humanoids is extremely difficult, demanding multilayer technological and interdisciplinary approaches to make the robots perceive, move and learn skills effectively for human-robot and robot-environment interactions.

Project GR00T is an initiative to develop accelerated libraries, foundation models and data pipelines to accelerate the global humanoid robot developer ecosystem.

Six new Project GR00T workflows provide humanoid developers with blueprints to realize the most challenging humanoid robot capabilities. They include:

GR00T-Gen for building generative AI-powered, OpenUSD-based 3D environments

GR00T-Mimic for robot motion and trajectory generation

GR00T-Dexterity for robot dexterous manipulation

GR00T-Control for whole-body control

GR00T-Mobility for robot locomotion and navigation

GR00T-Perception for multimodal sensing

Humanoid robots are the next wave of embodied AI, said Jim Fan, senior research manager of embodied AI at NVIDIA. NVIDIA research and engineering teams are collaborating across the company and our developer ecosystem to build Project GR00T to help advance the progress and development of global humanoid robot developers.

New Development Tools for World Model Builders Today, robot developers are building world models - AI representations of the world that can predict how objects and environments respond to a robot's actions. Building these world models is incredibly compute- and data-intensive, with models requiring thousands of hours of real-world, curated image or video data.

NVIDIA Cosmos tokenizers provide efficient, high-quality encoding and decoding to simplify the development of these world models. They set a new standard of minimal distortion and temporal instability, enabling high-quality video and image reconstructions.

Providing high-quality compression and up to 12x faster visual reconstruction, the Cosmos tokenizer paves the path for scalable, robust and efficient development of generative applications across a broad spectrum of visual domains.

1X, a humanoid robot company, has updated the 1X World Model Challenge dataset to use the Cosmos tokenizer.

NVIDIA Cosmos tokenizer achieves really high temporal and spatial compression of our data while still retaining visual fidelity, said Eric Jang, vice president of AI at 1X Technologies. This allows us to train world models with long horizon video generation in an even more compute-efficient manner.

Other humanoid and general-purpose robot developers, including XPENG Robotics and Hillbot, are developing with the NVIDIA Cosmos tokenizer to manage high-resolution images and videos.

NeMo Curator now includes a video processing pipeline. This enables robot developers to improve their world-model accuracy by processing large-scale text, image and video data.

Curating video data poses challenges due to its massive size, requiring scalable pipelines and efficient orchestration for load balancing across GPUs. Additionally, models for filtering, captioning and embedding need optimization to maximize throughput.

NeMo Curator overcomes these challenges by streamlining data curation with automatic pipeline orchestration, reducing processing time significantly. It supports linear scaling across multi-node, multi-GPU systems, efficiently handling over 100 petabytes of data. This simplifies AI development, reduces costs and accelerates time to market.

Advancing the Robot Learning Community at CoRL The nearly two dozen research papers the NVIDIA robotics team released with CoRL cover breakthroughs in integrating vision language models for improved environmental understanding and task execution, temporal robot navigati
LINK: https://blogs.nvidia.com/blog/robot-learning-humanoid-development/...
See more stories from nvidia

North America Stories

15/05/2026

Seattle Sounders FC and Reign FC Announce Seattle Soccer Celebration at Waterfront Park

Seattle Sounders FC and Seattle Reign FC, in partnership with RAVE Foundation an...

15/05/2026

How Sound Designer Dan Brumm Built Blueys Audio World with Sennheiser and Neumann

Dan Brumm has served as sound designer on Bluey, the Australian children's t...

15/05/2026

Applications Close May 31 for Mark Brunner Professional Audio Scholarship

The Professional Audio Manufacturers Alliance (PAMA) and Shure Incorporated are accepting applications for the 6th annual Mark Brunner Professional Audio Schola...

15/05/2026

Netflix Expands NFL Coverage With Additional Games Starting in 2026

Netflix has announced an expanded NFL schedule for 2026 and beyond under a four-year partnership extension with the NFL through the 2029-30 season. Each season,...

15/05/2026

Ateme Supports TVRIs SRT-Based Live Sports Contribution and Distribution Workflow

Ateme is supporting TVRI (Televisi Republik Indonesia) with a contribution and d...

15/05/2026

Concacaf Launches New Website and Mobile App Powered by Deltatre

Concacaf has announced the launch of a new website and mobile app built on Deltatre's FORGE platform. Concacaf.com and the mobile app, available on iOS and ...

15/05/2026

Qatar Media Corporation Launches QBC Business Channel in 4K via Eutelsat

Eutelsat has announced the launch of QBC Business Economic Channel by Qatar Media Corporation, broadcasting in 4K/UHD via Eutelsat's 7/8 West video neighbo...

15/05/2026

Amazon to Serve as Exclusive Launch Home of MLS Original Series Cup Dreams on May 14

Major League Soccer has announced four original content series timed to the 2026...

15/05/2026

AIMS to Focus on IPMX Education at InfoComm 2026

The Alliance for IP Media Solutions (AIMS) has announced it will exhibit and present at InfoComm 2026, taking place June 13-19 at the Las Vegas Convention Cente...

15/05/2026

InfoComm 2026To Feature Sports, Broadcast, and Live Event Technologies

InfoComm 2026 will take place June 13-19 (exhibits June 17-19) at the Las Vegas Convention Center. The show will include sessions and exhibits covering broadcas...

15/05/2026

Tracy McGradys Ones Basketball League Signs First Streaming Agreement with Fubo Sports Network

Tracy McGrady's Ones Basketball League (OBL) and FuboTV Inc. have announced ...

15/05/2026

Disguise and Creative Technology Return for Eighth Year at Eurovision Song Contest 2026

Disguise has partnered with Creative Technology (CT) to deliver visual playback ...

15/05/2026

Sony Announces Alpha 7R VI Camera and FE 100-400mm F4.5 GM OSS Lens

Sony Electronics has announced two new products for professional imaging: the Alpha 7R VI full-frame mirrorless camera and the FE 100-400mm F4.5 GM OSS super-te...

15/05/2026

SVG GameDay, Ep. 15: New Jersey Devils Joe Kuchie - Growing the Game in the Garden State

In-venue and creative video staffers at the professional and collegiate level ha...

15/05/2026

Ratings Roundup: ESPN Secures Top Viewed Second Round Game 4 of Stanley Cup on Cable; NBA Draft Lottery Viewership Up 23%

Ratings Roundup is a rundown of recent rating news and is derived from press rel...

15/05/2026

The Future of Sports Analytics: Building Trust and Intelligence With SmerSports and Cisco

For sports organizations, the most valuable assets are often the most sensitive:...

15/05/2026

NFL Broadcast Schedule Roundup: Breaking Down CBS, ESPN, FOX, NBC, Netflix, and Prime Lineups

The NFL's broadcast partners released their 2026 regular season schedules ye...

15/05/2026

Netflix Steps Into the Cage for First MMA Production With Rousey-Carano Showdown at Intuit Dome

When MMA icons Ronda Rousey and Gina Carano meet inside the Hexagon at Intuit Do...

15/05/2026

Dustin Hoffman and Leo Woodall Bring the Noise in Daniel Roher's Tuner

Daniel Roher attends the Tuner Premiere during the 2026 Sundance Film Festival at Eccles Theatre on January 22, 2026 in Park City, Utah. (Photo by Neilson Bar...

15/05/2026

CTV's Data Gap Holding Back Bigger Ad Budgets, New Gracenote Research Finds

86% of media planners would move more linear TV budget to CTV if they had show-level targeting and reporting - and 65% would also shift dollars from programmati...

15/05/2026

Scripps Completes Station Swaps with Gray Media

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Clear-Com Takes Communications Further at InfoComm 2026

Clear-Com will showcase new communications solutions and major platform updates at InfoComm 2026 (Booth N7005), June 17-19, in the North and Central Halls of t...

15/05/2026

Rise AV Launches Second Year of UK Elevate Programme Foll...

Following an outstanding inaugural year in 2025, Rise AV is proud to announce the return of its flagship leadership initiative, Elevate. The programme continues...

15/05/2026

Berklee Announces Lineup for Inaugural AI Music Summit

Berklee Announces Lineup for Inaugural AI Music Summit The three-day event puts musicians at the center of the future of music creation, ethics, and the indus...

15/05/2026

Lightware Highlights Scalable USB-C and AV-over-IP Innova...

Lightware returns to InfoComm 2026 with a focused showcase of scalable USB-C connectivity, next-generation AV-over-IP solutions, and technologies that help over...

15/05/2026

IAB Releases Campaign Data Standards 1.0 for Public Comment

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

ARRI Expands Management Board

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Gray Media Names Joanie Vasiliadis SVP of Transformation

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Study: Data and Measurement Problems Reduce CTV Ad Budgets

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

Upfronts: WBD Expands Advanced Ad Capabilities and AI Ad Tech

Share Copy link Facebook X Linkedin Bluesky Email...

15/05/2026

VLAST Powers PLAVEs Asia Tour Encore with AJA Gear

Delivering a live, arena-scale production of a massively popular band is no small feat. Between expansive in-arena LED walls and a global live stream fed to onl...

15/05/2026

Sun Broadcast Futureproofs Dayalbaghs Multimedia Van with...

Connection is the heartbeat of any strong community, and with live streaming becoming more accessible in the modern era, it's much easier for faith-based or...

15/05/2026

Disguise and Creative Technology Power Eurovision for the...

Powered by GX 3 media servers, optimised IP-VFC workflows and on-site engineering expertise, the production delivers high-performance visuals for one of the wor...

15/05/2026

A Mother, Two Daughters and One Big Scandal: Netflix's Crime-Comedy 'Maa Behen' Premieres June 4

Back to All News A Mother, Two Daughters and One Big Scandal: Netflixs Crime-Co...

14/05/2026

Sweetwater and Airstream Unveil Mobile Dolby Atmos Recording Studio

Sweetwater and Airstream have announced a custom-built Dolby Atmos mobile recording studio inside an Airstream trailer, set to tour music festivals, schools, tr...

14/05/2026

American Association of Professional Baseball Expands Broadcast Distribution for 2026 Season

The American Association of Professional Baseball (AAPB) has announced a new par...

14/05/2026

ESPN to Establish Week-Long Super Bowl LXI Broadcast Center on Santa Monica Beach

ESPN has announced plans to transform Santa Monica Beach into a broadcast hub du...

14/05/2026

Amagi Announces Major Upgrade to CLOUDPORT Cloud Broadcast Platform

Amagi has announced a significant update to Amagi CLOUDPORT, its cloud-based broadcast playout platform. The update includes 250-plus features shipped in FY25-2...

14/05/2026

Clear-Com to Showcase New Products and Platform Updates at InfoComm 2026

Clear-Com will exhibit at InfoComm 2026 (Booth N7005, June 17-19, Las Vegas Convention Center), introducing a new product that builds on Arcadia Central Station...

14/05/2026

Ikegami to Exhibit at BroadcastAsia 2026 with Two New Viewfinder Premieres

Ikegami will exhibit at BroadcastAsia 2026 (Stand 5D3-1, Singapore Expo, May 20-22), introducing two new viewfinders alongside its existing camera, control, and...

14/05/2026

dB Broadcast Delivers IP-Based OB Trucks for Cloudbass Featuring Grass Valley LDX 100 Cameras

Grass Valley has announced that dB Broadcast has delivered new IP-based outside ...

14/05/2026

NAGRAVISION and WPBSA Launch Play Snooker Digital Platform

NAGRAVISION, a Kudelski Group company, has announced a partnership with the World Professional Billiards and Snooker Association (WPBSA) to launch Play Snooker,...

14/05/2026

Belden To Acquire RUCKUS Networks for Approximately $1.85 Billion

Belden Inc. has announced a definitive agreement to acquire RUCKUS Networks from Vistance Networks for approximately $1.85 billion. The transaction has been app...

14/05/2026

NVIDIA Releases Content Localization Blueprint for AI-Assisted Dubbing and Speech Translation

NVIDIA has released the Content Localization Blueprint, a modular reference arch...

14/05/2026

Disney+ To Stream Banana Bowl Championship Live This October

Disney has announced that Disney will be the exclusive U.S. streaming home of the Banana Bowl, the Banana Ball league season championship, streaming live this ...

14/05/2026

Arkona and Manifold To Exhibit on Magna Systems Stand at BroadcastAsia 2026

Arkona technologies and technology partner manifold will demonstrate their production solutions on the Magna Systems and Engineering stand (Booth 5D1-1) at Broa...

14/05/2026

Haivision Webinar: New Broadcast Contribution Products Featuring Minor League Baseball Case Study

Haivision will host a webinar on Thursday, May 21 at 10 a.m. ET / 4 p.m. CET cov...

14/05/2026

The CW Network and ESPN Announce ACC Sublicense Agreement Through 2030-31 Season

The CW Network and ESPN have announced a sublicense broadcast agreement for The CW to televise ACC football and men's and women's college basketball gam...

14/05/2026

Detroit Pistons Ink Local Media Rights Deal With Scripps Sports

The agreement marks Scripps Sports' first NBA local rights deal...

14/05/2026

Report: Broadcasting Among Hardest-Hit Industries as AI Reshapes the Workforce

A new report from education-technology company Wiingy testing post-ChatGPT predictions against three years of real-world data has identified broadcasting as one...