
What makes a robot gripper useful isn't that it can pick up one object - it's that it can pick up the next one, and the one after that, with a tool it's never held before.
What makes an autonomous vehicle system safe isn't just that it can reason through a situation - it's that it can do so quickly enough on the hardware actually installed in the car.
What makes a virtual agent capable is exposure to as many different environments as possible before it faces the real world.
At this year's Computer Vision and Pattern Recognition (CVPR) conference, NVIDIA Research is presenting three papers that address each of these challenges - and share a common theme: training at scale creates systems that generalize across diverse applications.
The three papers cover different challenges in physical AI research:
GraspGen-X, the first foundation model for zero-shot grasping, was trained on billions of simulated grasps to work with any gripper it's shown.
LCDrive introduces a model that replaces expensive text-based reasoning with compact latent representations, letting autonomous vehicles think faster on embedded hardware.
NitroGen is a generalized gameplay AI foundation model that harnesses the NVIDIA Isaac GR00T robot foundation model architecture to help train embodied agents in virtual environments across tens of thousands of hours of interaction.
NVIDIA also unveiled at CVPR new physical AI agent skills that help researchers and developers speed the development of autonomous vehicles, robots and vision AI systems.
NitroGen and another NVIDIA-authored paper, PixelDIT, were named best paper finalists at the conference - an accolade given to just 15 of over 4,000 accepted papers at CVPR.
The First Foundation Model for Grasping Most AI systems for robotic grasping are specialists.
A vision-language-action policy trained for a two-finger gripper only learns to grasp with those two fingers. Similarly, a policy for dextrous grasping will only work for the bespoke multi-fingered gripper it's trained on. For every new embodiment, the process typically needs to be repeated - requiring new training data, fine-tuning and validation. This constraint means most robotics companies pick a gripper, train for it and stick with it.
GraspGen-X is the first foundation model for grasping built to eliminate this bottleneck.
Like a large language model that can apply its understanding of language to a new task without retraining, GraspGen-X applies its understanding of geometry and contact to any robotic gripper it encounters. Given the geometry of a new gripper and an unknown object it's never seen before, the model generates reliable grasp pose proposals to enable the robot to grasp the object.
https://blogs.nvidia.com/wp-content/uploads/2026/06/GraspGenX.mp4
To get there, the researchers needed a dataset that's impossible to collect in the real world at scale. They generated 2 billion simulated grasps across thousands of object shapes and synthetic gripper configurations, spanning the diversity of form factors a deployed robot might encounter.
For robot developers, this foundation model eliminates the need for per-gripper training cycles and can be applied out of the box for several commonly used grippers. GraspGenX can be used in conjunction with curoboV2, a new CUDA-accelerated motion planning library, to achieve these grasp poses in unknown environments.
Building on the GraspGen research foundation, another paper, Grasp-MPC - presented at ICRA 2026 - advances the next step in the pipeline: moving from grasp generation to closed-loop grasp execution.
Teaching Autonomous Vehicles to Think Faster In recent years, researchers have found that letting an AI reason - generating intermediate thinking steps before committing to an answer - reliably improves its decision-making.
For autonomous vehicles, the challenge is doing that reasoning on the hardware inside an actual vehicle. Text-based chain-of-thought reasoning generates words, and every word is a token that takes time to produce. On the processor running inside a car, token count is a real constraint on how fast the system can respond.
LCDrive tackles this problem by replacing words with compressed latent representations.
Instead of generating human-readable reasoning steps, the system thinks in a compact latent space - states that capture spatial information rather than producing text. The architecture alternates between two kinds of thinking: proposing candidate actions, then predicting what the world will look like if those actions are taken.
It uses that predicted world state to refine its next step. It's the same reasoning loop - just in a more computationally efficient form than natural language.
The result: comparable output trajectory quality to text-based reasoning, using roughly half the tokens.
The model was built on NVIDIA Alpamayo and trained using supervision derived from existing vehicle data.
Embodied Agents Trained in Virtual Worlds Isaac GR00T - NVIDIA's open foundation model for humanoid robots - is built on a simple principle: expose a model to enough diverse situations, and it will generalize to ones it hasn't seen.
NitroGen extends that principle to virtual environments, using the GR00T architecture to train a foundation model for embodied agents across a breadth of virtual worlds.
Video games offer something that's hard to build from scratch: structured, varied worlds with defined goals and well-specified success conditions. They're high-quality training environments, available at scale.
NitroGen treats them that way - as a training ground for agents that will eventually be trained to handle novel real- or simulated-world situations, like powering a robot that helps with housework based on broad instructions such as, Put these items away in the
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
09/10/2026
September 10 2026, 06:00 (PDT) Dolby Expands Dolby OptiView Platform with New Capabilities at IBC 2026
New Sports Intelligence helps providers better unders...
07/10/2026
Dalet, a leading technology and service provider for media-rich organizations, today announced the latest Long-Term Supported (LTS) release of Dalet Flex. Build...
23/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
23/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
23/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
23/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
23/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
23/09/2026
Westminster Schools, a K through 12 school in Atlanta, Georgia with approximately 2,000 students, has transformed its student broadcast and digital storytelling...
23/09/2026
The 2026 edition of CABSAT and SATExpo may have a new home and dates in the diary, but it still retains a strong presence from UK companies in the GREAT Britain...
23/09/2026
Glensound will return to CABSAT with new developments in professional monitoring and Dante interfacing, alongside its established commentary, esports and parlia...
22/09/2026
From Malaysia to Mexico and from DR Congo to Ukraine, this year's 12 finalists for the 2026 Thomson Foundation Young Journalist of the Year award tell extra...
22/09/2026
The annual preseason content-gathering event brought 36 NHL players to Las Vegas, where ESPN used 2.5 kilowatts of laser power and a motion-control camera syste...
22/09/2026
Neumann has introduced the KK 104 A (cardioid) and KK 105 A (supercardioid) wireless capsule heads, succeeding the KK 204 and KK 205. The new capsules are compa...
22/09/2026
The NFL and TMRW Sports have announced a new annual Pro Flag Showcase to be held during Super Bowl LXI week in Los Angeles. The event will feature top flag foot...
22/09/2026
Washington State University has unveiled a new end zone video display and audio ...
22/09/2026
Disney's SVP of Live Operations and Engineering details the technology behin...
22/09/2026
Cosm has announced two senior promotions. Mazen Alawar has been promoted to SVP,...
22/09/2026
Mississippi State University's University Television Center has purchased four KOKUSAI DENKI SK-UHD7000-S2 cameras since 2024, with plans to add two more, a...
22/09/2026
Net Insight has announced that Christine Kjellvard will assume the role of interim Chief Financial Officer on November 1, 2026. Kjellvard has previously held in...
22/09/2026
Net Insight has announced Zyntai TimeNode Access, a compact PTP Grandmaster designed for distributed edge sites. The product carries secure, precise time across...
22/09/2026
Sam Hazeldine, Austin Amelio, Padraic McKinley, Julia Jones, and Ethan Hawke attend the premiere of The Weight at The Ray Theatre on January 26, 2026, in Park...
22/09/2026
Second high-gain tone plug-in released
Following on from their debut release, Chainsaw Suite, Avalanche Tones have announced the launch of their second plug...
22/09/2026
Uncompressed audio with below 6ms of latency
Lewitt's have just introduced their first pair of headphones which, despite being aimed primarily at studio...
22/09/2026
Two new capsule heads for Sennheiser wireless systems
Neumann have announced the upcoming launch of the KK 104 A and KK 105 A, a pair of new premium condens...
22/09/2026
Apogee Plugins Are Moving to Plugin Alliance We're excited to share an important update about the future of Apogee plugins.
Apogee plugins are moving to Pl...
22/09/2026
Parkland, Florida, US, September 22, 2026: SipRadius, widely recognized for making content processing and connectivity secure and seamless, has been honored wit...
22/09/2026
VanityFilter 3.1.5 adds project files and adjustable protection around lips, bro...
22/09/2026
Chile's TV Modernizes Broadcast Operations with Blackmagic Design
Brie Clayton September 22, 2026
0 Comments
URSA Broadcast G2 brings versatility...
22/09/2026
OWC Shares Sneak Peek at Amazon Prime Big Deal Days Pricing
Brie Clayton September 22, 2026
0 Comments
Image courtesy of Deposit Photos
Business and ...
22/09/2026
Berklee's First Homecoming Block Party Brings Alumni Home for Global Music F...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
22/09/2026
NVIDIA AI Day Singapore, which takes place Sept. 22-23 at the Raffles City Conve...
22/09/2026
From Academy Award-winning Passion Pictures and BAFTA-winning director Marian Mo...
22/09/2026
Tuesday 22 September 2026
Portrait Artist of the Year returns to Sky Arts with ...
22/09/2026
Whether it be from the sun or from our phones and other devices, we are all exposed to electromagnetic radiation in some shape or form. With the dramatic growth...
22/09/2026
Comscore's latest AI Intelligence Report Shows Competition Intensifying as C...
22/09/2026
The Factory-X Project Sets Standards for Data Sovereignty in Industry
Arvato Systems Brings Digital Sovereignty to Industrial Practice
Factory-X Successfully...
22/09/2026
LONDON Apple today announced the opening of Apple Music Hall, a brand-new state-...
22/09/2026
RT , Creative Ireland and Shared Island Initiative ask young artists
to respon...
22/09/2026
To build and deploy sophisticated robotics applications that can perceive, reason and act in dynamic environments, developers need new physical AI models and to...