Sony Pixel Power calrec Sony

Research Galore From 2024: Recapping AI Advancements in 3D Simulation, Climate Science and Audio Engineering

30/12/2024

The pace of technology innovation has accelerated in the past year, most dramatically in AI. And in 2024, there was no better place to be a part of creating those breakthroughs than NVIDIA Research.

NVIDIA Research is comprised of hundreds of extremely bright people pushing the frontiers of knowledge, not just in AI, but across many areas of technology.

In the past year, NVIDIA Research laid the groundwork for future improvements in GPU performance with major research discoveries in circuits, memory architecture and sparse arithmetic. The team's invention of novel graphics techniques continues to raise the bar for real-time rendering. And we developed new methods for improving the efficiency of AI - requiring less energy, taking fewer GPU cycles and delivering even better results.

But the most exciting developments of the year have been in generative AI.

We're now able to generate, not just images and text, but 3D models, music and sounds. We're also developing better control over what is generated: to generate realistic humanoid motion and to generate sequences of images with consistent subjects.

The application of generative AI to science has resulted in high-resolution weather forecasts that are more accurate than conventional numerical weather models. AI models have given us the ability to accurately predict how blood glucose levels respond to different foods. Embodied generative AI is being used to develop autonomous vehicles and robots.

And that was just this year. What follows is a deeper dive into some of NVIDIA Research's greatest generative AI work in 2024. Of course, we continue to develop new models and methods for AI, and expect even more exciting results next year.

ConsiStory: AI-Generated Images With Main Character Energy ConsiStory, a collaboration between researchers at NVIDIA and Tel Aviv University, makes it easier to generate multiple images with a consistent main character - an essential capability for storytelling use cases such as illustrating a comic strip or developing a storyboard.

The researchers' approach introduced a technique called subject-driven shared attention, which reduces the time it takes to generate consistent imagery from 13 minutes to around 30 seconds.

Read the ConsiStory paper.

ConsiStory is capable of generating a series of images featuring the same character. Edify 3D: Generative AI Enters a New Dimension NVIDIA Edify 3D is a foundation model that enables developers and content creators to quickly generate 3D objects that can be used to prototype ideas and populate virtual worlds.

Edify 3D helps creators quickly ideate, lay out and conceptualize immersive environments with AI-generated assets. Novice and experienced content creators can use text and image prompts to harness the model, which is now part of the NVIDIA Edify multimodal architecture for developing visual generative AI.

Read the Edify 3D paper and watch the video on YouTube.

Fugatto: Flexible AI Sound Machine for Music, Voices and More A team of NVIDIA researchers recently unveiled Fugatto, a foundational generative AI model that can create or transform any mix of music, voices and sounds based on text or audio prompts.

The model can, for example, create music snippets based on text prompts, add or remove instruments from existing songs, modify the accent or emotion in a voice recording, or generate completely novel sounds. It could be used by music producers, ad agencies, video game developers or creators of language learning tools.

Read the Fugatto paper.

GluFormer: AI Predicts Blood Sugar Levels Four Years Out Researchers from the Weizmann Institute of Science, Tel Aviv-based startup Pheno.AI and NVIDIA led the development of GluFormer, an AI model that can predict an individual's future glucose levels and other health metrics based on past glucose monitoring data.

The researchers showed that, after adding dietary intake data into the model, GluFormer can also predict how a person's glucose levels will respond to specific foods and dietary changes, enabling precision nutrition. The research team validated GluFormer across 15 other datasets and found it generalizes well to predict health outcomes for other groups, including those with prediabetes, type 1 and type 2 diabetes, gestational diabetes and obesity.

Read the GluFormer paper.

LATTE3D: Enabling Near-Instant Generation, From Text to 3D Shape Another 3D generator released by NVIDIA Research this year is LATTE3D, which converts text prompts into 3D representations within a second - like a speedy, virtual 3D printer. Crafted in a popular format used for standard rendering applications, the generated shapes can be easily served up in virtual environments for developing video games, ad campaigns, design projects or virtual training grounds for robotics.

Read the LATTE3D paper.

MaskedMimic: Reconstructing Realistic Movement for Humanoid Robots To advance the development of humanoid robots, NVIDIA researchers introduced MaskedMimic, an AI framework that applies inpainting - the process of reconstructing complete data from an incomplete, or masked, view - to descriptions of motion.

Given partial information, such as a text description of movement, or head and hand position data from a virtual reality headset, MaskedMimic can fill in the blanks to infer full-body motion. It's become part of NVIDIA Project GR00T, a research initiative to accelerate humanoid robot development.

Read the MaskedMimic paper.

StormCast: Boosting Weather Prediction, Climate Simulation In the field of climate science, NVIDIA Research announced StormCast, a generative AI model for emulating atmospheric dynamics. While other machine learning models trained on global data have a spatial resolution of about 30 kilometers and a temporal resolution of six hours, StormCast achieves a 3-kilometer, hourly scale.
LINK: https://blogs.nvidia.com/blog/ai-research-2024/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

19/03/2026

How Chris Bolte Builds the Sound of PlayStation Games

How Chris Bolte Builds the Sound of PlayStation Games Chris Bolte '16 is a technical sound designer at Sucker Punch Productions, where he helps create the...

18/03/2026

Net Insight Names Larissa GrnerMeeus CPO

Newly named chief product officer (CPO), Larissa G rner Meeus will return to Net Insight on May 4. Larissa G rner Meeus will become CPO of Net Insight on May 4...

18/03/2026

SVG Europe's The Football Summit 2026: Select Sessions Now Available to Watch on SVG PLAY

SVG Europe's The Football Summit 2026 explored how the sports broadcasting i...

18/03/2026

SVG's SportsTech@NABShow Blog Goes Live as Countdown to Vegas Heats Up

The 2026 NAB Show kicks off one month from today and SVG is once again set to cover the show from every angle. SVG's SportsTech@NAB Show Blog is now live - ...

18/03/2026

The Premiere of Dead Lover Showcases Grave Robbing and Sex With a Giant Finger

(L-R) The cast and crew of Dead Lover at The Ray Theater for its premiere at the 2025 Sundance Film Festival. (Photo by Robin Marshall/Shutterstock for Sundan...

18/03/2026

Music Row Piano from Wiltone Productions

Yamaha C7 captured in Nashville Wiltone Productions have announced the release of Music Row Piano, a deeply sampled Yamaha C7 piano library that's been ...

18/03/2026

Techivation introduce T-Warmer Mk2

Bass-enhancement plug-in upgraded Along with a steady stream of new releases, Techivation have recently been revisiting some of the older plug-ins in their ...

18/03/2026

Tonal Balance Control 3 from iZotope

Mix-reference plug-in overhauled iZotope's powerful mix-referencing plug-in has just reached its third major version, and now boasts a new capture proce...

18/03/2026

Toontrack Transistor Organ EKX

Latest EZKeys 2 expansion arrives Toontrack's staggering collection of EZKeys 2 expansions has grown once again, and the latest instalment delivers a on...

18/03/2026

Jess Ho explores the politics of food in new SBS Audio podcast For The Culture

Jess Ho explores the politics of food in new SBS Audio podcast For The Culture 18 March, 2026 Media releases New SBS Audio podcast For The Culture, hosted ...

18/03/2026

Future-ready broadcasts with ATSC 3.0 and enhanced services from Rohde & Schwarz at NAB 2026

Future-ready broadcasts with ATSC 3.0 and enhanced services from Rohde & Schwarz...

18/03/2026

LCTWS 2026

London Calling - When The Industry Convened to Help Streaming Find its MoJo In this blog, Laura Rognoni reflects on key discussions from the Connected TV World...

18/03/2026

SMPTE Unveils 2026 NAB Show Educational Presentations

SMPTE Unveils 2026 NAB Show Educational Presentations Brie Clayton March 18, 2026 0 Comments SMPTE , the home of media professionals, technologists, a...

18/03/2026

Auditel Ad Campaign Shot on Blackmagic PYXIS 12K

Auditel Ad Campaign Shot on Blackmagic PYXIS 12K Brie Clayton March 18, 2026 0 Comments LED wall virtual production blends 12K open gate acquisition w...

18/03/2026

Brainstorm transforms productivity and sustainability with Suite 7 at NAB Show 2026

Brainstorm transforms productivity and sustainability with Suite 7 at NAB Show 2...

18/03/2026

Neutrik To Showcase opticalCON ADVANCED Connectors At 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

SMPTE Details 2026 NAB Show Educational Sessions

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

Ben Bradshaw Joins PSSI as Director, Product and Network Development

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

Peter Thordarson Joins ASG as Technical Account Executive

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

Survey: Voters Trust TV News Over AI, Social and Search

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

2026 NAB Show Exhibitor Insight: Amazon Web Services (AWS)

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

SMPTE Unveils 2026 NAB Show Educational Presentations

SMPTE , the home of media professionals, technologists, and engineers, today unveiled its educational presentations for the 2026 NAB Show. This year SMPTE will ...

18/03/2026

Maxon Marks Its Official Entry Into the AEC Market With I...

Maxon, maker of powerful, approachable software solutions for creators working in 2D and 3D design, motion graphics, visual effects, gaming, and more, today ann...

18/03/2026

Digital Alert Systems NAB Preview 2026

Digital Alert Systems Preview 2026 NAB Show April 19 - 22 Booth C3452 At the 2026 NAB Show, Digital Alert Systems will showcase Version 6.0 of its DASDEC ...

18/03/2026

Setplex Transforms Video Streaming with AI and Super Aggr...

Setplex today announced that it will showcase its complete, fully integrated Zapflex platform for the first time at the 2026 NAB Show, introducing powerful new ...

18/03/2026

SES Announces Extension of Tender Offer

THIS ANNOUNCEMENT RELATES TO THE DISCLOSURE OF INFORMATION THAT QUALIFIED OR MAY HAVE QUALIFIED AS INSIDE INFORMATION WITHIN THE MEANING OF ARTICLE 7(1) OF THE ...

18/03/2026

COW Jobs: Seeking DP for Low Budget Dramedy - Chicago

COW Jobs: Seeking DP for Low Budget Dramedy - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Director of Photography for Low Budget Dramedy Fe...

18/03/2026

COW Jobs: Seeking Gaffer for Low Budget Dramedy - Chicago

COW Jobs: Seeking Gaffer for Low Budget Dramedy - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Gaffer for Low Budget Dramedy Feature Film- I...

18/03/2026

COW Jobs: Seeking Location, Sound for Low Budget Dramedy - Chicago

COW Jobs: Seeking Location, Sound for Low Budget Dramedy - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Location/Sound for Low Budget Dramed...

18/03/2026

COW Jobs: Seeking Child Wrangler for Low Budget Film - Chicago

COW Jobs: Seeking Child Wrangler for Low Budget Film - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Child Wrangler for Low Budget Dramedy Fe...

18/03/2026

Calrec Redefines Broadcast Workflows at NAB 2026 with its Most Powerful Hardware, Virtual and Hybrid Audio Lineup Yet

Calrec Redefines Broadcast Workflows at NAB 2026 with its Most Powerful Hardware...

18/03/2026

Oscar Nominated Two People Exchanging Saliva Posted with DaVinci Resolve Studio

Oscar Nominated Two People Exchanging Saliva Posted with DaVinci Resolve Studio Brie Clayton March 17, 2026 0 Comments DaVinci Resolve Studio handle...

18/03/2026

Boston Conservatory Presents Celebrated Musical Satire Urinetown

Boston Conservatory Presents Celebrated Musical Satire Urinetown Performances for this Center Stage production will take place at Boston Conservatory Theater ...

18/03/2026

Charlie Puth Joins Switched On Pop at Berklee NYC

Charlie Puth Joins Switched on Pop at Berklee NYC The Berklee alum spoke with host and Berklee NYC professor Charlie Harding for a live taping, answering audi...

18/03/2026

X-Rite Pantone Demonstrates Advanced Color Management Solutions to Support Smart Manufacturing at MAX and American Coatings Show

X-Rite Pantone Demonstrates Advanced Color Management Solutions to Support Smart...

18/03/2026

BAFTA winner James McAvoy to star in Sky Original series Meantime, based on the novel by Frankie Boyle

Wednesday 18 March 2026 BAFTA winner James McAvoy to star in Sky Original serie...

18/03/2026

Sky Sports and Zuffa Boxing announce multi-year agreement for the UK and Ireland

Wednesday 18 March 2026 Sky Sports and Zuffa Boxing announce multi-year agreement for the UK and Ireland Sky Sports and Zuffa Boxing have announced a new mult...

18/03/2026

Sky Unveils Poster and Trailer for Original Feature Film Shaun the Sheep: The Beast of Mossy Bottom

Wednesday 18 March 2026 Sky Unveils Poster and Trailer for Original Feature Fil...

18/03/2026

UP NEXT: Sky presents a bold slate of new, premium acquired content

Wednesday 18 March 2026 UP NEXT: Sky presents a bold slate of new, premium acquired content Sky reveals slate of newly acquired premium drama and comedy, as i...

18/03/2026

UP NEXT: Sky unveils an unmissable genre-spanning slate of new shows, exclusive previews and first looks for 2026 and beyond

Wednesday 18 March 2026 UP NEXT: Sky unveils an unmissable genre-spanning slate...

18/03/2026

Preparatevi Alla Terza E Ultima Stagione De 'La Legge di Lidia Pot' In Arrivo Solo Su Netflix Dal 15 Aprile

Back to All News Preparatevi Alla Terza E Ultima Stagione De La Legge di Lidia ...

18/03/2026

Netflix Drops Teaser and First Look Images for The Chestnut Man: Hide and Seek'

Back to All News Netflix Drops Teaser and First Look Images for The Chestnut M...

18/03/2026

Harmonic Enhances XOS Advanced Media Processor to Streamline Next-Generation Broadcast Distribution

New Playout-to-Delivery Capabilities Elevate ATSC 3.0 Experiences, Lower DTV In...

17/03/2026

NASA+ Prepares To Live Stream Historic Artemis II Mission, Bringing Deep-Space Exploration to Global Audiences

NASA+'s Rebecca Sirmons and Brittany Brown offer unique look at live streami...