Sony Pixel Power calrec Sony

Research Galore From 2024: Recapping AI Advancements in 3D Simulation, Climate Science and Audio Engineering

30/12/2024

The pace of technology innovation has accelerated in the past year, most dramatically in AI. And in 2024, there was no better place to be a part of creating those breakthroughs than NVIDIA Research.

NVIDIA Research is comprised of hundreds of extremely bright people pushing the frontiers of knowledge, not just in AI, but across many areas of technology.

In the past year, NVIDIA Research laid the groundwork for future improvements in GPU performance with major research discoveries in circuits, memory architecture and sparse arithmetic. The team's invention of novel graphics techniques continues to raise the bar for real-time rendering. And we developed new methods for improving the efficiency of AI - requiring less energy, taking fewer GPU cycles and delivering even better results.

But the most exciting developments of the year have been in generative AI.

We're now able to generate, not just images and text, but 3D models, music and sounds. We're also developing better control over what is generated: to generate realistic humanoid motion and to generate sequences of images with consistent subjects.

The application of generative AI to science has resulted in high-resolution weather forecasts that are more accurate than conventional numerical weather models. AI models have given us the ability to accurately predict how blood glucose levels respond to different foods. Embodied generative AI is being used to develop autonomous vehicles and robots.

And that was just this year. What follows is a deeper dive into some of NVIDIA Research's greatest generative AI work in 2024. Of course, we continue to develop new models and methods for AI, and expect even more exciting results next year.

ConsiStory: AI-Generated Images With Main Character Energy ConsiStory, a collaboration between researchers at NVIDIA and Tel Aviv University, makes it easier to generate multiple images with a consistent main character - an essential capability for storytelling use cases such as illustrating a comic strip or developing a storyboard.

The researchers' approach introduced a technique called subject-driven shared attention, which reduces the time it takes to generate consistent imagery from 13 minutes to around 30 seconds.

Read the ConsiStory paper.

ConsiStory is capable of generating a series of images featuring the same character. Edify 3D: Generative AI Enters a New Dimension NVIDIA Edify 3D is a foundation model that enables developers and content creators to quickly generate 3D objects that can be used to prototype ideas and populate virtual worlds.

Edify 3D helps creators quickly ideate, lay out and conceptualize immersive environments with AI-generated assets. Novice and experienced content creators can use text and image prompts to harness the model, which is now part of the NVIDIA Edify multimodal architecture for developing visual generative AI.

Read the Edify 3D paper and watch the video on YouTube.

Fugatto: Flexible AI Sound Machine for Music, Voices and More A team of NVIDIA researchers recently unveiled Fugatto, a foundational generative AI model that can create or transform any mix of music, voices and sounds based on text or audio prompts.

The model can, for example, create music snippets based on text prompts, add or remove instruments from existing songs, modify the accent or emotion in a voice recording, or generate completely novel sounds. It could be used by music producers, ad agencies, video game developers or creators of language learning tools.

Read the Fugatto paper.

GluFormer: AI Predicts Blood Sugar Levels Four Years Out Researchers from the Weizmann Institute of Science, Tel Aviv-based startup Pheno.AI and NVIDIA led the development of GluFormer, an AI model that can predict an individual's future glucose levels and other health metrics based on past glucose monitoring data.

The researchers showed that, after adding dietary intake data into the model, GluFormer can also predict how a person's glucose levels will respond to specific foods and dietary changes, enabling precision nutrition. The research team validated GluFormer across 15 other datasets and found it generalizes well to predict health outcomes for other groups, including those with prediabetes, type 1 and type 2 diabetes, gestational diabetes and obesity.

Read the GluFormer paper.

LATTE3D: Enabling Near-Instant Generation, From Text to 3D Shape Another 3D generator released by NVIDIA Research this year is LATTE3D, which converts text prompts into 3D representations within a second - like a speedy, virtual 3D printer. Crafted in a popular format used for standard rendering applications, the generated shapes can be easily served up in virtual environments for developing video games, ad campaigns, design projects or virtual training grounds for robotics.

Read the LATTE3D paper.

MaskedMimic: Reconstructing Realistic Movement for Humanoid Robots To advance the development of humanoid robots, NVIDIA researchers introduced MaskedMimic, an AI framework that applies inpainting - the process of reconstructing complete data from an incomplete, or masked, view - to descriptions of motion.

Given partial information, such as a text description of movement, or head and hand position data from a virtual reality headset, MaskedMimic can fill in the blanks to infer full-body motion. It's become part of NVIDIA Project GR00T, a research initiative to accelerate humanoid robot development.

Read the MaskedMimic paper.

StormCast: Boosting Weather Prediction, Climate Simulation In the field of climate science, NVIDIA Research announced StormCast, a generative AI model for emulating atmospheric dynamics. While other machine learning models trained on global data have a spatial resolution of about 30 kilometers and a temporal resolution of six hours, StormCast achieves a 3-kilometer, hourly scale.
LINK: https://blogs.nvidia.com/blog/ai-research-2024/...
See more stories from nvidia

Most recent headlines

06/10/2025

France Tlvisions Wins Prestigious 2025 EBU Technology & Innovation Award in Groundbreaking Collaboration with Dalet

France T l visions, France's leading broadcaster, has received the 2025 EBU ...

04/09/2025

Monumental Sports & Entertainment and Dalet Win Prestigious 2025 NAB Show Project of the Year Award

Monumental Sports & Entertainment (MSE), in collaboration with Dalet, has been a...

16/06/2025

Tegna Announces Major Expansion of Local News Programming

TYSONS, Va. Tegna Inc. is embarking on a notable expansion of their already substantial local news programming by launching live and on-demand, local newscasts ...

16/06/2025

Netflix Expands Programmatic Ad Sales with Yahoo DSP

Netflix has announced that it is expanding its global programmatic ad offerings by partnering with Yahoo DSP. This will enable brands to buy Netflix advertising...

16/06/2025

Sub51 & Soundtrax announce Drop Pad 3

Instrument now boasts full NKS support Sub51 and Soundtrax have just announced the launch of an updated and improved version of their innovative sample-base...

16/06/2025

Roku, Amazon Team Up to Dominate CTV Ad Market

NEW YORK In a landmark agreement to overtake the burgeoning connected TV (CTV) advertising market, Amazon Ads and Roku today announced a new integration that gi...

16/06/2025

EdgeBeam Wireless Names Conrad Clemson CEO

ATLANTA, BALTIMORE, CINCINNATI and IRVING, Texas The four major broadcast groups behind the ATSC 3.0-based EdgeBeam Wireless datacasting joint venture today nam...

16/06/2025

Amazon MGM Studios to Deploy Avid Tools on AWS

BURLINGTON, Mass. Avid today announced an extended agreement with Amazon MGM Studios to integrate Avid's Media Composer and Avid NEXIS on Amazon Web Service...

16/06/2025

Maxon Epic Sale Drops June 16

Maxon, maker of powerful, approachable software for creators working in 2D and 3D design, motion graphics, visual effects, gaming and more, today announced the ...

16/06/2025

Alfalite launches Skypix a new ceiling-mounted Led panel...

Alfalite, the only European manufacturer of LED displays, announces the launch of SKYPIX RGBW & IM, a new series of ceiling-mounted LED panels designed specifi...

16/06/2025

ALM/Busy Circuits launch Pip Filter & LFO

Two new compact 4HP modules introduced ALM/Busy Circuits have just announced the launch of two new Eurorack modules, the Pip Filter and Pip LFO, both of whi...

16/06/2025

Run With Ray in Cork, Waterford, Kilkenny, Drogheda and Dublin as The Ray D'Arcy Show hits the road

Run with Ray is back! RT Radio 1's The Ray D'Arcy Show hits the road th...

15/06/2025

Music Production for Women free in-person workshops

July 2025 in Dublin, Berlin, Amsterdam & London Photo: Thea Martre Music Production for Women (MPW) have announced that they will be running a series of fo...

15/06/2025

Jason's Piano & API Drums instruments from Sulcata Sound

Composer/producer launches free virtual instruments Sulcata Sound is the latest venture of Jason Graves, a two-time British Academy Award-winnning composer,...

14/06/2025

Pluto TV Adds All Womens Sports Network's FAST Channel

NEW YORK Pluto TV and the All Womens Sports Network have launched a free ad-supported streaming TV (FAST) AWSN channel in the U.S., Canada, the U.K. and the Nor...

14/06/2025

Scripps Inks Multiyear Agreement for WNBA Games on Ion

NEW YORK and CINCINNATI E.W. Scripps has announced a new, multiyear agreement with the WNBA that will continue Ions regular-season coverage of the league on Fri...

14/06/2025

NAB Highlights Hidden Importance of Spectrum in Major Sports Broadcasting

WASHINGTON The National Association of Broadcasters highlighted the hidden importance of spectrum in the production of major sporting events and described wha...

14/06/2025

1.0 Sunset, BPS and NextGen Broadcast's Potential Dominate ATSC Meeting

WASHINGTON Sunsetting ATSC 1.0, expanding business opportunities for NextGen Broadcast and increasing international adoption of the ATSC 3.0 standard were top o...

14/06/2025

Samba TV and Acxiom Announce Massive 40-market Global Expansion

SAN FRANCISCO Samba TV and Acxiom have announced that they will dramatically expand their longstanding relationship....

14/06/2025

MPW announce free in-person workshops

July 2025 in Dublin, Berlin, Amsterdam & London Photo: Thea Martre Music Production for Women (MPW) have announced that they will be running a series of fo...

14/06/2025

San Francisco State University's School of Cinema Uses Blackmagic Design

San Francisco State University's School of Cinema Uses Blackmagic Design Brie Clayton June 13, 2025 0 Comments More than 40 Blackmagic Design came...

14/06/2025

Boris FX Mocha Pro Adds New AI Tools To Tackle VFX Tasks Fast

Boris FX Mocha Pro Adds New AI Tools To Tackle VFX Tasks Fast Jessie Electa Petrov June 13, 2025 0 Comments The 2025.5 release helps artists work more...

14/06/2025

AJA Debuts DRM2-Plus Mini-Converter Frame at InfoComm 2025

AJA Debuts DRM2-Plus Mini-Converter Frame at InfoComm 2025 Brie Clayton June 13, 2025 0 Comments Next-gen frame addresses diverse rackmount needs wit...

13/06/2025

Prime Minister: A Behind-the-Scenes Look at a Leader Who Champions Kindness

(L-R) Lindsay Utz, Michelle Walshe, and The Right Honourable Dame Jacinda Ardern attend the 2025 Sundance Film Festival premiere of Prime Minister at Eccles T...

13/06/2025

Materialists' Director Celine Song Reveals the Inspirations Behind the Film's Soundtrack

Photo credit: Atsushi Nishijima If you're a true lover of rom-coms, chances...

13/06/2025

Pure Drama and Fierce Rivalries set to dominate the world's most iconic sporting event

Pure Drama and Fierce Rivalries set to dominate the world's most iconic spor...

13/06/2025

Press Release: NFVF Opens Call for Public Film Screenings on GBV Awareness as South Africa Confronts Ongoing Femicide Crisis

Johannesburg, 12 June 2025 - The National Film and Video Foundation (NFVF), an a...

13/06/2025

Central Texas Storm Knocks out KTXS Tower, Severely Damages Building

ABILENE. Texas A severe storm knocked down the tower and severely damaged the news studio and main facility of Sinclair-owned KTXS here on Sunday, June 8....

13/06/2025

Berklee's Music Business/Management Department Recognized by the Music Biz Association

Berklee's Music Business/Management Department Recognized by the Music Biz A...

13/06/2025

ATSC Honors Aldo Cugnini, Clarence Hau

WASHINGTON The ATSC, the Broadcast Standards Association, honored veteran technologist Aldo Cugnini and Clarence Hau, Senior Vice President of Standards, Policy...

13/06/2025

ESPN Doubles Down on Immersive Fan Experience for UFL Championship

(Editor's note: The 2025 UFL Championship Game between the D.C. Defenders and Michigan Panthers kicks off Saturday, June 14, at 8 p.m. Eastern. The game wil...

13/06/2025

Soulyft Audio release Chime

New iPad/iPhone synth App announced Following on from last year's release of Gradient Synth - which reached #6 on the App Store's Paid Music charts ...

13/06/2025

HBO Max Plans July Launches in 12 New Markets

LONDON Warner Bros. Discovery has announced that HBO Max will launch direct-to-consumer in multiple new countries this July as the streamer becomes available in...

13/06/2025

Verbit Launches Speaker Identification for Live ASR Broadcast Captions

AI voice transcription and captioning platform Verbit has added a new feature to its Captivate ASR solution the ability to identify specific features in automat...

13/06/2025

FCC's Anna Gomez Meets with TV Networks, Studio Execs and Unions

WASHINGTON Federal Communications Commission member Anna Gomez has wrapped up two weeks in California visiting broadcasters, television studio executives, enter...

13/06/2025

House Passes Rescission Package That Would Claw Back CPB Funding

WASHINGTON The U.S. House of Representatives voted mostly along party lines to approve a rescission package that would cancel $9.4 billion in previously approve...

13/06/2025

AJA Debuts DRM2-Plus Mini-Converter Frame at InfoComm 202...

At InfoComm 2025, AJA Video Systems announced DRM2-Plus, an intuitive, high-capacity 3RU frame that can neatly house up to 24 AJA Mini-Converters. Tailored to s...

13/06/2025

National CineMedia Selects Operatives Suite of AI-Based C...

Cinema advertising leader to leverage AOS and suite of AI-enabled solutions to optimize forecasting, yield management, and streamlined ad sales and operations a...

13/06/2025

Manfrotto Unveils Unique New Tripod System for Mirrorless...

Manfrotto has launched the ONE Hybrid Tripod, a new support system designed specifically for professional content creators working with mirrorless cameras acros...

13/06/2025

Synamedia unveils security enhancements to its Media Edge...

Leading video software provider, Synamedia, today announced that its Media Edge Gateway (MEG), an ATSC 3.0 software-based IRD, now supports Device Security requ...

13/06/2025

LiveU Strengthens Presence in DACH Region with Strategic...

LiveU, the global leader in live IP-video contribution, production and distribution solutions, is deepening its commitment to the German-speaking market with th...

13/06/2025

Chaos Adds AI Tools Faster Animation Renders and More Sty...

Chaos, the leader in architectural visualisation software, today announces Chaos Corona 13, giving archviz designers new ways to add eye-catching style and flai...

13/06/2025

PALI's Nena Music Video Shot with Blackmagic Design

PALI's Nena Music Video Shot with Blackmagic Design Brie Clayton June 12, 2025 0 Comments Blackmagic Cinema Camera 6K and DaVinci Resolve Studio b...

13/06/2025

OddBeast Powers Up iRobot's Newest Roombas with Suite of CGI Launch Assets

OddBeast Powers Up iRobot's Newest Roombas with Suite of CGI Launch Assets Brie Clayton June 12, 2025 0 Comments The motion design and production ...

13/06/2025

On Chick Coreas Birthday, a Newly Uncovered Archival Release

On Chick Coreas Birthday, a Newly Uncovered Archival Release The Visitors, composed by Corea and performed by vibraphonist Gary Burton and pianist Kirill Gers...

13/06/2025

RT Seeks Expressions of Interest for on-air presenting roles on RT Radio 1 and RT News & Current Affairs Output

In fulfilment of a recommendation by the Government's Expert Advisory Commit...

13/06/2025

SVG Sit-Down: Backblaze's Gleb Budman Talks Products, Partnerships, and the Growth in Cloud Storage

SVG Sit-Down: Backblaze's Gleb Budman Talks Products, Partnerships, and the ...

13/06/2025

SVG Sit-Down: DAZN's Walker Jacobs Calls Streaming the FIFA Club World Cup The Most Ambitious Thing We've Ever Done in the U.S.'

SVG Sit-Down: DAZN's Walker Jacobs Calls Streaming the FIFA Club World Cup ...

13/06/2025

New Sponsor Spotlight: Vecima Networks' Paul Strickland on How Improving QoE for Sports Fans Can Deliver Real ROI

New Sponsor Spotlight: Vecima Networks' Paul Strickland on How Improving QoE...

13/06/2025

Pitch Perspective: Where's Next for Specialty Cameras in Soccer?

Pitch Perspective: Where's Next for Specialty Cameras in Soccer? Leaders from Sky Austria and ACS discuss the possibilities of camera placement pitchside B...