Sony Pixel Power calrec Sony

Speak Like a Native: NVIDIA Parlays Win in Voice Challenge

14/02/2024

Thanks to their work driving AI forward, Akshit Arora and Rafael Valle could someday speak to their spouses' families in their native languages.

Arora and Valle - along with colleagues Sungwon Kim and Rohan Badlani - won the LIMMITS '24 challenge which asks contestants to recreate in real time a speaker's voice in English or any of six languages spoken in India with the appropriate accent. Their novel AI model only required a three-second speech sample.

The NVIDIA team advanced the state of the art in an emerging field of personalized voice interfaces for more than a billion native speakers of Bengali, Chhattisgarhi, Hindi, Kannada, Marathi and Telugu.

Making Voice Interfaces Realistic The technology for personalized text-to-speech translation is a work in progress. Existing services sometimes fail to accurately reflect the accents of the target language or nuances of the speaker's voice.

The challenge judged entries by listening for the naturalness of models' resulting speech and its similarity to the original speaker's voice.

The latest improvements promise personalized, realistic conversations and experiences that break language barriers. Broadcasters, telcos, universities, as well as e-commerce and online gaming services are eager to deploy such technology to create multilingual movies, lectures and virtual agents.

We demonstrated we can do this at a scale not previously seen, said Arora, who has two uses close to his heart.

Breaking Down Linguistic Barriers A senior data scientist who supports one of NVIDIA's biggest customers, Arora speaks Punjabi, while his wife and her family are native Tamil speakers.

It's a gulf he's long wanted to bridge for himself and others. I had classmates who knew their native languages much better than the Hindi and English used in school, so they struggled to understand class material, he said.

The gulf crosses continents for Valle, a native of Brazil whose wife and family speak Gujarati, a language popular in west India.

It's a problem I face every day, said Valle, an AI researcher with degrees in computer music and machine listening and improvisation. We've tried many products to help us have clearer conversations.

Badlani, an AI researcher, said living in seven different Indian states, each with its own popular language, inspired him to work in the field.

A Race to the Finish Line The initiative started nearly two years ago when Arora and Badlani formed the four-person team to work on the very different version of the challenge that would be held in 2023.

Their efforts generated a working code base for the so-called Indic languages. But getting to the win announced in January required a full-on sprint because the 2024 challenge didn't get on the team's radar until 15 days before the deadline.

Luckily, Kim, a deep learning researcher in NVIDIA's Seoul office, had been working for some time on an AI model well suited to the challenge.

A specialist in text-to-speech voice synthesis, Kim was designing a so-called P-Flow model prior to starting his second internship at NVIDIA in 2023. P-Flow models borrow the technique large language models employ of using short voice samples as prompts so they can respond to new inputs without retraining.

I created the model for English, but we were able to generalize it for any language, he said.

We were talking and texting about this model even before he started at NVIDIA, said Valle, who mentored Kim in two internships before he joined full time in January.

Giving Others a Voice P-Flow will soon be part of NVIDIA Riva, a framework for building multilingual speech and translation AI software, included in the NVIDIA AI Enterprise software platform.

The new capability will let users deploy the technology inside their data centers, on personal systems or in public or private cloud services. Today, voice translation services typically run on public cloud services.

I hope our customers are inspired to try this technology, Arora said. I enjoy being able to showcase in challenges like this one the work we do every day.

The contest is part of an initiative to develop open-source datasets and AI models for nine languages most widely spoken in India.

Hear Arora and Badlani share their experiences in a session at GTC next month.

And listen to the results of the team's model below, starting with a three-second sample of a native Kannada speaker:

https://blogs.nvidia.com/wp-content/uploads/2024/02/pr_kannada_f_indictts_prompt_3s-1.mp3

Here's a similar-sounding synthesized voice reading the first sentence of this blog in Hindi:

https://blogs.nvidia.com/wp-content/uploads/2024/02/pr_kannada_f_indictts_speaking_hindi_3-2.mp3

And then in English:

https://blogs.nvidia.com/wp-content/uploads/2024/02/pr_kannada_f_indictts_speaking_english-1.mp3 See notice regarding software product information.
LINK: https://blogs.nvidia.com/blog/generative-voice-challenge/...
See more stories from nvidia

More from Nvidia

22/04/2024

Climate Tech Startups Integrate NVIDIA AI for Sustainability Applications

Whether they're monitoring miniscule insects or delivering insights from satellites in space, NVIDIA-accelerated startups are making every day Earth Day. S...

18/04/2024

Wide Open: NVIDIA Accelerates Inference on Meta Llama 3

NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM). The open mod...

18/04/2024

Up to No Good: No Rest for the Wicked' Early Access Launches on GeForce NOW

It's time to get a little wicked. Members can now stream No Rest for the Wicked from the cloud. It leads six new games joining the GeForce NOW library of m...

18/04/2024

NVIDIA Honors Partners of the Year in Europe, Middle East, Africa

NVIDIA today recognized 18 partners in Europe, the Middle East and Africa for their achievements and commitment to driving AI adoption. The recipients were hon...

17/04/2024

Seeing Beyond: Living Optics CEO Robin Wang on Democratizing Hyperspectral Imaging

Step into the realm of the unseen with Robin Wang, CEO of Living Optics. The sta...

17/04/2024

Moving Pictures: Transform Images Into 3D Scenes With NVIDIA Instant NeRF

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

16/04/2024

New NVIDIA RTX A400 and A1000 GPUs Enhance AI-Powered Design and Productivity Workflows

AI integration across design and productivity applications is becoming the new s...

16/04/2024

To Cut a Long Story Short: Video Editors Benefit From DaVinci Resolve's New AI Features Powered by RTX

Editor's note: This post is part of our In the NVIDIA Studio series, which c...

15/04/2024

AI Is Tech's Greatest Contribution to Social Elevation,' NVIDIA CEO Tells Oregon State Students

AI promises to bring the full benefits of the digital revolution to billions acr...

10/04/2024

The Building Blocks of AI: Decoding the Role and Significance of Foundation Models

Editor's note: This post is part of the AI Decoded series, which demystifies...

10/04/2024

Combating Corruption With Data: Cleanlab and Berkeley Research Group on Using AI-Powered Investigative Analytics

Talk about scrubbing data. Curtis Northcutt, cofounder and CEO of Cleanlab, and ...

09/04/2024

NVIDIA Joins $110 Million Partnership to Help Universities Teach AI Skills

The Biden Administration has announced a new $110 million AI partnership between Japan and the United States that includes an initiative to fund research throug...

09/04/2024

Broadcasting Breakthroughs: NVIDIA Holoscan for Media, Available Now, Transforms Live Media With Easy AI Integration

Whether delivering live sports programming, streaming services, network broadcas...

09/04/2024

Start Up Your Engines: NVIDIA and Google Cloud Collaborate to Accelerate AI Development

NVIDIA and Google Cloud have announced a new collaboration to help startups arou...

04/04/2024

NVIDIA Ranked by Fortune at No. 3 on 100 Best Companies to Work For' List

NVIDIA jumped to No. 3 on the latest list of America's 100 Best Companies to Work For by Fortune magazine and Great Place to Work. It's the company'...

04/04/2024

The Elder Scrolls Online' Joins GeForce NOW for Game's 10th Anniversary

Rain or shine, a new month means new games. GeForce NOW kicks off April with nearly 20 new games, seven of which are available to play this week. GFN Thursday ...

03/04/2024

A New Lens: Dotlumen CEO Cornel Amariei on Assistive Technology for the Visually Impaired

Dotlumen is illuminating a new technology to help people with visual impairments...

03/04/2024

Coming Up ACEs: Decoding the AI Technology That's Enhancing Games With Realistic Digital Humans

Editor's note: This post is part of the AI Decoded series, which demystifies...

28/03/2024

Greater Scope: Doctors Get Inside Look at Gut Health With AI-Powered Endoscopy

From humble beginnings as a university spinoff to an acquisition by the leading global medtech company in its field, Odin Vision has been on an accelerated jour...

28/03/2024

Get Cozy With Palia' on GeForce NOW

Ease into spring with the warm, cozy vibes of Palia, coming to the cloud this GFN Thursday. It's part of six new titles joining the GeForce NOW library of ...

27/03/2024

Software Developers Launch OpenUSD and Generative AI-Powered Product Configurators Built on NVIDIA Omniverse

From designing dream cars to customizing clothing, 3D product configurators are ...

27/03/2024

NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf

It's official: NVIDIA delivered the world's fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks...

27/03/2024

Viome's Guru Banavar Discusses AI for Personalized Health

In the latest episode of NVIDIA's AI Podcast, Viome Chief Technology Officer Guru Banavar spoke with host Noah Kravitz about how AI and RNA sequencing are r...

27/03/2024

Unlocking Peak Generations: TensorRT Accelerates AI on RTX PCs and Workstations

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

26/03/2024

Boom in AI-Enabled Medical Devices Transforms Healthcare

The future of healthcare is software-defined and AI-enabled. Around 700 FDA-cleared, AI-enabled medical devices are now on the market - more than 10x the number...

26/03/2024

Model Innovators: How Digital Twins Are Making Industries More Efficient

A manufacturing plant near Hsinchu, Taiwan's Silicon Valley, is among facilities worldwide boosting energy efficiency with AI-enabled digital twins. A virt...

26/03/2024

Into the Omniverse: Groundbreaking OpenUSD Advancements Put NVIDIA GTC Spotlight on Developers

Editor's note: This post is part of Into the Omniverse, a series focused on ...

25/03/2024

NVIDIA Blackwell and Automotive Industry Innovators Dazzle at NVIDIA GTC

Generative AI, in the data center and in the car, is making vehicle experiences safer and more enjoyable. The latest advancements in automotive technology were...

21/03/2024

AI's New Frontier: From Daydreams to Digital Deeds

Imagine a world where you can whisper your digital wishes into your device, and poof, it happens. That world may be coming sooner than you think. But if you...

21/03/2024

You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper

Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by...

21/03/2024

Instant Latte: NVIDIA Gen AI Research Brews 3D Shapes in Under a Second

NVIDIA researchers have pumped a double shot of acceleration into their latest text-to-3D generative AI model, dubbed LATTE3D. Like a virtual 3D printer, LATTE...

21/03/2024

Here Be Dragons: Dragon's Dogma 2' Comes to GeForce NOW

Arise for a new adventure with Dragon's Dogma 2, leading two new titles joining the GeForce NOW library this week. Set Forth, Arisen Fulfill a forgotten de...

20/03/2024

AI Decoded From GTC: The Latest Developer Tools and Apps Accelerating AI on PC and Workstation

Editor's note: This post is part of the AI Decoded series, which demystifies...

19/03/2024

NVIDIA Celebrates Americas Partners Driving AI-Powered Transformation

NVIDIA recognized 14 partners in the Americas for their achievements in transforming businesses with AI, this week at GTC. The winners of the NVIDIA Partner Ne...

19/03/2024

Climate Pioneers: 3 Startups Harnessing NVIDIA's AI and Earth-2 Platforms

To help mitigate climate change - one of humanity's greatest challenges - researchers are turning to AI and sustainable computing to accelerate and operatio...

19/03/2024

Secure by Design: NVIDIA AIOps Partner Ecosystem Blends AI for Businesses

In today's complex business environments, IT teams face a constant flow of challenges, from simple issues like employee account lockouts to critical securit...

19/03/2024

Generation Sensation: New Generative AI and RTX Tools Boost Content Creation

Editor's note: This post is part of our In the NVIDIA Studio series, which celebrates featured artists, offers creative tips and tricks, and demonstrates ho...

19/03/2024

NVIDIA, Huang Win Top Honors in Innovation, Engineering

NVIDIA today was named the world's most innovative company by Fast Company magazine. The accolade comes on the heels of company founder and CEO Jensen Huan...

18/03/2024

NVIDIA Edify Unlocks 3D Generative AI, New Image Controls for Visual Content Providers

NVIDIA Edify, a multimodal architecture for visual generative AI, is entering a ...

18/03/2024

From Atoms to Supercomputers: NVIDIA, Partners Scale Quantum Computing

The latest advances in quantum computing include investigating molecules, deploying giant supercomputers and building the quantum workforce with a new academic ...

18/03/2024

New NVIDIA Storage Partner Validation Program Streamlines Enterprise AI Deployments

A sharp increase in generative AI deployments is driving business innovation for...

18/03/2024

NVIDIA Unveils Digital Blueprint for Building Next-Gen Data Centers

Designing, simulating and bringing up modern data centers is incredibly complex, involving multiple considerations like performance, energy efficiency and scala...

18/03/2024

Generative AI Developers Harness NVIDIA Technologies to Transform In-Vehicle Experiences

Cars of the future will be more than just modes of transportation; they'll b...

18/03/2024

All Eyes on AI: Automotive Tech on Full Display at GTC 2024

All eyes across the auto industry are on GTC - the global AI conference running in San Jose, Calif., and online through Thursday, March 21 - as the world's ...

18/03/2024

All Aboard: NVIDIA Scores 23 World Records for Route Optimization

With nearly two dozen world records to its name, NVIDIA cuOpt now holds the top spot for 100% of the largest routing benchmarks in the last three years. And thi...

18/03/2024

We Created a Processor for the Generative AI Era,' NVIDIA CEO Says

Generative AI promises to revolutionize every industry it touches - all that's been needed is the technology to meet the challenge. NVIDIA founder and CEO ...

14/03/2024

NVIDIA GTC 2024: A Glimpse Into the Future of AI With Jensen Huang

NVIDIA's GTC 2024 AI conference will set the stage for another leap forward in AI. At the heart of this highly anticipated event: the opening keynote by Je...

14/03/2024

Reach for the Stars: Eight Out-of-This-World Games Join the Cloud

The stars align this GFN Thursday as more top titles from Ubisoft and Square Enix join the cloud. Star Wars Outlaws will be coming to the GeForce NOW library a...

13/03/2024

Currents of Change: ITIF President Daniel Castro on Energy-Efficient AI and Climate Change

AI-driven change is in the air, as are concerns about the technology's envir...

13/03/2024

AI Decoded: Demystifying Large Language Models, the Brains Behind Chatbots

Editor's note: This post is part of our AI Decoded series, which aims to demystify AI by making the technology more accessible, while showcasing new hardwar...