Sony Pixel Power calrec Sony

The King's Swedish: AI Rewrites the Book in Scandinavia

19/06/2022

If the King of Sweden wants help drafting his annual Christmas speech this year, he could ask the same AI model that's available to his 10 million subjects.

As a test, researchers prompted the model, called GPT-SW3, to draft one of the royal messages, and it did a pretty good job, according to Magnus Sahlgren, who heads research in natural language understanding at AI Sweden, a consortium kickstarting the country's journey into the machine learning era.

Later, our minister of digitalization visited us and asked the model to generate arguments for political positions and it came up with some really clever ones - and he intuitively understood how to prompt the model to generate good text, Sahlgren said.

Early successes inspired work on an even larger and more powerful version of the language model they hope will serve any citizen, company or government agency in Scandinavia.

A Multilingual Model The current version packs 3.6 billion parameters and is smart enough to do a few cool things in Swedish. Sahlgren's team aims to train a state-of-the-art model with a whopping 175 billion parameters that can handle all sorts of language tasks in the Nordic languages of Swedish, Danish, Norwegian and, it hopes, Icelandic, too.

For example, a startup can use it to automatically generate product descriptions for an e-commerce website given only the products' names. Government agencies can use it to quickly classify and route questions from citizens.

Companies can ask it to rapidly summarize reports so they can react fast. Hospitals can run distilled versions of the model privately on their own systems to improve patient care.

It's a foundational model we will provide as a service for whatever tasks people want to solve, said Sahlgren, who's been working at the intersection of language and machine learning since he earned his Ph.D. in computational linguistics in 2006.

Permission to Speak Freely It's a capability increasingly seen as a strategic asset, a keystone of digital sovereignty in a world that speaks thousands of languages across nearly 200 countries.

Most language services today focus on Chinese or English, the world's two most-spoken tongues. They're typically created in China or the U.S., and they aren't free.

It's important for us to have models built in Sweden for Sweden, Sahlgren said.

Small Team, Super System We're a small country and a core team of about six people, yet we can build a state-of-the-art resource like this for people to use, he added.

That's because Sweden has a powerful engine in BerzeLiUs, a 300-petaflops AI supercomputer at Link ping University. It trained the initial GPT-SW3 model using just 16 of the 60 nodes in the NVIDIA DGX SuperPOD.

The next model may exercise all the system's nodes. Such super-sized jobs require super software like the NVIDIA NeMo Megatron framework.

It lets us scale our training up to the full supercomputer, and we've been lucky enough to have access to experts in the NeMo development team - without NVIDIA it would have been so much more complicated to come this far, he said.

A Workflow for Any Language NVIDIA's engineers created a recipe based on NeMo and an emerging process called p-tuning that optimizes massive models fast, and it's geared to work with any language.

In one early test, a model nearly doubled its accuracy after NVIDIA engineers applied the techniques.

Magnus Sahlgren What's more, it requires one-tenth the data, slashing the need for tens of thousands of hand-labeled records. That opens the door for users to fine-tune a model with the relatively small, industry-specific datasets they have at hand.

We hope to inspire a lot of entrepreneurship in industry, startups and the public using our technology to develop their own apps and services, said Sahlgren.

Writing the Next Chapter Meanwhile, NVIDIA's developers are already working on ways to make the enabling software better.

One test shows great promise for training new capabilities using widely available English datasets into models designed for any language. In another effort, they're using the p-tuning techniques in inference jobs so models can learn on the fly.

Zenodia Charpy, a senior solutions architect at NVIDIA based in Gothenburg, shares the enthusiasm of the AI Sweden team she supports. We've only just begun trying new and better methods to tackle these large language challenges - there's much more to come, she said.

The GPT-SW3 model will be made available by the end of year via an early access program. To apply, contact francisca.hoyer@ai.se.
LINK: https://blogs.nvidia.com/blog/2022/06/19/ai-sweden-nlp/...
See more stories from nvidia

More from Nvidia

25/04/2024

Into the Omniverse: Unlocking the Future of Manufacturing With OpenUSD on Siemens Teamcenter X

Editor's note: This post is part of Into the Omniverse, a series focused on ...

25/04/2024

Blast From the Past: Stream StarCraft' and Diablo' on GeForce NOW

Support for Battle.net on GeForce NOW expands this GFN Thursday, as titles from the iconic StarCraft and Diablo series come to the cloud. StarCraft Remastered,...

24/04/2024

Rays Up: Decoding AI-Powered DLSS 3.5 Ray Reconstruction

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

24/04/2024

Forecasting the Future: AI2's Christopher Bretherton Discusses Using Machine Learning for Climate Modeling

Can machine learning help predict extreme weather events and climate change? Chr...

24/04/2024

NVIDIA to Acquire GPU Orchestration Software Provider Run:ai

To help customers make more efficient use of their AI computing resources, NVIDIA today announced it has entered into a definitive agreement to acquire Run:ai, ...

24/04/2024

How Virtual Factories Are Making Industrial Digitalization a Reality

To address the shift to electric vehicles, increased semiconductor demand, manufacturing onshoring, and ambitions for greater sustainability, manufacturers are ...

23/04/2024

Small and Mighty: NVIDIA Accelerates Microsoft's Open Phi-3 Mini Language Models

NVIDIA announced today its acceleration of Microsoft's new Phi-3 Mini open l...

22/04/2024

Climate Tech Startups Integrate NVIDIA AI for Sustainability Applications

Whether they're monitoring miniscule insects or delivering insights from satellites in space, NVIDIA-accelerated startups are making every day Earth Day. S...

18/04/2024

Wide Open: NVIDIA Accelerates Inference on Meta Llama 3

NVIDIA today announced optimizations across all its platforms to accelerate Meta Llama 3, the latest generation of the large language model (LLM). The open mod...

18/04/2024

Up to No Good: No Rest for the Wicked' Early Access Launches on GeForce NOW

It's time to get a little wicked. Members can now stream No Rest for the Wicked from the cloud. It leads six new games joining the GeForce NOW library of m...

18/04/2024

NVIDIA Honors Partners of the Year in Europe, Middle East, Africa

NVIDIA today recognized 18 partners in Europe, the Middle East and Africa for their achievements and commitment to driving AI adoption. The recipients were hon...

17/04/2024

Seeing Beyond: Living Optics CEO Robin Wang on Democratizing Hyperspectral Imaging

Step into the realm of the unseen with Robin Wang, CEO of Living Optics. The sta...

17/04/2024

Moving Pictures: Transform Images Into 3D Scenes With NVIDIA Instant NeRF

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

16/04/2024

New NVIDIA RTX A400 and A1000 GPUs Enhance AI-Powered Design and Productivity Workflows

AI integration across design and productivity applications is becoming the new s...

16/04/2024

To Cut a Long Story Short: Video Editors Benefit From DaVinci Resolve's New AI Features Powered by RTX

Editor's note: This post is part of our In the NVIDIA Studio series, which c...

15/04/2024

AI Is Tech's Greatest Contribution to Social Elevation,' NVIDIA CEO Tells Oregon State Students

AI promises to bring the full benefits of the digital revolution to billions acr...

10/04/2024

The Building Blocks of AI: Decoding the Role and Significance of Foundation Models

Editor's note: This post is part of the AI Decoded series, which demystifies...

10/04/2024

Combating Corruption With Data: Cleanlab and Berkeley Research Group on Using AI-Powered Investigative Analytics

Talk about scrubbing data. Curtis Northcutt, cofounder and CEO of Cleanlab, and ...

09/04/2024

NVIDIA Joins $110 Million Partnership to Help Universities Teach AI Skills

The Biden Administration has announced a new $110 million AI partnership between Japan and the United States that includes an initiative to fund research throug...

09/04/2024

Broadcasting Breakthroughs: NVIDIA Holoscan for Media, Available Now, Transforms Live Media With Easy AI Integration

Whether delivering live sports programming, streaming services, network broadcas...

09/04/2024

Start Up Your Engines: NVIDIA and Google Cloud Collaborate to Accelerate AI Development

NVIDIA and Google Cloud have announced a new collaboration to help startups arou...

04/04/2024

NVIDIA Ranked by Fortune at No. 3 on 100 Best Companies to Work For' List

NVIDIA jumped to No. 3 on the latest list of America's 100 Best Companies to Work For by Fortune magazine and Great Place to Work. It's the company'...

04/04/2024

The Elder Scrolls Online' Joins GeForce NOW for Game's 10th Anniversary

Rain or shine, a new month means new games. GeForce NOW kicks off April with nearly 20 new games, seven of which are available to play this week. GFN Thursday ...

03/04/2024

A New Lens: Dotlumen CEO Cornel Amariei on Assistive Technology for the Visually Impaired

Dotlumen is illuminating a new technology to help people with visual impairments...

03/04/2024

Coming Up ACEs: Decoding the AI Technology That's Enhancing Games With Realistic Digital Humans

Editor's note: This post is part of the AI Decoded series, which demystifies...

28/03/2024

Greater Scope: Doctors Get Inside Look at Gut Health With AI-Powered Endoscopy

From humble beginnings as a university spinoff to an acquisition by the leading global medtech company in its field, Odin Vision has been on an accelerated jour...

28/03/2024

Get Cozy With Palia' on GeForce NOW

Ease into spring with the warm, cozy vibes of Palia, coming to the cloud this GFN Thursday. It's part of six new titles joining the GeForce NOW library of ...

27/03/2024

Software Developers Launch OpenUSD and Generative AI-Powered Product Configurators Built on NVIDIA Omniverse

From designing dream cars to customizing clothing, 3D product configurators are ...

27/03/2024

NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf

It's official: NVIDIA delivered the world's fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks...

27/03/2024

Viome's Guru Banavar Discusses AI for Personalized Health

In the latest episode of NVIDIA's AI Podcast, Viome Chief Technology Officer Guru Banavar spoke with host Noah Kravitz about how AI and RNA sequencing are r...

27/03/2024

Unlocking Peak Generations: TensorRT Accelerates AI on RTX PCs and Workstations

Editor's note: This post is part of the AI Decoded series, which demystifies AI by making the technology more accessible, and which showcases new hardware, ...

26/03/2024

Boom in AI-Enabled Medical Devices Transforms Healthcare

The future of healthcare is software-defined and AI-enabled. Around 700 FDA-cleared, AI-enabled medical devices are now on the market - more than 10x the number...

26/03/2024

Model Innovators: How Digital Twins Are Making Industries More Efficient

A manufacturing plant near Hsinchu, Taiwan's Silicon Valley, is among facilities worldwide boosting energy efficiency with AI-enabled digital twins. A virt...

26/03/2024

Into the Omniverse: Groundbreaking OpenUSD Advancements Put NVIDIA GTC Spotlight on Developers

Editor's note: This post is part of Into the Omniverse, a series focused on ...

25/03/2024

NVIDIA Blackwell and Automotive Industry Innovators Dazzle at NVIDIA GTC

Generative AI, in the data center and in the car, is making vehicle experiences safer and more enjoyable. The latest advancements in automotive technology were...

21/03/2024

AI's New Frontier: From Daydreams to Digital Deeds

Imagine a world where you can whisper your digital wishes into your device, and poof, it happens. That world may be coming sooner than you think. But if you...

21/03/2024

You Transformed the World,' NVIDIA CEO Tells Researchers Behind Landmark AI Paper

Of GTC's 900+ sessions, the most wildly popular was a conversation hosted by...

21/03/2024

Instant Latte: NVIDIA Gen AI Research Brews 3D Shapes in Under a Second

NVIDIA researchers have pumped a double shot of acceleration into their latest text-to-3D generative AI model, dubbed LATTE3D. Like a virtual 3D printer, LATTE...

21/03/2024

Here Be Dragons: Dragon's Dogma 2' Comes to GeForce NOW

Arise for a new adventure with Dragon's Dogma 2, leading two new titles joining the GeForce NOW library this week. Set Forth, Arisen Fulfill a forgotten de...

20/03/2024

AI Decoded From GTC: The Latest Developer Tools and Apps Accelerating AI on PC and Workstation

Editor's note: This post is part of the AI Decoded series, which demystifies...

19/03/2024

NVIDIA Celebrates Americas Partners Driving AI-Powered Transformation

NVIDIA recognized 14 partners in the Americas for their achievements in transforming businesses with AI, this week at GTC. The winners of the NVIDIA Partner Ne...

19/03/2024

Climate Pioneers: 3 Startups Harnessing NVIDIA's AI and Earth-2 Platforms

To help mitigate climate change - one of humanity's greatest challenges - researchers are turning to AI and sustainable computing to accelerate and operatio...

19/03/2024

Secure by Design: NVIDIA AIOps Partner Ecosystem Blends AI for Businesses

In today's complex business environments, IT teams face a constant flow of challenges, from simple issues like employee account lockouts to critical securit...

19/03/2024

Generation Sensation: New Generative AI and RTX Tools Boost Content Creation

Editor's note: This post is part of our In the NVIDIA Studio series, which celebrates featured artists, offers creative tips and tricks, and demonstrates ho...

19/03/2024

NVIDIA, Huang Win Top Honors in Innovation, Engineering

NVIDIA today was named the world's most innovative company by Fast Company magazine. The accolade comes on the heels of company founder and CEO Jensen Huan...

18/03/2024

NVIDIA Edify Unlocks 3D Generative AI, New Image Controls for Visual Content Providers

NVIDIA Edify, a multimodal architecture for visual generative AI, is entering a ...

18/03/2024

From Atoms to Supercomputers: NVIDIA, Partners Scale Quantum Computing

The latest advances in quantum computing include investigating molecules, deploying giant supercomputers and building the quantum workforce with a new academic ...

18/03/2024

New NVIDIA Storage Partner Validation Program Streamlines Enterprise AI Deployments

A sharp increase in generative AI deployments is driving business innovation for...

18/03/2024

NVIDIA Unveils Digital Blueprint for Building Next-Gen Data Centers

Designing, simulating and bringing up modern data centers is incredibly complex, involving multiple considerations like performance, energy efficiency and scala...

18/03/2024

Generative AI Developers Harness NVIDIA Technologies to Transform In-Vehicle Experiences

Cars of the future will be more than just modes of transportation; they'll b...