Sony Pixel Power calrec Sony

Math Test? No Problems: NVIDIA Team Scores Kaggle Win With Reasoning Model

15/04/2025

The final days of the AI Mathematical Olympiad's latest competition were a transcontinental relay for team NVIDIA.

Every evening, two team members on opposite ends of the U.S. would submit an AI reasoning model to Kaggle - the online Olympics of data science and machine learning. They'd wait a tense five hours before learning how well the model tackled a sample set of 50 complex math problems.

After seeing the results, the U.S. team would pass the baton to teammates waking up in Armenia, Finland, Germany and Northern Ireland, who would spend their day testing, modifying and optimizing different model versions.

Every night I'd be so disappointed in our score, but then I'd wake up and see the messages that came in overnight from teammates in Europe, said Igor Gitman, senior applied scientist. My hopes would go up and we'd try again.

While the team was disheartened by their lack of improvement on the public dataset during the competition's final days, the real test of an AI model is how well it can generalize to unseen data. That's where their reasoning model leapt to the top of the leaderboard - correctly answering 34 out of 50 Olympiad questions within a five-hour time limit using a cluster of four NVIDIA L4 GPUs.

We got the magic in the end, said Northern Ireland-based team member Darragh Hanley, a Kaggle grandmaster and senior large language model (LLM) technologist.

Building a Winning Equation The NVIDIA team competed under the name NemoSkills - a nod to their use of the NeMo-Skills collection of pipelines for accelerated LLM training, evaluation and inference. The seven members each contributed different areas of expertise, spanning LLM training, model distillation and inference optimization.

For the Kaggle challenge, over 2,200 participating teams submitted AI models tasked with solving 50 math questions - complex problems at the National Olympiad level, spanning algebra, geometry, combinatorics and number theory - within five hours.

https://blogs.nvidia.com/wp-content/uploads/2025/04/Sample-Reasoning-AI.mp4

The team's winning model uses a combination of natural language reasoning and Python code execution.

To complete this inference challenge on the small cluster of NVIDIA L4 GPUs available via Kaggle, the NemoSkills team had to get creative.

Their winning model used Qwen2.5-14B-Base, a foundation model with chain-of-thought reasoning capabilities which the team fine-tuned on millions of synthetically generated solutions to math problems.

These synthetic solutions were primarily generated by two larger reasoning models - DeepSeek-R1 and QwQ-32B - and used to teach the team's foundation model via a form of knowledge distillation. The end result was a smaller, faster, long-thinking model capable of tackling complex problems using a combination of natural language reasoning and Python code execution.

To further boost performance, the team's solution reasons through multiple long-thinking responses in parallel before determining a final answer. To optimize this process and meet the competition's time limit, the team also used an innovative early-stopping technique.

A reasoning model might, for example, be set to answer a math problem 12 different times before picking the most common response. Using the asynchronous processing capabilities of NeMo-Skills and NVIDIA TensorRT-LLM, the team was able to monitor and exit inference early if the model had already converged at the correct answer four or more times.

TensorRT-LLM also enabled the team to harness FP8 quantization, a compression method that resulted in a 1.5x speedup over using the more commonly used FP16 format. ReDrafter, a speculative decoding technique developed by Apple, was used for a further 1.8x speedup.

The final model performed even better on the competition's unseen final dataset than it did on the public dataset - a sign that the team successfully built a generalizable model and avoided overfitting their LLM to the sample data.

Even without the Kaggle competition, we'd still be working to improve AI reasoning models for math, said Gitman. But Kaggle gives us the opportunity to benchmark and discover how well our models generalize to a third-party dataset.

Sharing the Wealth The team will soon release a technical report detailing the techniques used in their winning solution - and plans to share their dataset and a series of models on Hugging Face. The advancements and optimizations they made over the course of the competition have been integrated into NeMo-Skills pipelines available on GitHub.

Key data, technology, and insights from this pipeline were also used to train the just-released NVIDIA Llama Nemotron Ultra model.

Throughout this collaboration, we used tools across the NVIDIA software stack, said Christof Henkel, a member of the Kaggle Grandmasters of NVIDIA, known as KGMON. By working closely with our LLM research and development teams, we're able to take what we learn from the competition on a day-to-day basis and push those optimizations into NVIDIA's open-source libraries.

After the competition win, Henkel regained the title of Kaggle World Champion - ranking No. 1 among the platform's over 23 million users. Another teammate, Finland-based Ivan Sorokin, earned the Kaggle Grandmaster title, held by just over 350 people around the world.

For their first-place win, the group also won a $262,144 prize that they're directing to the NVIDIA Foundation to support charitable organizations.

Meet the full team - Igor Gitman, Darragh Hanley, Christof Henkel, Ivan Moshkov, Benedikt Schifferer, Ivan Sorokin and Shubham Toshniwal - in the video below:

Sample math questions in the featured visual above are from the 2025 American Invitational Mathematics Examination. Find the full set of questions and solutions on the Art
LINK: https://blogs.nvidia.com/blog/reasoning-ai-math-olympiad/...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

31/01/2026

ISE: Sony Electronics Launches BRAVIA Professional Displays BZ-P Series

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

31/01/2026

Fox Sports Details FIFA World Cup 2026 Broadcast Schedule

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

31/01/2026

ISE: LYNX Technik Debuts New 4-Channel 12G-SDI Fiber Converters

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

31/01/2026

PBS Promotes Scott Nourse to Chief Technology Officer

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

2026 Sundance Film Festival Announces Award Winners

Top L-R: The Friend's House is Here, Josephine, The Lake, Bedford Park, Who Killed Alex Odeh? Second Row L-R: Take Me Home, American Pachuco: The Legend of...

30/01/2026

Dnyada her 14 kiiden biri Trk mzii dinliyor

Spotify, Haziran ay sonunda kadar stanbul'da yeni bir ofis a aca n ve T rkiye pazar n y netmek zere yeni bir atama ger ekle tirdi ini duyurdu. Bu kaps...

30/01/2026

Artemis II Wet Dress Rehearsal: Taking Things Into the Home Stretch

The Artemis II wet dress rehearsal will simulate the launch countdown, fully loading fuel and verifying systems ahead of the first SLS and Orion crewed flight....

30/01/2026

Copper Leaf Media Merges With Dimension PR

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

MXL: Aligning Broadcast Production With Software

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

Grass Valley and NETGEAR Partner to Accelerate Enterprise...

Grass Valley , the leading technology provider for live production solutions, and NETGEAR Inc. (NASDAQ: NTGR), a global leader in network solutions, today anno...

30/01/2026

tvONE Appoints Amit Singh as Regional Sales Manager for I...

tvONE, a leading video processor, signal distribution technology and media server developer, announces the expansion of Amit Singh's role to Regional Sales ...

30/01/2026

Mike Aiton Brings Stories to Life with NUGEN Audio

With a career that spans four decades across television, film and post-production, Freelance Sound Designer and Post-production Sound Mixer Mike Aiton has built...

30/01/2026

DPA Showcases its Complete Wireless Microphone Ecosystem...

DPA Microphones will feature its new, fully integrated wireless microphone ecosystem, designed to let audio professionals work faster, cleaner and with total co...

30/01/2026

Ventum Tech and Emergent Announce a Strategic Partnership...

As the Middle East continues to accelerate investment in next-generation media, broadcast, and immersive content technologies, Ventum Tech today announced a str...

30/01/2026

MRMC Broadcast to Highlight High-Precision Motion Control...

Mark Roberts Motion Control (MRMC), a Nikon company and global leader in robotic camera systems, today announced its participation at Integrated Systems Europe ...

30/01/2026

Peacock Hits 44 Million Subs, Lost $552 Million in Q4

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

ATSC Board Leadership Re-Elected For 2026

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

MRC Issues Final Digital Advertising Auction Transparency Standards

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

CBS Atlanta Launches New Weekday Morning News Show with AR/VR Set

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

30/01/2026

Boston Conservatory at Berklee Hosts the National Opera Association's 2026 Conference

Boston Conservatory at Berklee Hosts the National Opera Association's 2026 C...

30/01/2026

Student Spotlight: Sriram Narayanan

Student Spotlight: Sriram Narayanan The classical pianist shares his experience growing up with a language disability and finding his voice through music. Ja...

30/01/2026

2026 Media Industry Trends to Watch: Change is Your Competitive Advantage

Heading into 2026, the pace of change across radio, TV, and digital media is reaching an inflection point. Audience behaviors continue to evolve, measurement mo...

30/01/2026

VEON Partners with MindBridge to Enhance Financial Analytics, Audit and Internal Controls with Augmented Intelligence Capabilities

30 Jan 2026 VEON Partners with MindBridge to Enhance Financial Analytics, Audit...

30/01/2026

Introducing the NEW Techtel.tv! | FEB 5% OFF Offer

Introducing the NEW Techtel.tv! | FEB 5% OFF Offer 30 Jan Written By Suzanne Costello Our Website & Online Store: Now Unified for a Seamless Experience.We&#...

30/01/2026

Easels at the ready! All new judging line up for series 13 of Portrait Artist of the Year

Friday 30 January 2026 Easels at the ready! All new judging line up for series ...

30/01/2026

Britain can switch off terrestrial TV in the 2030s, with targeted support to close the digital divide

Friday 30 January 2026 Britain can switch off terrestrial TV in the 2030s, with...

30/01/2026

The Danish Crime Series The Asset' Returns for a Second Season

Back to All News The Danish Crime Series The Asset' Returns for a Second Season Entertainment 30 January 2026 GlobalDenmark Link copied to clipboard ...

30/01/2026

Building Trust in Retail Media: Lessons from the IAB Europe Agency Breakfast

Two key themes came through strongly: Inconsistent measurement remains a major barrier to comparing performance across Retail Media Networks Independent cer...

29/01/2026

Extension of Invitation to Submit Proposals for Micro-Budget Film Projects 2026 Deadline to 2 February 2026

The National Film and Video Foundation (NFVF), in collaboration with a distribut...

29/01/2026

Hitachi Europe Appoints Michele Fracchiolla as President

Michele Fracchiolla Succeeds Andrew Barr as President of EMEA region from April 1, 2026 London, January 29, 2026 Hitachi Europe Ltd. today announces the appoi...

29/01/2026

L3Harris Technologies Reports Strong Full Year and Fourth Quarter 2025 Results, Initiates 2026 Guidance

MELBOURNE, Fla., January 29, 2026 - L3Harris Technologies (NYSE: LHX) reports fu...

29/01/2026

Nielsen Announces 2025 ARTEY Award Winners Following Record-Breaking Year of Streaming

Bluey' Wins Second Consecutive Top Streaming Title of the Year with 45 Billi...

29/01/2026

Report: Performance TV Ties With Social Media in Driving Ad Results

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

ISE: NDI and OBSBOT Expand Partnership

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

NTCA Asks FCC to Block Nexstar, Tegna Deal

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

FCC Announces Tentative Agenda for February Open Meeting

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

CBS Sports AFC Championship Game Attracts 48.6 Million Viewers

Share Share by: Copy link Facebook X Linkedin Bluesky Email...

29/01/2026

Boston Conservatory Orchestra Presents East Coast Premiere of Peter and Leonardo Dugan Piano Concerto

Boston Conservatory Orchestra Presents East Coast Premiere of Peter and Leonardo...

29/01/2026

Kyivstar Announces Pricing of Secondary Offering of Common Shares Held by VEON

29 Jan 2026 Kyivstar Announces Pricing of Secondary Offering of Common Shares Held by VEON NEW YORK, New York, January 29, 2026 -- VEON Ltd. (Nasdaq: VEON), a ...

29/01/2026

Mercedes-Benz Unveils New S-Class Built on NVIDIA DRIVE AV, Which Enables an L4-Ready Architecture

Mercedes-Benz is marking 140 years of automotive innovation with a new S-Class b...

29/01/2026

X-Rite Pantone Appoints Cindy Cooperman as Vice President and General Manager of Pantone

X-Rite Pantone Appoints Cindy Cooperman as Vice President and General Manager of...

29/01/2026

Outback Terror: The Falconio Murder

New two-part true crime documentary, OUTBACK TERROR: THE FALCONIO MURDER, aims to shed new light on a case that continues to intrigue on both sides of the world...

29/01/2026

'Love is Blind: Sweden' Returns for a Third Season - Premiering on March 12

Back to All News Love is Blind: Sweden Returns for a Third Season - Premiering ...

29/01/2026

Unmask Bridgerton' Season 4 With Our Complete Coverage Guide

Back to All News Unmask Bridgerton' Season 4 With Our Complete Coverage Guide Yerin Ha as Sophie Baek and Luke Thompson as Benedict Bridgerton in Season ...

29/01/2026

Extraordinary Crime Mysteries, Mythical Worlds and High-Stakes Psychological Thrillers: Inside Netflix's 2026 Chinese-Language Slate

Back to All News Extraordinary Crime Mysteries, Mythical Worlds and High-Stakes...

29/01/2026

FOX Sports Unveils Historic FIFA World Cup 2026 Broadcast Schedule

FOX Sports Unveils Historic FIFA World Cup 2026 Broadcast Schedule Monumental Slate Features 340 Hours of Live First-Run Programming Across FOX Sports Platfo...