
Pinterest has engineered a way to serve its photo-sharing community more of the images they love.
The social-image service, with more than 400 million monthly active users, has trained bigger recommender models for improved accuracy at predicting people's interests.
Pinterest handles hundreds of millions of user requests an hour on any given day. And it must also narrow down relevant images from roughly 300 billion images on the site to roughly 50 for each person.
The last step - ranking the most relevant and engaging content for everyone using Pinterest - required a leap in acceleration to run heftier models, with minimal latency, for better predictions.
Pinterest has improved the accuracy of its recommender models powering people's home feeds and other areas, increasing engagement by as much as 16%.
The leap was enabled by switching from CPUs to NVIDIA GPUs, which could easily be applied next to other areas, including advertising images, according to Pinterest.
Normally we would be happy with a 2% increase, and 16% is just a beginning for home feeds. We see additional gains - it opens a lot of doors for opportunities, said Pong Eksombatchai, a software engineer at Pinterest.
Transformer models capable of better predictions are shaking up industries from retail to entertainment and advertising. But their leaps in performance gains of the past few years have come with a need to serve models that are some 100x bigger as their number of model parameters and computations skyrockets.
Huge Inference Gains, Same Infrastructure Cost Like many, Pinterest engineers wanted to tap into state-of-the-art recommender models to increase engagement. But serving these massive models on CPUs presented a 100x increase in cost and latency. That wasn't going to maintain its magical user experience - fresh and more appealing images - occurring within a fraction of a second.
If that latency happened, then obviously our users wouldn't like that very much because they would have to wait forever, said Eksombatchai. We are pretty close to the limit of what we can do on CPU basically.
The challenge was to serve these hundredfold larger recommender models within the same cost and latency constraints.
Working with NVIDIA, Pinterest engineers began architectural changes to optimize their inference pipeline and recommender models to enable the transition from CPU to GPU cloud instances. The technology transition began late last year and required major changes to how the company manages workloads. The result is a 100x gain in inference efficiency on the same IT budget, meeting their goals.
We are starting to use really, really big models now. And that is where the GPU comes in - to help make these models possible, Eksombatchai said.
Tapping Into cuCollections Switching from CPUs to GPUs required rethinking its inference systems architecture. Among other issues, engineers had to change how they send workloads to their inference servers. Fortunately, there are tools to assist in making the transition easier.
The Pinterest inference server built for CPUs had to be altered because it was set up to send smaller batch sizes to its servers. GPUs can handle much larger workloads, so it's necessary to set up larger batch requests to increase efficiency.
One area where this comes into play is with its embedding table lookup module. Embedding tables are used to track interactions between various context-specific features and interests of user profiles. They can track where you navigate, and what people Pin on Pinterest, share or numerous other actions, helping refine predictions on what users might like to click on next.
They are used to incrementally learn user preference based on context in order to make better content recommendations to those using Pinterest. Its embedding table lookup module required two computation steps repeated hundreds of times because of the number of features tracked.
Pinterest engineers greatly reduced this number of operations using a GPU-accelerated concurrent hash table from NVIDIA cuCollections. And they set up a custom consolidated embedding lookup module so they could merge requests into a single lookup. Better results were seen immediately.
Using cuCollections helped us to remove bottlenecks, said Eksombatchai.
Enlisting CUDA Graphs Pinterest relied on CUDA Graphs to eliminate what was remaining of the small batch operations, further optimizing its inference models.
CUDA Graphs helps reduce the CPU interactions when launching on GPUs. They're designed to enable workloads to be defined as graphs rather than single operations. They provide a mechanism to launch multiple GPU operations through a single CPU operation, reducing CPU overheads.
Pinterest enlisted CUDA Graphs to represent the model inference process as a static graph of operation instead of as those individually scheduled. This enabled the computation to be handled as a single unit without any kernel launching overhead.
The company now supports CUDA Graph as a new backend of its model server. When a model is first loaded, the model server runs the model inference once to build the graph instance. This graph can then be run repeatedly in inference to show content on its app or site.
Implementing CUDA Graphs helped Pinterest to significantly reduce inference latency of its recommender models, according to its engineers.
GPUs have enabled Pinterest to do something that was impossible with CPUs on the same budget, and by doing this they can make changes that have a direct impact on various business metrics.
Learn about Pinterest's GPU-driven inference and optimizations at its GTC session, Serving 100x Bigger Recommender Models, and in the Pinterest Engineering blog.
Register for GTC, running Sept. 19-22, for free to attend sessions with NVIDIA and dozens of industry leaders.
Most recent headlines
09/11/2025
Dalet today announced a transformative leap forward for media operations: Agentic Artificial Intelligence (AI) that unifies the Dalet ecosystem under one natura...
06/10/2025
France T l visions, France's leading broadcaster, has received the 2025 EBU ...
16/09/2025
At DSEI 2025, James Dunne of L3Harris Maritime UK chaired a panel on aligning the supply chain to the warfighter, where leaders discussed modernising support fo...
16/09/2025
Calrec has strengthened its collaboration with audio metering expert RTW by integrating RTW's new TMxCore metering platform across its full range of Argo IP...
16/09/2025
College Football Scores Top Telecast in August with 16M+ Viewers on FOX, Followe...
16/09/2025
Collaboration marks the first SSP integration of Gracenote IDs, enabling show-le...
16/09/2025
AMSTERDAM The organizers of IBC2025 are reporting that 43,858 visitors from more than 170 countries attended the event, which had more than 1,300 exhibitors and...
16/09/2025
Wooden Camera announces the release of its new Accessory Collection for the FUJIFILM GFX ETERNA 55. The highlights of this collection include vital power soluti...
16/09/2025
Anton/Bauer, a leading manufacturer of mobile power solutions for broadcast and cinematic equipment, has announced the launch of Anton/Bauer Fleet Management, a...
16/09/2025
Teradek, a leading provider of video transmission and live production solutions, today announced the launch of Prism Jetpack, a groundbreaking 5G video contribu...
16/09/2025
Astera, the leader in wireless LED lighting solutions, announces the ultra-versatile SolaBulb. Building on the success of the Astera bulb family, SolaBulb intro...
16/09/2025
As the world gathered at TED2025 to explore the provocative theme "Humanity Reimagined", Clear-Com , supported by NETGEAR networking infrastructure, delivered f...
16/09/2025
Bitfocus' Buttons platform celebrated as a catalyst for AV and broadcast convergence
Bitfocus has been named winner of the IABM Impact Award Pro AV Chan...
16/09/2025
Record entries and outstanding innovation celebrated at IBC2025 as the MediaTech community honors its leading people, projects and organizations
IABM announced...
16/09/2025
Harmonic (NASDAQ: HLIT) today announced that SKY Perfect JSAT Corporation (SJC), a leading satellite operator and pay-TV provider in Japan, has partnered with H...
16/09/2025
Telestream congratulates Sky Group, which has been awarded the prestigious IBC Innovation Award for Content Distribution for its MediaMesh platform on Sunday, S...
16/09/2025
Leading space solutions company will use optical ground stations to deliver faster, more secure data from space
Luxembourg, September 15, 2025 - SES, a leading...
16/09/2025
NOVI, Mich. ENCO has unveiled Raptor, a cloud-based live streaming captioning encoder that injects the speed, power and reliability of real-time AI capabilities...
16/09/2025
NEW YORK Comcast NBCUniversal and NBCUniversal Local have announced that a total of $2.5 million has been awarded in 2025 to 69 nonprofit organizations servicin...
16/09/2025
AMSTERDAM Martin Euredjian has joined Atomos as director of advanced imaging and will lead innovation for advanced display technology....
16/09/2025
AMSTERDAM Calrec introduced a 48-faced Argo M and showcased its largest Argo software updates at the recently concluded IBC2025....
16/09/2025
AMSTERDAM Lawo introduced its HOME Audio Shuffler app, a replacement for a traditional baseband audio matrix within an IP-based Dynamic Media Facility, during t...
16/09/2025
CBS is reporting that the 77TH Emmy Awards hosted by Nate Bargatze on Sunday Sept. 14 was seen by more than 7.42 million viewers on the CBS Television Network a...
16/09/2025
Back to All News
Meet the Adversaries That Will Clash in Rulers of Fortune, a S...
16/09/2025
Back to All News
Would You Take 30 Million Baht from a Dead Persons Account? E...
16/09/2025
Back to All News
Netflix Unveils the Official Trailer and Poster for She Walks ...
16/09/2025
Comscore Unveils The Scoreboard: An Interactive Destination Surfacing Consumer B...
15/09/2025
Steve Zahn, Winona Ryder, Ethan Hawke, and Janeane Garofalo star in Ben Stiller&...
15/09/2025
Global K-Pop sensation aespa is redefining what it means to be rich with the r...
15/09/2025
Every day, millions of people around the world turn to Spotify to enjoy the audi...
15/09/2025
After months of intensive planning and implementation, Brembo SGL Carbon Ceramic...
15/09/2025
A U.S. Marine launches a Javelin shoulder-fired anti-tank missile during a training exercise. (Photo credit: U.S. Marine Corps)...
15/09/2025
AMSTERDAM TV Tech has named its Best of Show Awards winners for IBC2025, which wraps up today. Entrants were judged by a panel of industry experts on the criter...
15/09/2025
Unique sports content orchestration platform builds momentum among SES's cus...
15/09/2025
Back to All News
Over 41 Million Global Viewers On Netflix Watch Terence Crawfo...
15/09/2025
Paris, France - September 15, 2025 -- Space42 (ADX: SPACE42), the UAE-based AI-p...
15/09/2025
Back to All News
The Great Flood' Teaser Trailer Previews A Gripping Struggle for Survival
Entertainment
15 September 2025
GlobalSouth Korea
Link copi...
15/09/2025
Back to All News
From Hardship to Hope: Typhoon Family' Presents a Tale of...
15/09/2025
Back to All News
Gloom Goes Global as Wednesday's' The Doom Tour Hit...
15/09/2025
-- Opens door to growth in renewable energy New Delhi, India - 15th September -- Global business and industry leaders from around the world are joining technol...
14/09/2025
Partnership to address business and technical challenges of DMF adoption
he Advanced Media Workflow Association (AMWA) and the European Broadcasting Union (EBU...
14/09/2025
WARNER BROS. TELEVISION'S THE PITT (HBO MAX) WINS FIVE EMMY AWARDS, INCLU...
14/09/2025
Back to All News
Mantis' Trailer Previews High-Stakes Rivalries to Be No. ...
13/09/2025
ATLANTA Cox Media Group has announced that the company's vice president of news, Misty Turnbull has been inducted into the National Academy of Television Ar...
13/09/2025
AMSTERDAM Shotoku Broadcast Systems, a major developer of robotic systems, has announced plans to take studio robotics to the next level at IBC2025 by debuting ...
13/09/2025
At IBC2025 in Amsterdam, Riedel Communications unveiled Bolero Mini, the company's lightest and flattest wireless intercom beltpack to date. Designed to del...
13/09/2025
Shotoku Broadcast Systems, the international developer of dependable, userfriendly robotic systems, is taking studio robotics to the next level at IBC 2025 with...
13/09/2025
Bitmovin, a leading provider of video streaming solutions, today released the 9th annual Video Developer Report 2025/26, offering an in-depth look at the evolvi...
13/09/2025
Bitmovin, the leading provider of video streaming solutions, today announced a strategic partnership with StreamShark, the trusted video platform for enterprise...
13/09/2025
Ikegami has chosen IBC 2025 in Amsterdam as the launch venue for a major addition to its range of viewfinders. The new VFE-P711AD is a 7-inch high resolution OL...