
With persistence and the right tools, Deborah Tylor was able to do the impossible.
A data scientist, she was tasked to comb a 3+ terabyte dataset at the Internal Revenue Service for patterns that might help uncover fraud. But even when she let the job run all night on a large bank of CPU servers the data refused to line up.
She returned in the morning to find the job had failed, so she tried again. It failed again.
About that time, Nasheb Ismaily of Cloudera knocked on the door of Rahul Tikekar, manager of a technical team that supports data analysts at the IRS. The Cloudera solutions engineer asked if Tikekar's team had any uses for Cloudera Data Platform (CDP), implementing Apache Spark 3.0 software accelerated by GPUs.
I jumped at the opportunity, said Tikekar. We have NVIDIA graphics cards on standalone servers, but using Spark to run them on a distributed cluster had eluded us for a while, so this was perfect timing for us and Deb had the perfect use case, he said.
A Nerdy Knot Untied A quick test of the software immediately speeded up many parts of Tylor's work up to 5x with no code changes, but a few pieces still lagged.
Ismaily called in a team of data scientists at NVIDIA to examine the guts of the code. They quickly determined a few tasks with particularly gnarly data structures were still running on CPUs. They wrote code to handle those jobs and inserted it into Spark's software interface for RAPIDS, the open library for running data analytics on GPUs.
Tylor ran another test, and boom, it all went on the GPUs in a distributed Spark cluster and the speedup was remarkable - Deb's running the whole program on a four-node cluster right now, said Tikekar.
The Cloudera and NVIDIA integration will empower us to use data-driven insights to power mission-critical use cases, said Joe Ansaldi, technical branch chief of the research and applied analytics and statistics division at the IRS and Tikekar's boss.
We're currently implementing this integration, and already seeing over 20x speed improvements at half the cost for our data engineering and data science workflows, he added.
Spark 3.0 + GPUs = New Horizons The work promises several payoffs the IRS team is already exploring.
With a Spark cluster of GPU-powered servers, the group can accelerate all its current jobs and run others previously thought impractical. And those jobs can tackle big datasets the team has at its disposal.
Before Spark 3.0, this was not possible, but now we're upping the ante with GPUs and we can dream of solving problems that were once impossible, said Tikekar.
Charting a Course to AI The team plans to apply what it learned with its success in data preparation, the so-called extract/transform/load (ETL) work of data analytics. Its next big step is accelerating full-blown AI inference jobs.
The partnership with Cloudera and NVIDIA helped us harness GPUs in clusters. When such advances come along, it takes a while to realize their power and develop apps that can use them, so Deb is really charting a new course for us - she's definitely the hero of the story, Tikekar said.
Specifically, the team aims to provide this distributed Spark-GPU infrastructure to analysts. Together, they will build large deep learning neural networks to tackle natural language processing and other analytics jobs currently impossible on a single server.
Many Apps for Machine Learning It's the kind of transformation many enterprises are seeking today with machine learning.
My personal feeling is that machine learning brings an incredible potential to make things that were difficult to achieve possible, said Tikekar, a Ph.D. in computer science who spent a decade teaching at Southern Oregon University before joining the IRS more than 13 years ago.
For example, today we scan in forms and then apply optical character recognition to read pieces of them, but with AI we can do a much better job of reading forms and finding patterns that can help find ID theft or reduce waste - a lot of applications can benefit from AI in numerous ways, he added.
To learn more about accelerating Cloudera's CDP 7.1.6 with NVIDIA GPUs, watch a GTC talk (free to view with registration) from October 2020, when the two companies announced their partnership.
And view Cloudera's demo below of a 44x speed increase on a data science workload using NVIDIA GPUs and RAPIDS compared to CPUs.
Most recent headlines
05/01/2027
Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...
04/08/2026
Dalet, a leading technology and service provider for media-rich organizations, t...
04/07/2026
April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...
01/06/2026
January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026
Throughout the week, Dolby brings to life the latest innovatio...
20/05/2026
Amp-simulation software expanded
Acustica Audio's latest release greatly expands on their amp-simulation platform, turning it into a complete amplifica...
20/05/2026
MainStage integration, Analog Lab improvements & more
Arturia have just announced the release of an update that brings an assortment of new features to thei...
20/05/2026
At the Annual General Meeting held on May 20, 2026, the shareholders of SGL Carb...
20/05/2026
From Gulkula to the nation: Yothu Yindi Foundation and NITV deepen national acce...
20/05/2026
SBS appoints David Fernandez as National Manager, Digital & TV Sales
20 May, 2026
Media releases
SBS has appointed David Fernandez as National Manager, Dig...
20/05/2026
Rohde & Schwarz and INFOZAHYST: A strategic alliance set to redefine modern defe...
20/05/2026
A Rotating Detonation Engine being hot fire tested at Purdue University's Zu...
20/05/2026
A U.S. Army VAMPIRE system, assigned to Bravo Battery, 1st Battalion, 51st Air Defense Artillery Regiment, 7th Infantry Division/Multi-Domain Command - Pacific ...
20/05/2026
Cable Captures Only Monthly Increase Among Viewing Categories in March, Earns it...
20/05/2026
April brought a symbolic decrease in the overall time spent in front of television screens. On average, Poles watched video content for 3 hours and 51 minutes a...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
The Royal Television Society Technology Centre today announces the launch of the RTS Technology Awards 2026, celebrating excellence, innovation and achievement ...
20/05/2026
In the heart of London's financial district, the new purpose-built Troubadour Canary Wharf Theatre invites audiences to experience Suzanne Collins' inte...
20/05/2026
LiveU, the leader in live IP-video solutions, today announced that production powerhouse BCC Live successfully deployed the new LU900Q intelligent production un...
20/05/2026
Nella Mente di Narciso Docuseries Uses Blackmagic Design Workflow
Brie Clayton May 19, 2026
0 Comments
PYXIS 6K full frame camera and DaVinci Resolve ...
20/05/2026
Beeble launches Canvas, a node-based AI compositor for VFX and Virtual Productio...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
Share
Copy link
Facebook
X
Linkedin
Bluesky
Email...
20/05/2026
First Nations Factual Co-production Development Fund launched to elevate Indigen...
20/05/2026
Back to All News
One Hundred Years of Solitude Concludes This August, With Part...
20/05/2026
Back to All News
Beloved NHK Dramas to Stream on Netflix Worldwide Starting June 22
Entertainment
20 May 2026
GlobalJapan
Link copied to clipboard
Netflix...
20/05/2026
May 20 2026, 06:00 (PDT) Dolby Recognized as 2025 Supplier of the Year and Over...
20/05/2026
Mayo's Dee Freney and Margaret Leahy from Galway have reached the final of RT Today's TV Home Cook competition.
Both contestants will cook again live...
20/05/2026
RT IN FULL BLOOM AT BORD BIA BLOOM 2026 WITH LIVE BROADCASTS, MUSIC, CHAT AND M...
19/05/2026
The winner of Thomson Foundation's Young Journalist of the Year 2025, Tracy Bonareri Onchoke, and runner up Wangu Kanuri enjoyed a three-day trip to London ...
19/05/2026
Cisco and the USGA have announced a multiyear extension of their partnership, which began in 2018. Cisco serves as the Official Technology Partner of the USGA, ...
19/05/2026
Urban Edge Network (UEN), a streaming platform for NAIA sports, has announced a partnership with Spiideo to provide streaming and production tools to UEN's ...
19/05/2026
Warner Bros. Discovery (WBD) will provide live coverage of all 900 Roland-Garros matches across its platforms beginning with qualifiers on May 18. In Europe, 21...
19/05/2026
Tubi, Fox Corporation's free streaming service, has announced the launch of the FIFA World Cup 2026 FOX Hub, a dedicated destination for World Cup programmi...
19/05/2026
Telef nica, in collaboration with Sony, has conducted a 5G connectivity trial at the Movistar Arena in Spain using the 26 GHz millimetre wave (mmWave) band. The...
19/05/2026
Ross Production Services (RPS) has installed a Calrec Argo M console into its new Hypermax-1 remote production truck, replacing one of three Argo S consoles pre...
19/05/2026
Panasonic Projector and Display Corporation has announced the acquisition of 100% of the shares of UK-based media technology company Hive Media Control Ltd. (HI...
19/05/2026
Globecast has announced the completion of a nine-month renovation of its Singapore facility, converting it from a traditional linear broadcast operation into a ...
19/05/2026
Grass Valley has announced a three-year enterprise agreement with Phoenix Broadc...
19/05/2026
Bitmovin has announced that Watch Brasil, a streaming platform operating across Brazil and Europe since 2018, has replaced its legacy systems with Bitmovin'...
19/05/2026
Ateme has announced the migration of Dish Home Nepal's Nepal Premier League (NPL) streaming infrastructure to Ateme's TITAN Live solution deployed on Ak...
19/05/2026
CMSI provided workflow, media management, and HDR support for ESPN during coverage of the NCAA Gymnastics Semifinals and Championships. The company supported fi...
19/05/2026
In advance of this year's Sports Emmy Awards, SVG is taking a deep dive into the six production-technologies nominated for this year's George Wensel Tec...
19/05/2026
In advance of this year's Sports Emmy Awards, SVG is taking a deep dive into the six production-technologies nominated for this year's George Wensel Tec...
19/05/2026
Featuring a fully IP infrastructure, Supershooter 11 is intended for large-scale events. Enabling remote and distributed workflows, Supershooter 65 joins the RE...
19/05/2026
By Jessica Herndon
The line wrapped around the building outside Denver's La...
19/05/2026
Podcasting continues to evolve, and so does Spotify. As we build what comes next, one thing remains constant: This is a medium built on connection. It lives in ...
19/05/2026
Popular design joins Inherit cartridge line-up
When GC Audio introduced their modular Inherit system, it was available with a selection of the company's...
19/05/2026
Resonance-suppression plug-in gets ground-up rebuild
Following on from its 10-year anniversary, oeksound's flagship plug-in has just reached its third m...
19/05/2026
Dedicated FL Studio controller keyboard range refreshed
Novation's dedicated FL Studio controller family has just been upgraded, with four new models ex...