Sony Pixel Power calrec Sony

Give AI a Look: Any Industry Can Now Search and Summarize Vast Volumes of Visual Data

04/11/2024

Enterprises and public sector organizations around the world are developing AI agents to boost the capabilities of workforces that rely on visual information from a growing number of devices - including cameras, IoT sensors and vehicles.

To support their work, a new NVIDIA AI Blueprint for video search and summarization will enable developers in virtually any industry to build visual AI agents that analyze video and image content. These agents can answer user questions, generate summaries and enable alerts for specific scenarios.

Part of NVIDIA Metropolis, a set of developer tools for building vision AI applications, the blueprint is a customizable workflow that combines NVIDIA computer vision and generative AI technologies.

Global systems integrators and technology solutions providers including Accenture, Dell Technologies and Lenovo are bringing the NVIDIA AI Blueprint for visual search and summarization to businesses and cities worldwide, jump-starting the next wave of AI applications that can be deployed to boost productivity and safety in factories, warehouses, shops, airports, traffic intersections and more.

Announced ahead of the Smart City Expo World Congress, the NVIDIA AI Blueprint gives visual computing developers a full suite of optimized software for building and deploying generative AI-powered agents that can ingest and understand massive volumes of live video streams or data archives.

Users can customize these visual AI agents with natural language prompts instead of rigid software code, lowering the barrier to deploying virtual assistants across industries and smart city applications.

NVIDIA AI Blueprint Harnesses Vision Language Models Visual AI agents are powered by vision language models (VLMs), a class of generative AI models that combine computer vision and language understanding to interpret the physical world and perform reasoning tasks.

The NVIDIA AI Blueprint for video search and summarization can be configured with NVIDIA NIM microservices for VLMs like NVIDIA VILA, LLMs like Meta's Llama 3.1 405B and AI models for GPU-accelerated question answering and context-aware retrieval-augmented generation. Developers can easily swap in other VLMs, LLMs and graph databases and fine-tune them using the NVIDIA NeMo platform for their unique environments and use cases.

Adopting the NVIDIA AI Blueprint could save developers months of effort on investigating and optimizing generative AI models for smart city applications. Deployed on NVIDIA GPUs at the edge, on premises or in the cloud, it can vastly accelerate the process of combing through video archives to identify key moments.

In a warehouse environment, an AI agent built with this workflow could alert workers if safety protocols are breached. At busy intersections, an AI agent could identify traffic collisions and generate reports to aid emergency response efforts. And in the field of public infrastructure, maintenance workers could ask AI agents to review aerial footage and identify degrading roads, train tracks or bridges to support proactive maintenance.

Beyond smart spaces, visual AI agents could also be used to summarize videos for people with impaired vision, automatically generate recaps of sporting events and help label massive visual datasets to train other AI models.

The video search and summarization workflow joins a collection of NVIDIA AI Blueprints that make it easy to create AI-powered digital avatars, build virtual assistants for personalized customer service and extract enterprise insights from PDF data.

NVIDIA AI Blueprints are free for developers to experience and download, and can be deployed in production across accelerated data centers and clouds with NVIDIA AI Enterprise, an end-to-end software platform that accelerates data science pipelines and streamlines generative AI development and deployment.

AI Agents to Deliver Insights From Warehouses to World Capitals Enterprise and public sector customers can also harness the full collection of NVIDIA AI Blueprints with the help of NVIDIA's partner ecosystem.

Global professional services company Accenture has integrated NVIDIA AI Blueprints into its Accenture AI Refinery, which is built on NVIDIA AI Foundry and enables customers to develop custom AI models trained on enterprise data.

Global systems integrators in Southeast Asia - including ITMAX in Malaysia and FPT in Vietnam - are building AI agents based on the video search and summarization NVIDIA AI Blueprint for smart city and intelligent transportation applications.

Developers can also build and deploy NVIDIA AI Blueprints on NVIDIA AI platforms with compute, networking and software provided by global server manufacturers.

Dell will use VLM and agent approaches with Dell's NativeEdge platform to enhance existing edge AI applications and create new edge AI-enabled capabilities. Dell Reference Designs for the Dell AI Factory with NVIDIA and the NVIDIA AI Blueprint for video search and summarization will support VLM capabilities in dedicated AI workflows for data center, edge and on-premises multimodal enterprise use cases.

NVIDIA AI Blueprints are also incorporated in Lenovo Hybrid AI solutions powered by NVIDIA.

Companies like K2K, a smart city application provider in the NVIDIA Metropolis ecosystem, will use the new NVIDIA AI Blueprint to build AI agents that analyze live traffic cameras in real time. This will enable city officials to ask questions about street activity and receive recommendations on ways to improve operations. The company also is working with city traffic managers in Palermo, Italy, to deploy visual AI agents using NIM microservices and NVIDIA AI Blueprints.

Discover more about the NVIDIA AI Blueprint for video search and summarization by visiting the NVIDIA booth at the Smart Cities Expo World Congress, taking place in Barcelona through Nov. 7.

Le
LINK: https://blogs.nvidia.com/blog/video-search-summarization-ai-agents/...
See more stories from nvidia

North America Stories

14/04/2026

Appear Expands X Platform from Core to Edge at NAB Show 2...

Appear launches include XM estate management and new X Platform processing enhancements to add density for next-generation hybrid & IP workflows, X5 is also now...

14/04/2026

Synamedia turns OTT content into TikTok-style feeds with...

Addressing the needs of a new generation's viewing habits, Synamedia launches GO Shorts. The AI-powered module turns existing catalogues into TikTok-style ...

14/04/2026

NAB 2026 - Vubiquity and Eluvio Showcase Streaming Soluti...

Vubiquity, an Amdocs company and global leader in technology-led media services, will be showcasing a new end-to-end streaming solution in collaboration with El...

14/04/2026

LiveU Announces Expanded Collaboration with Sony at NAB S...

LiveU today announced a significant expansion of its collaboration with Sony Corporation, introducing integrated support for Sony's file-based workflow solu...

14/04/2026

BBC World Service TV selects Open Broadcast Systems for I...

Open Broadcast Systems (https://www.obe.tv/) has announced that BBC World Service has selected its decoders for IP Television distribution. The high-quality, lo...

14/04/2026

Blackmagic Design Announces DaVinci Resolve 21

Blackmagic Design Announces DaVinci Resolve 21 Brie Clayton April 14, 2026 0 Comments Major update adds new Photo page bringing Hollywood's most a...

14/04/2026

Sinclair's WTOV Taps Brightline for Lighting Upgrade

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Grass Valley Showcases Alliance Ecosystem at 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

FCC Selects New Lead Administrator for U.S. Cyber Trust Mark Program

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Gray Media Names Jim Hays GM of WTHI

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

NAB Blasts CTA in FCC Sports Probe Comments

Share Copy link Facebook X Linkedin Bluesky Email...

14/04/2026

Wowza to Showcase AI-Powered Video Workflows and Emerging...

Wowza will return to NAB Show 2026 with a set of live demonstrations focused on how video infrastructure is evolving for a new generation of AI-powered and oper...

14/04/2026

Stegawave Debuts Real-Time Forensic Watermarking to Tackle Piracy in Live Sports Streaming

Stegawave Debuts Real-Time Forensic Watermarking to Tackle Piracy in Live Sports...

14/04/2026

Living in Boston: A Guide for Incoming Boston Conservatory Students

Living in Boston: A Guide for Incoming Boston Conservatory Students From navigating the T to balancing school with professional gigs, a current student shar...

14/04/2026

Just What Is Genre These Days, Anyway?

Just What Is Genre These Days, Anyway? Understanding the business and art of genre-bending in 2026. April 10, 2026 By Bryan Parys Illustration by Jack Fla...

14/04/2026

Lenora Helm Hammonds Is Turning Passion Into Plan A

Lenora Helm Hammonds Is Turning Passion Into Plan A The dean of the Professional Education Division has seen the industry from all sides. Now shes bringing it...

14/04/2026

How Michelle Zalabak Found Her Dream Career in Music and Finance

How Michelle Zalabak Found Her Dream Career in Music and Finance The Warner Music Group deal analysis manager helps determine what artist catalogs are worth a...

13/04/2026

Jnger Audio Joins EBU ADM Implementers Group as Founding Member

Telos Alliance has announced that J nger Audio has joined the EBU ADM Implementers Group (ADM-IG) as a founding member. The group is focused on advancing ADM an...

13/04/2026

NAB 2026: Grass Valley to Showcase Alliance Partner Ecosystem

Grass Valley will demonstrate its Alliance Partner ecosystem at NAB Show 2026 (Booth C2408, Central Hall, April 19-22), showing AMPP integrations across live pr...

13/04/2026

NAB 2026: Media Links to Demonstrate IP Transport Solutions

Media Links will exhibit at NAB Show 2026 (Booth W2033), demonstrating IP transport solutions for live production including hitless protection technology, Xscen...

13/04/2026

NBC Sports Partners with Overtime for OT7 Football League and Navy All-American Bowl

NBC Sports has announced a programming, distribution, and sales partnership with...

13/04/2026

FloSports Promotes Jayar Donlan from COO to President

FloSports has promoted Chief Operating Officer Jayar Donlan to President, effective immediately. In his new role, Donlan will lead the company's commercial,...

13/04/2026

MASV Case Study: PanCam Pictures Uses MASV for Remote Post-Production at Senior Bowl 2026

PanCam Pictures, the documentary production company founded by Paul Camarata, us...

13/04/2026

NAB 2026: Mimir to Showcase Cloud Production Platform

Mimir will exhibit at NAB Show 2026 (North Hall, Booth N2850), demonstrating its cloud-native media production platform with new capabilities including Mimir Cu...

13/04/2026

NAB 2026: BBright Adds RIST Protocol Support to IP Gateway

BBright has announced that its IP Gateway now supports the Reliable Internet Stream Transport (RIST) protocol. The addition will be introduced at NAB Show 2026 ...

13/04/2026

Net Insight Awarded ESA NAVISP Development Project for PNT Technology

Net Insight has been awarded a development project through the European Space Agency's Navigation Innovation and Support Program (NAVISP), with co-funding f...

13/04/2026

NAB 2026: intoPIX to Showcase JPEG XS, IPMX, and SMPTE 2110 Solutions

intoPIX will exhibit at NAB Show 2026, marking the company's 20th anniversary. The company will demonstrate its JPEG XS compression portfolio and IPMX-appro...

13/04/2026

Inside the Launch of BravesVision: How Braves, Raycom Sports Pulled Off One of the Most Ambitious Efforts in Regional-Sports-Media History

Starting from scratch, the team built an in-house content platform comprising ga...

13/04/2026

NAB 2026: AI Will Make Its Presence Felt in Audio Offerings, Presentations

Here's a look at some of the new products and updates, along with audio-centric conferences, that attendees will find next week at the show When the 2026 N...

13/04/2026

NAB 2026: Avid to Demonstrate Integrated Newsroom Capabilities

Avid will launch new integrated newsroom capabilities for Avid for News at NAB Show 2026 (Booth N2226, April 18-22), demonstrating how Avid Content Core connect...

13/04/2026

NAB 2026: Synamedia Launches Cloud-Controlled Edge Playout Version of Quortex PowerVu

Synamedia has announced a new version of Quortex PowerVu, an IP-native, software...

13/04/2026

NAB 2026: Mediaproxy Adds AI Brand and Advertisement Tracking to LogServer

Mediaproxy has developed a suite of AI-powered tools for brand and advertisement tracking, integrated into its LogServer compliance logging and analysis platfor...

13/04/2026

NAB 2026: Disguise to Demonstrate Media Server and Software Integrations

Disguise will demonstrate its media servers and software at NAB Show 2026, appearing across five partner booths in Central Hall: MRMC, B&H, Planar, CarbonBlack,...

13/04/2026

NAB 2026: OpenDrives Introduces Edge Hybrid Cloud-Edge Performance Accelerator

OpenDrives is introducing OpenDrives Edge at NAB Show 2026, a hybrid cloud-edge performance accelerator for distributed video and rich media workflows. The prod...

13/04/2026

ESPN Returns to The Shed for 2026 WNBA Draft, Expanding Camera Arsenal and Deepening Fan Coverage

The show will deploy 18 cameras across two sets and the draft floor, including a...

13/04/2026

When Missiles Move at 5X the Speed of Sound, Timing Is Everything

L3Harris is accelerating the development of infrared payloads for Space Development Agency's Tranche 2 Tracking Layer, to help meet urgent national defense ...

13/04/2026

US Army Selects L3Harris for Next-Generation Night-Vision System

By leveraging cutting-edge unfilmed Gen III image intensifier technology, NOVA delivers unmatched clarity, range, and reliability in low-light environments - en...

13/04/2026

Harvey Arnold Represents the Best of Broadcast Engineering

Share Copy link Facebook X Linkedin Bluesky Email...

13/04/2026

Ross Video and HighField AI to Deliver AI-Assisted Graphics Creation

Share Copy link Facebook X Linkedin Bluesky Email...

13/04/2026

Disguise to Showcase Cutting-Edge Experience Tech for Bro...

Explore new Disguise plugins, including Sony's VP integration; Listen to panels across partner booths at Sony and B&H Disguise, the company powering everyt...

13/04/2026

TAG Video Systems Joins MXL Interoperability Initiative t...

TAG Video Systems, the leading IP-native Realtime Media Platform, has announced its participation in the Media Exchange Layer (MXL) interop initiative. TAG has ...

13/04/2026

Chaos Launches Free V-Ray for Blender Community Edition a...

Today, Chaos launched V-Ray for Blender Community Edition at BCON Austin 2026, making its production-proven 3D renderer free for all Blender users. The same Aca...

13/04/2026

LTN Appoints Mark Romano as Vice President Multichannel P...

Additions strengthen LTN's leadership as broadcasters scale satellite-to-IP transition LTN today announced the appointments of Mark Romano as Vice Presiden...

13/04/2026

NUGEN Audio Updates Halo Vision With New Precision Analys...

LEEDS, UK, APRIL 13, 2026 NUGEN Audio releases Halo Vision v1.2, a significant update to its real time, customizable audio analysis suite for 3D, surround and...

13/04/2026

Atomos to Acquire Flanders Scientific

Atomos today announced the acquisition of Flanders Scientific (FSI), one of the most respected names in professional reference monitoring. This strategic move r...

13/04/2026

How Mei Semones Built Her Sound from J-Pop, Jazz, and Bilingual Songwriting

How Mei Semones Built Her Sound from J-Pop, Jazz, and Bilingual Songwriting The indie-pop artist combines agile guitar lines, rhythmic shifts, and lyrics that...

13/04/2026

Cue the Change: Jonathon Heyward Is Making Classical Music More Relatable

Cue the Change: Jonathon Heyward Is Making Classical Music More Relatable Nicknamed the Converse Conductor, the Boston Conservatory alum holds top conductin...

13/04/2026

Heat Wave: Inside Miamis Sizzling, Boundary-Blurring Latin Music Scene

Heat Wave: Inside Miamis Sizzling, Boundary-Blurring Latin Music Scene In a city shaped by migration and exchange, Berklee alumni are helping drive a Latin mu...

13/04/2026

TikToK, Major Ad Groups Back Influencer Certification Program

Share Copy link Facebook X Linkedin Bluesky Email...

13/04/2026

DHD Marks 30th Anniversary with Brand Relaunch

DHD audio, developer and manufacturer of digital audio systems for professional broadcast, has launched a comprehensive brand update to mark its 30th anniversar...