Sony Pixel Power calrec Sony

Why Vidispine becomes cognitive

04/05/2020

Vision and hearing are your main senses when experiencing a movie. You recognize the actors, you understand the spoken language even if it is not your native language. You follow the story and enjoy the amazing film photo of the different environments - and by sharing all these experiences you can not only relate to the movie itself but also convince your friend to see the movie.

Wouldn't it be great if your Media Asset Management system could possess similar capabilities when managing your media? Being able to understand the language? Recognize actors, detect and define parts of the image - maybe also differentiate between genres? But how?

In order to do this you need a system that can actually see and listen what's inside your media - you need a system that has cognitive capabilities like yourself and can store that info in - yes, you guessed right - metadata.

But how do we navigate our vastly growing archives of file-based media?

Media files themselves today includes a lot of metadata already in a descriptive format. In here, there is room for all general metadata as well as technical metadata describing the actual file structures. MAM (Media Asset Management) systems make use of this existing metadata along with additional layers of metadata frameworks to help you navigate, find and tag not only media files themselves but also the time-based intervals of the media.

Because of this, you can argue that the true definition of a media file must include an audio-visual asset AND an associated metadata description. Without one or the other - the asset is not complete.

Cognitive Metadata to boldly go where no MAM has gone before Traditionally, the common notion is that while a machine can read and act on the associated text-based metadata of a media file, a human can understand the storyline. We can detect lipsync, recognize actors, emotions, and all the visual objects inside a frame. We can also listen to the language spoken, understand the story and do a translation into a new language.

Because of this common view on the differences between machine capabilities and human capabilities, it is still also quite common that production companies and similar, divide many tasks in a media supply chain between man and machine this way.

But times are changing, and they are changing fast. For any Content Owner, CTO or technical strategist building a modern media workflow, it is vital to challenge this traditional view on what machines can and cannot do.

Interview with Ralf Jansen Product Manager and Software Architect at Arvato / Vidispine To find out more on this subject, we talked to Ralf Jansen, Product Manager and Software Architect at Arvato / Vidispine AB. Ralf Jansen has a strong technical background, finished computer science degree with a Thesis Diploma at Fraunhofer Institute and has since worked as a developer and software architect in the industry for nearly the last 20 years. Today Ralf Jansen is managing the development of the new Vidinet Cognitive Services (VCS) and is part of the Vidinet partner success team.

So, Ralf, why is cognitive services important? Cognitive services allow the machine to find information inside the video and audio frame itself, very much like we humans can interpret the same content. This of course opens up important new possibilities depending on what type of workflow you are managing. A channel distributor can use cognitive services to automatically find (new) types of information in a huge amount of media content that could not be processed manually before - and thus use or present that insights to the viewer as a program, highlights, suggested shows or even as autogenerated trailers. Cognitive services carry this new information as metadata and give your MAM system new and much more granular methods of managing your media files. This is very important in the process of optimizing the performance and capabilities of your evolving media supply chain.

Revenue and how we can improve revenue are, of course, a driver for the advancement and adaption of cognitive services like for most other technology. And once you are getting familiar with the idea of challenging your common view on what machines can do - the subject of revenue by technology gets even more interesting.

In what areas could cognitive services improve existing revenue streams? Knowing and understanding the inside of your media opens many new opportunities that can improve revenue and help customers to monetize their owned media assets. The first one that comes to mind is of course speech to text - where cognitive services can in best case reach or even exceed the magical benchmark of human understanding (which is roughly at 5% error rate) depending on how purely spoken and what known vocabulary was used with automatic transcribe functionality already today. Automatic speech to text at this level not only free up human resources and saves money otherwise spent on external subtitling services, but also enables a new layer of time based metadata where you actually can navigate in time to find deep linked subjects, names and topics by simply searching the contents of your subtitling in your MAM systems accurate search capabilities and in our case powered by Elastic Search. And this is of course just one of many examples.

It is important to understand the value of temporal metadata since captured reality stored into the video (and audio) file changes every 30-60 frames per second or more - and because of temporal metadata we are able to define accurate time spans for different video and audio content detected by cognitive services. A post house ingesting reality content normally uses human resources for logging and preparing projects for the editors. In these and similar production workflows, the challenge is the huge amount of incoming raw footage that needs to be sorted a
LINK: https://www.vidispine.com/blog/why-vidispine-becomes-cognitive...
See more stories from vidispine

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

17/03/2026

PMVG's TechConnect Goes Virtual for 2026

Share Copy link Facebook X Linkedin Bluesky Email...

17/03/2026

Miris unlocks high-fidelity 3D asset streaming at scale

3D streaming infrastructure provider Miris today announced the launch of a public beta for its new 3D asset streaming platform. Miris is building the infrastruc...

17/03/2026

Tedial Powers the Future of Media Operations at NAB Show...

As media organizations face mounting pressure to produce more content, faster, while maximizing value and operational efficiency, Tedial, a leading provider of ...

17/03/2026

Brainstorm transforms productivity and sustainability wit...

Brainstorm, a leading manufacturer of real-time graphics, augmented and virtual production, is launching the newest version of its platform, Brainstorm Suite 7,...

17/03/2026

Limecraft Introduces New Platform Update Adding Greater C...

Limecraft today announces the release of Limecraft 2026.2, the second platform update in its 2026 release cycle. Limecraft is an AI-powered production platform ...

17/03/2026

Pioneering the Next Era of Sports Broadcasting - Broadcas...

Broadcast Solutions, a leading system integrator and provider of innovative solutions for the broadcast and media industry, showcased its latest broadcast and V...

17/03/2026

SES Launches Cash Tender Offer

THIS ANNOUNCEMENT RELATES TO THE DISCLOSURE OF INFORMATION THAT QUALIFIED OR MAY HAVE QUALIFIED AS INSIDE INFORMATION WITHIN THE MEANING OF ARTICLE 7(1) OF THE ...

17/03/2026

FCC Announces TV Translator Call Sign Changes

Share Copy link Facebook X Linkedin Bluesky Email...

17/03/2026

2026 NAB Show Offering Free Show Floor Passes to Creators

Share Copy link Facebook X Linkedin Bluesky Email...

17/03/2026

QuickLink's Latest StudioEdge Models to Make North American Debut at NAB 202

QuickLink's Latest StudioEdge Models to Make North American Debut at NAB 202 Brie Clayton March 16, 2026 0 Comments The Multi-platform Remote Gues...

17/03/2026

Frankenstein Graded with DaVinci Resolve Studio

Frankenstein Graded with DaVinci Resolve Studio Brie Clayton March 16, 2026 0 Comments Sonnenfeld enhances the controlled interplay between warm and c...

17/03/2026

New Voyavox from Link Electronics with Real-Time Speech-to-Text Captioning to be Featured in NAB Booth #W2910

New Voyavox from Link Electronics with Real-Time Speech-to-Text Captioning to be...

17/03/2026

Berklee City Music Stewards META Fellowship Supporting Massachusetts Music Educators

Berklee City Music Stewards META Fellowship Supporting Massachusetts Music Educa...

17/03/2026

Snap Decisions: How Open Libraries for Accelerated Data Processing Boost A/B Testing for Snapchat

The features on social media apps like Snapchat evolve nearly as fast as what...

17/03/2026

GTC Spotlights NVIDIA RTX PCs and DGX Sparks Running Latest Open Models and AI Agents Locally

The paradigm of consumer computing has revolved around the concept of a personal...

16/03/2026

DAZN to Stream NCAA Men's and Women's Basketball Tourneys Free in Select International Markets

DAZN will allow fans in select international territories to watch the NCAA men&#...

16/03/2026

IDM and Skate Board Association Announce Arena and Training Complex Planned for Big Bear Lake

IDM and The Skate Board Association (SBA) have announced a partnership with Coop...

16/03/2026

NAB 2026: Solid State Logic Introduces ST 2110-to-Dante Converter

Solid State Logic (SSL) will debut the Net I/O ST 2110 Bridge at NAB 2026 (booth C6907), a standalone unit that converts between ST 2110 and Dante audio formats...

16/03/2026

NAB 2026: Marshall Electronics Launches First 4K All-IP Weatherproof NDI Camera

Marshall Electronics (Booth C8339) is introducing its first all-IP 4K POV camera, the CV574-WP, at NAB 2026. The camera carries an IP67 weatherproof rating for ...

16/03/2026

Sony Expands Camera Authenticity Solution to Support Video

Sony Electronics' Camera Verify (beta), a feature of its Camera Authenticity Solution which enables news organizations to share content authenticity informa...

16/03/2026

FloSports and Storied Sports Partner on Women's and College Sports Content

FloSports has announced a partnership with Storied Sports, a content and IP studio founded by former espnW and The Players' Tribune executives, to develop s...

16/03/2026

Montreux Jazz Festival Names Gravity Media as A/V Production Provider

Montreux Jazz Festival has announced a multi-year collaboration with Gravity Media, who will become the Festival's Audio Visual Production Provider followin...

16/03/2026

USSI Global Names Ralph Annunziata Senior Vice President of Operations

USSI Global, a provider of customized network, broadcast and digital signage systems and services, has announced Ralph Annunziata joined the company on Jan. 5 a...

16/03/2026

NAB 2026: Boland Communications to Show New OLED Displays and Video Wall Applications

Boland Communications (booth C3519) will exhibit at NAB Show 2026 in Las Vegas, ...

16/03/2026

ST 2110 On The Go? A Peek Inside BRISK, FOX Sports' Broadcast Remote IP Studio Kit

Built in partnership with Diversified, the system As the sports broadcast indus...

16/03/2026

Amagi Report: FAST Viewership Up 21%, AI Adoption Growing Across Media Operations

Global FAST (Free Ad-supported Streaming TV) viewership grew 21% year-over-year ...

16/03/2026

Behind The Mic: Netflix, NBC Tap Matt Vasgersian to Call MLB and Tony Dungy Is Out at NBC

Behind The Mic provides a roundup of recent news regarding on-air talent, includ...

16/03/2026

Cloudvocal launch the SonoFlex instrument mic

Promises studio-grade fidelity for the stage Cloudvocal have announced the launch of a new instrument mic designed for professional live performers and engi...

16/03/2026

Kenton reveal the USB Solo Mk2

Popular MIDI/CV converter & interface overhauled Kenton have announced the launch of the USB Solo Mk2, a new and improved version of their compact MIDI to C...

16/03/2026

Sonarworks Spring Sale

Running from 16-29 March 2026 Starting from today (16 March) and running until 29 March 2026, Sonarworks are offering discounts of up to 40% across their ra...

16/03/2026

L3Harris Carries Goddard's Legacy Into a New Era

Dr. Robert H. Goddard and a liquid oxygen-gasoline rocket in the frame from which it was fired on March 16, 1926, at Auburn, Massachusetts. Credit: NASA....

16/03/2026

L3Harris Military GPS Receiver Deliveries Surpass 100,000 Units

Precision-guided munitions shown in production illustrate one of many operational systems benefiting from modernized M-Code GPS, supporting assured positioning,...

16/03/2026

A+E Global Media Signs New Multiyear Deal With Nielsen Covering Audience Measurement and Media Intelligence

NEW YORK - March 16, 2026 - A E Global Media and Nielsen today announced a new,...

16/03/2026

aconnic ramping up delivery of commercial 100-gigabit system

aconnic AG (ISIN: DE000A0LBKW6), Munich, is delivering the first commercial 100-Gigabit systems following successful validation and certification for customer n...

16/03/2026

Spectrum Launches Multiview for March Madness

Share Copy link Facebook X Linkedin Bluesky Email...

16/03/2026

Ikegami To Spotlight Latest UNICAM 4K-UHD Cameras At 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

16/03/2026

A+E Global Media Signs New Agreement With Nielsen

Share Copy link Facebook X Linkedin Bluesky Email...

16/03/2026

Shotoku Brings Broadcast-Grade Control to PTZ with New Au...

Shotoku USA, Shotoku Broadcast Systems' North American operation, will unveil significant additions to its platform at NAB 2026. Topping the list is the wor...

16/03/2026

Ikegami to Showcase Latest Generation TV Production Camer...

Ikegami USA will demonstrate the latest additions to its wide range of broadcast-quality cameras, controllers and monitors on Central Hall booth C3819 during th...

16/03/2026

[Updated] Carr Threatens Broadcast Licenses Over Iran War Coverage

Share Copy link Facebook X Linkedin Bluesky Email...

16/03/2026

ELEMENTS launches GRID at NAB Show 2026

ELEMENTS launches GRID at NAB Show 2026 Brie Clayton March 15, 2026 0 Comments North Hall, Booth N1717 ELEMENTS returns to NAB Show 2026, with an exp...

16/03/2026

Blackmagic Design Cameras Capture Artist Salavat Fidai's Micro Sculptures

Blackmagic Design Cameras Capture Artist Salavat Fidai's Micro Sculptures Brie Clayton March 15, 2026 0 Comments 6K sensor and open gate capabilit...

16/03/2026

DHD to Introduce Latest Generation Broadcast Audio Mixers at NAB 2026, Las Vegas

DHD to Introduce Latest Generation Broadcast Audio Mixers at NAB 2026, Las Vegas Brie Clayton March 15, 2026 0 Comments Hero image: Front of DHD RM1 P...

16/03/2026

VEON Files its 2025 Annual Report on Form 20-F

16 Mar 2026 VEON Files its 2025 Annual Report on Form 20-F Dubai and New York, March 16, 2026 - VEON Ltd. (Nasdaq: VEON), a global digital operator ( VEON'...