Why Vidispine becomes cognitive

04/05/2020

Vision and hearing are your main senses when experiencing a movie. You recognize the actors, you understand the spoken language even if it is not your native language. You follow the story and enjoy the amazing film photo of the different environments - and by sharing all these experiences you can not only relate to the movie itself but also convince your friend to see the movie.

Wouldn't it be great if your Media Asset Management system could possess similar capabilities when managing your media? Being able to understand the language? Recognize actors, detect and define parts of the image - maybe also differentiate between genres? But how?

In order to do this you need a system that can actually see and listen what's inside your media - you need a system that has cognitive capabilities like yourself and can store that info in - yes, you guessed right - metadata.

But how do we navigate our vastly growing archives of file-based media?

Media files themselves today includes a lot of metadata already in a descriptive format. In here, there is room for all general metadata as well as technical metadata describing the actual file structures. MAM (Media Asset Management) systems make use of this existing metadata along with additional layers of metadata frameworks to help you navigate, find and tag not only media files themselves but also the time-based intervals of the media.

Because of this, you can argue that the true definition of a media file must include an audio-visual asset AND an associated metadata description. Without one or the other - the asset is not complete.

Cognitive Metadata to boldly go where no MAM has gone before Traditionally, the common notion is that while a machine can read and act on the associated text-based metadata of a media file, a human can understand the storyline. We can detect lipsync, recognize actors, emotions, and all the visual objects inside a frame. We can also listen to the language spoken, understand the story and do a translation into a new language.

Because of this common view on the differences between machine capabilities and human capabilities, it is still also quite common that production companies and similar, divide many tasks in a media supply chain between man and machine this way.

But times are changing, and they are changing fast. For any Content Owner, CTO or technical strategist building a modern media workflow, it is vital to challenge this traditional view on what machines can and cannot do.

Interview with Ralf Jansen Product Manager and Software Architect at Arvato / Vidispine To find out more on this subject, we talked to Ralf Jansen, Product Manager and Software Architect at Arvato / Vidispine AB. Ralf Jansen has a strong technical background, finished computer science degree with a Thesis Diploma at Fraunhofer Institute and has since worked as a developer and software architect in the industry for nearly the last 20 years. Today Ralf Jansen is managing the development of the new Vidinet Cognitive Services (VCS) and is part of the Vidinet partner success team.

So, Ralf, why is cognitive services important? Cognitive services allow the machine to find information inside the video and audio frame itself, very much like we humans can interpret the same content. This of course opens up important new possibilities depending on what type of workflow you are managing. A channel distributor can use cognitive services to automatically find (new) types of information in a huge amount of media content that could not be processed manually before - and thus use or present that insights to the viewer as a program, highlights, suggested shows or even as autogenerated trailers. Cognitive services carry this new information as metadata and give your MAM system new and much more granular methods of managing your media files. This is very important in the process of optimizing the performance and capabilities of your evolving media supply chain.

Revenue and how we can improve revenue are, of course, a driver for the advancement and adaption of cognitive services like for most other technology. And once you are getting familiar with the idea of challenging your common view on what machines can do - the subject of revenue by technology gets even more interesting.

In what areas could cognitive services improve existing revenue streams? Knowing and understanding the inside of your media opens many new opportunities that can improve revenue and help customers to monetize their owned media assets. The first one that comes to mind is of course speech to text - where cognitive services can in best case reach or even exceed the magical benchmark of human understanding (which is roughly at 5% error rate) depending on how purely spoken and what known vocabulary was used with automatic transcribe functionality already today. Automatic speech to text at this level not only free up human resources and saves money otherwise spent on external subtitling services, but also enables a new layer of time based metadata where you actually can navigate in time to find deep linked subjects, names and topics by simply searching the contents of your subtitling in your MAM systems accurate search capabilities and in our case powered by Elastic Search. And this is of course just one of many examples.

It is important to understand the value of temporal metadata since captured reality stored into the video (and audio) file changes every 30-60 frames per second or more - and because of temporal metadata we are able to define accurate time spans for different video and audio content detected by cognitive services. A post house ingesting reality content normally uses human resources for logging and preparing projects for the editors. In these and similar production workflows, the challenge is the huge amount of incoming raw footage that needs to be sorted a

LINK:	https://www.vidispine.com/blog/why-vidispine-becomes-cognitive...
	See more stories from vidispine

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

19/03/2026

Study: Creator Content Plays Growing Role in Streaming Habits

Share Copy link Facebook X Linkedin Bluesky Email...

19/03/2026

NAB Show: Studio Technologies to Debut New StudioComm System

Share Copy link Facebook X Linkedin Bluesky Email...

19/03/2026

Ateme Verified for YouTube Live

Share Copy link Facebook X Linkedin Bluesky Email...

19/03/2026

Roku Launches NCAA March Madness Zone

Share Copy link Facebook X Linkedin Bluesky Email...

19/03/2026

Harmonic Makes NextGen TV Upgrades to XOS Media Processor

Share Copy link Facebook X Linkedin Bluesky Email...

19/03/2026

FCC Removes More Drones from Covered List of Banned Products

Share Copy link Facebook X Linkedin Bluesky Email...

19/03/2026

How Chris Bolte Builds the Sound of PlayStation Games

How Chris Bolte Builds the Sound of PlayStation Games Chris Bolte '16 is a technical sound designer at Sucker Punch Productions, where he helps create the...

19/03/2026

Foo Fighters' Other Voices performance to air on RT this Easter Monday

Other Voices presents a legendary night of music as Foo Fighters bring their stadium anthems to St James' Church, An Daingean. The surprise performance teas...

19/03/2026

RTS Ireland Awards / Gradaim RTS 2026 Shortlist Announced

***Issued by RT on behalf of RTS Ireland Awards / Gradaim RTS 2026...

19/03/2026

Smooth Moves: 90 Frames-Per-Second Virtual Reality Arrives on GeForce NOW

It's a double feature on GFN Thursday. This week, GeForce NOW offers smoother sights in virtual reality (VR) and a sprawling new land to conquer. Streaming...

18/03/2026

Net Insight Names Larissa GrnerMeeus CPO

Newly named chief product officer (CPO), Larissa G rner Meeus will return to Net Insight on May 4. Larissa G rner Meeus will become CPO of Net Insight on May 4...

18/03/2026

SVG Europe's The Football Summit 2026: Select Sessions Now Available to Watch on SVG PLAY

SVG Europe's The Football Summit 2026 explored how the sports broadcasting i...

18/03/2026

SVG's SportsTech@NABShow Blog Goes Live as Countdown to Vegas Heats Up

The 2026 NAB Show kicks off one month from today and SVG is once again set to cover the show from every angle. SVG's SportsTech@NAB Show Blog is now live - ...

18/03/2026

The Premiere of Dead Lover Showcases Grave Robbing and Sex With a Giant Finger

(L-R) The cast and crew of Dead Lover at The Ray Theater for its premiere at the 2025 Sundance Film Festival. (Photo by Robin Marshall/Shutterstock for Sundan...

18/03/2026

Music Row Piano from Wiltone Productions

Yamaha C7 captured in Nashville Wiltone Productions have announced the release of Music Row Piano, a deeply sampled Yamaha C7 piano library that's been ...

18/03/2026

Techivation introduce T-Warmer Mk2

Bass-enhancement plug-in upgraded Along with a steady stream of new releases, Techivation have recently been revisiting some of the older plug-ins in their ...

18/03/2026

Tonal Balance Control 3 from iZotope

Mix-reference plug-in overhauled iZotope's powerful mix-referencing plug-in has just reached its third major version, and now boasts a new capture proce...

18/03/2026

Toontrack Transistor Organ EKX

Latest EZKeys 2 expansion arrives Toontrack's staggering collection of EZKeys 2 expansions has grown once again, and the latest instalment delivers a on...

18/03/2026

Jess Ho explores the politics of food in new SBS Audio podcast For The Culture

Jess Ho explores the politics of food in new SBS Audio podcast For The Culture 18 March, 2026 Media releases New SBS Audio podcast For The Culture, hosted ...

18/03/2026

Future-ready broadcasts with ATSC 3.0 and enhanced services from Rohde & Schwarz at NAB 2026

Future-ready broadcasts with ATSC 3.0 and enhanced services from Rohde & Schwarz...

18/03/2026

LCTWS 2026

London Calling - When The Industry Convened to Help Streaming Find its MoJo In this blog, Laura Rognoni reflects on key discussions from the Connected TV World...

18/03/2026

SMPTE Unveils 2026 NAB Show Educational Presentations

SMPTE Unveils 2026 NAB Show Educational Presentations Brie Clayton March 18, 2026 0 Comments SMPTE , the home of media professionals, technologists, a...

18/03/2026

Auditel Ad Campaign Shot on Blackmagic PYXIS 12K

Auditel Ad Campaign Shot on Blackmagic PYXIS 12K Brie Clayton March 18, 2026 0 Comments LED wall virtual production blends 12K open gate acquisition w...

18/03/2026

Brainstorm transforms productivity and sustainability with Suite 7 at NAB Show 2026

Brainstorm transforms productivity and sustainability with Suite 7 at NAB Show 2...

18/03/2026

Neutrik To Showcase opticalCON ADVANCED Connectors At 2026 NAB Show

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

SMPTE Details 2026 NAB Show Educational Sessions

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

Ben Bradshaw Joins PSSI as Director, Product and Network Development

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

Peter Thordarson Joins ASG as Technical Account Executive

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

Survey: Voters Trust TV News Over AI, Social and Search

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

2026 NAB Show Exhibitor Insight: Amazon Web Services (AWS)

Share Copy link Facebook X Linkedin Bluesky Email...

18/03/2026

SMPTE Unveils 2026 NAB Show Educational Presentations

SMPTE , the home of media professionals, technologists, and engineers, today unveiled its educational presentations for the 2026 NAB Show. This year SMPTE will ...

18/03/2026

Maxon Marks Its Official Entry Into the AEC Market With I...

Maxon, maker of powerful, approachable software solutions for creators working in 2D and 3D design, motion graphics, visual effects, gaming, and more, today ann...

18/03/2026

Digital Alert Systems NAB Preview 2026

Digital Alert Systems Preview 2026 NAB Show April 19 - 22 Booth C3452 At the 2026 NAB Show, Digital Alert Systems will showcase Version 6.0 of its DASDEC ...

18/03/2026

Setplex Transforms Video Streaming with AI and Super Aggr...

Setplex today announced that it will showcase its complete, fully integrated Zapflex platform for the first time at the 2026 NAB Show, introducing powerful new ...

18/03/2026

SES Announces Extension of Tender Offer

THIS ANNOUNCEMENT RELATES TO THE DISCLOSURE OF INFORMATION THAT QUALIFIED OR MAY HAVE QUALIFIED AS INSIDE INFORMATION WITHIN THE MEANING OF ARTICLE 7(1) OF THE ...

18/03/2026

COW Jobs: Seeking DP for Low Budget Dramedy - Chicago

COW Jobs: Seeking DP for Low Budget Dramedy - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Director of Photography for Low Budget Dramedy Fe...

18/03/2026

COW Jobs: Seeking Gaffer for Low Budget Dramedy - Chicago

COW Jobs: Seeking Gaffer for Low Budget Dramedy - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Gaffer for Low Budget Dramedy Feature Film- I...

18/03/2026

COW Jobs: Seeking Location, Sound for Low Budget Dramedy - Chicago

COW Jobs: Seeking Location, Sound for Low Budget Dramedy - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Location/Sound for Low Budget Dramed...

18/03/2026

COW Jobs: Seeking Child Wrangler for Low Budget Film - Chicago

COW Jobs: Seeking Child Wrangler for Low Budget Film - Chicago Brie Clayton March 17, 2026 0 Comments Seeking Child Wrangler for Low Budget Dramedy Fe...

18/03/2026

Calrec Redefines Broadcast Workflows at NAB 2026 with its Most Powerful Hardware, Virtual and Hybrid Audio Lineup Yet

Calrec Redefines Broadcast Workflows at NAB 2026 with its Most Powerful Hardware...

18/03/2026

Oscar Nominated Two People Exchanging Saliva Posted with DaVinci Resolve Studio

Oscar Nominated Two People Exchanging Saliva Posted with DaVinci Resolve Studio Brie Clayton March 17, 2026 0 Comments DaVinci Resolve Studio handle...

18/03/2026

Freelance Developer Releases RizomUV Link for Cinema 4D - A Three-Edition Plugin That Brings Fileless UV Bridging, Scripting, and Full Batch Processing to C4D Artists

18/03/2026

Boston Conservatory Presents Celebrated Musical Satire Urinetown

Boston Conservatory Presents Celebrated Musical Satire Urinetown Performances for this Center Stage production will take place at Boston Conservatory Theater ...

18/03/2026

Charlie Puth Joins Switched On Pop at Berklee NYC

Charlie Puth Joins Switched on Pop at Berklee NYC The Berklee alum spoke with host and Berklee NYC professor Charlie Harding for a live taping, answering audi...

18/03/2026

X-Rite Pantone Demonstrates Advanced Color Management Solutions to Support Smart Manufacturing at MAX and American Coatings Show

X-Rite Pantone Demonstrates Advanced Color Management Solutions to Support Smart...

View most recent headlines