Sony Pixel Power calrec Sony

AI Can Be Leveraged to Simplify, Enhance STT Services

01/07/2020

AI Can Be Leveraged to Simplify, Enhance STT Services

Author:Guy Finley Artificial intelligence (AI) can be used by media and entertainment companies to simplify and enhance all of their subtitling, translation and transcription (STT) services in the cloud, according to M&E technology firm Digital Nirvana.

Digital Nirvana's Russell Wise, SVP of sales and marketing, and Ed Hauber, its business development manager, used the June 24 webinar Leveraging AI for Speed & Efficiency in M&E STT to detail how Trance - the company's enterprise-level, cloud-based closed captioning and translation solution - can simplify the process, as a managed or self-service STT tool.

Bloomberg, Turner and other major media organizations are already using the plug-and-play, AI-powered offering to produce captions at record speed, improving productivity by 50% and more, according to Digital Nirvana. The workflow can be used across the industry, with media, post and caption service providers all able to take advantage.

Trance is a cloud-based, enterprise-level Software-as-a-Service (SaaS) platform that is used to generate automated transcripts, to create closed captions, to translate those captions into alternate languages and also to export captioned files in all known industry-supported formats, Hauber pointed out.

Trance is also fully web-based, he noted, adding: It's accessible via a LAN, WAN or even a basic Internet connection. As an enterprise tool, Trance is fully configurable for an unlimited number of users, groups and roles.

Administrators, meanwhile, can manage multiple projects, they can create manage users, define roles and permissions, as well as establish system presets, he said, while giving viewers a demonstration of Trance.

The Manage Presets section gives users the ability to define caption attributes, such as the number of lines, the line length and the total number of characters, he pointed out during the demo.

To get media into Trance, we have a tool that we use called Media Services Portal and, like Trance, Media Services Portal - also called MSP - [is] a cloud-based platform, which allows users to ingest any number of common audio and video file formats into Trance, he said. MSP can directly integrate with both FTP and Amazon S3, he also noted.

Digital Nirvana also offers an open application programming interface (API) to integrate Media Services Portal directly into large enterprise media systems, he pointed out. Using our API, those operators don't need to create a secondary workflow process to move media into and out of Trance - and this is a really big time-saving and productivity advantage of Trance, he said.

The Trance speech-to-text engine has created a highly-accurate transcript of the media that we just imported, he also showed during the demo, noting that eliminates the necessity of doing the manual transcribing of content and delivers huge productivity gains over conventional transcription methods. It is also highly accurate - between roughly 90 to 95 percent accurate - based on good good-quality content, he noted.

The transcript interface includes text on the right side of the screen and a media player on the left with intuitive controls to play back audio and video, he demonstrated. Also featured are tools that help provide fast text editing, including an auto highlight of potentially misspelled words and spell check, he showed. Users can also create captions in more than one language, he noted.

During the Q&A, he said: Unlike other providers, we're not limited to one specific speech-to-text engine. In fact, we, by design, do not operate that way. We constantly evaluate and measure the performance of all the best speech-to-text engines that exist in the marketplace today. And so, we're not limited to just one. And the reason that that's important is this technology is progressing and developing and advancing very quickly and so being tied to one or the other is inherently limiting. We would rather take the approach of using them all and continually measuring and evaluating them.

So, as an example, if we detect that Engine A' is performing better in scenarios - say where there is sports content, and we can even be more specific: domestic American basketball - we see that speech-to-text Engine A' is performing better in this application, we automatically in the background route that content based on machine learning capability to say we're going to route this client's content through this speech-to-text engine because we see it now as performing better than the other options, he explained.

There is a great degree of accuracy that we can accomplish by using that process, he noted.

Although Trance is currently not a live captioning solution, he was quick to say: It is on our roadmap and it is something that we're actively developing. So, live captioning with the ability to run our speech-to-text engine, to collapse the time of that speech-to-text process down to near real-time, or essentially real-time, giving an operator the ability to make very quick edits within a few seconds of live and be able to do that on the fly. That's something that we're evaluating and we're working towards as the technology matures and there's a degree of reliability and consistency that we can bring to the market that is on the roadmap for sure. Not today - but coming soon.

He went on to point out: We're constantly developing the product . This company really adheres to a philosophy and a down-to-earth principle in being very, very agile. And, as much as this is an enterprise tool, the product operates on a very agile basis, meaning it's able to take and respond to customer requests very, very quickly.

There is a long history at Digital Nirvana of continual development an
LINK: https://digital-nirvana.com/ai-can-be-leveraged-to-simplify-enhance-st...
See more stories from digitalnirvana

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

31/03/2026

Spectrum TV App to Launch on Amazon Fire TV Devices

Share Copy link Facebook X Linkedin Bluesky Email...

31/03/2026

CueScript to Showcase Innovations at NAB that Add Mobilit...

CueScript, a leading international developer of professional teleprompting solutions, is enhancing its platform with three innovations that elevate flexibility,...

31/03/2026

GALSNGEAR Returns to the 2026 NAB Show with Expanded Meet...

#GALSNGEAR will host a series of meetups, panel discussions, and networking events during the 2026 NAB Show, taking place April 18 22 in Las Vegas. The organi...

31/03/2026

Imagine Introduces AI Assisted Scheduling at 2026 NAB Sho...

At the 2026 NAB Show (April 19-22, Las Vegas Convention Center, Booth #N1328), Imagine Communications is introducing AI assisted scheduling capabilities within ...

31/03/2026

PSSI boosts live sports flexibility and security with the...

High-density upgrade for mobile fleet and PSSI International Teleport gives PSSI more flexibility for premium live event coverage PSSI and Appear today announc...

31/03/2026

Solid State Logic Enhances Virtualized System T Platform to Deliver Scalable Cloud and On-Prem Production for NAB 2026

Solid State Logic Enhances Virtualized System T Platform to Deliver Scalable Clo...

31/03/2026

OWC to Showcase Storage, Workflow Acceleration, and Reliability Solutions for Creatives and Post-Production Professionals at NAB 2026

OWC to Showcase Storage, Workflow Acceleration, and Reliability Solutions for Cr...

31/03/2026

PSSI Taps Appear X for Live Coverage of FIFA World Cup

Share Copy link Facebook X Linkedin Bluesky Email...

31/03/2026

OpenDrives Shows Off Sports Expertise in Sports Business...

OpenDrives, a leader in software-defined video and rich media storage management solutions, will demonstrate several new innovations at the 2026 NAB Show in Apr...

31/03/2026

SDVI to Showcase Platform Enhancements New Technology Int...

At the 2026 NAB Show, SDVI will demonstrate how its Rally media supply chain management platform continues to give media operations teams the tools they need to...

31/03/2026

Boland Communications Introduces QD4K315HDR10 QD-OLED Ser...

Boland Communications today announced that at the 2026 NAB Show, it will introduce its new QD4K315HDR10, a 31.5-inch QD-OLED monitor delivering exceptional colo...

31/03/2026

NUGEN Audio CEO Dr Paul Tapper to Lead Presentation About...

Dr. Paul Tapper, CEO of NUGEN Audio, will lead an essential industry discussion at the 2026 NAB Show, examining whether dialog intelligibility is poised to beco...

31/03/2026

BCNEXXT at NAB 2026 - Turning Industry Pressure into Oppo...

BCNEXXT will use NAB 2026 to focus on how broadcasters can turn today's operational and market pressures into opportunity. Meeting with customers in the IAB...

31/03/2026

Mediagenix Showcases Semantic Intelligence Powered Title...

Mediagenix, a global leader in smart content solutions to profitably connect the right content to the right audience, will showcase its latest innovations at th...

31/03/2026

DHD Introduces AI-Based Audio Noise Reduction to XD3 IP C...

DHD announces that AI-based audio noise reduction is now available as a powerful new option for its XD3 IP Core. Developed in cooperation with ai-coustics, audi...

31/03/2026

PTZOptics to showcase intelligent video and camera contro...

PTZOptics will showcase its vision for intelligent video and advanced camera automation at NAB Show 2026 (#N1902). Through immersive demonstrations and technolo...

31/03/2026

Experience Commerce Wins Social Media Mandate for ICONIQA...

Experience Commerce, a leading full-service digital marketing agency and part of the Cheil SWA Group, has secured the social media mandate for ICONIQA Hotels & ...

31/03/2026

Fabric to Showcase Cloud-Native Origin Platform and Next-...

Fabric, the entertainment industry's leader in data and operations solutions, will showcase a major expansion of its media technology platform at NAB 2026, ...

31/03/2026

Boland Communications Introduces New QD-OLED Series Monitors

Share Copy link Facebook X Linkedin Bluesky Email...

31/03/2026

Carr Warns NFL Over Streaming Rights, Consumer Costs

Share Copy link Facebook X Linkedin Bluesky Email...

31/03/2026

Farhan Khan Appointed FCC Chief Information Officer

Share Copy link Facebook X Linkedin Bluesky Email...

31/03/2026

Federal Judge Pauses Nexstar/Tegna Merger

Share Copy link Facebook X Linkedin Bluesky Email...

31/03/2026

Boris FX Acquires Vegas Pro, Sound Forge, and Acid Pro

Boris FX Acquires Vegas Pro, Sound Forge, and Acid Pro Jessie Electa Petrov March 30, 2026 0 Comments The acclaimed VFX developer bolsters its award-w...

31/03/2026

Berklee India Exchange Presents Anirudh Varma Collective

Berklee India Exchange Presents Anirudh Varma Collective The collective will present a workshop on Hindustani classical music and a concert in collaboration w...

31/03/2026

Efficiency at Scale: NVIDIA, Energy Leaders Accelerating PowerFlexible AI Factories to Fortify the Grid

CERAWeek - dubbed the Davos of energy - is where policymakers, producers, techno...

30/03/2026

NAB 2026: Manifold to Demonstrate 400GbE COTS FPGA Support

Manifold Technologies, a Germany-based provider of cloud infrastructure for live broadcast production, will demonstrate support for 400GbE COTS FPGA accelerator...

30/03/2026

NAB 2026: Boland Communications Introduces QD-OLED Series Monitors

Boland Communications will introduce its QD4K315HDR10, a 31.5-inch QD-OLED monitor, at NAB Show 2026 (Booth C3519, April 18-22). The company is also introducing...

30/03/2026

NAB 2026: PTZOptics to Showcase Move 4K and Horizon Platform

PTZOptics will demonstrate its Move 4K PTZ cameras and Horizon web-based control platform at NAB Show 2026 (Booth N1902). Move 4K with Horizon is now available...

30/03/2026

NAB 2026: Net Insight to Showcase Updated Nimbra Edge

Net Insight will demonstrate the next version of Nimbra Edge, its orchestration and control layer for live media services across multi-domain environments, at N...

30/03/2026

NAB 2026: Appear to Showcase Live Production Processing

Appear ASA will exhibit at NAB Show 2026 (Booth W1531, April 19-22, Las Vegas). The company completed an IPO in November 2025. Our customer-first approach is ...

30/03/2026

NAB 2026: Harmonic Announces New Live Sports Streaming Capabilities

Harmonic has announced new capabilities for its sports streaming platform, covering multiview, programmatic advertising, in-stream advertising, and content wate...

30/03/2026

NAB 2026: Ateme to Showcase GenAI, Agentic AI, and Streaming

Ateme (Booth W1723) will demonstrate broadcast, streaming, and AI-driven media workflow solutions at NAB Show 2026. GenAI and Agentic AI Ateme will demonstrat...

30/03/2026

NAB 2026: Bitmovin's Player Web X Adds Advertising Support, Vertical Video, and Proprietary ABR Algorithm

Bitmovin has announced new capabilities for Player Web X, its web video player, ...

30/03/2026

NAB 2026: Brazil's Minister of Communications and FCC Commissioner To Speak

The 2026 NAB Show (April 18-22, exhibits April 19-22, Las Vegas Convention Center) will host Brazil's Minister of Communications, Frederico de Siqueira Filh...

30/03/2026

NAB 2026: EVS To Showcase Expanded Live Production Ecosystem

EVS will exhibit at NAB Show 2026 (Booth N1841), highlighting new products and updates across its live production portfolio, including the debut of T-Motion med...

30/03/2026

NAB 2026: Solid State Logic To Demonstrate Expanded Virtual System T Platform

Solid State Logic will demonstrate its virtualized System T platform at NAB Show 2026 (Booth C6907). Demonstrations will include the VTE1 virtual DSP engine, ne...

30/03/2026

NAB 2026: Globecast To Showcase Managed Media Services Approach

Globecast will exhibit at NAB Show 2026 (Booth W3335), highlighting its hybrid service model spanning satellite, IP, fiber, and cloud. The company will demonst...

30/03/2026

NAB 2026: IP Showcase Returns as IPMX Moves to Deployment

The Alliance for IP Media Solutions (AIMS), Advanced Media Workflow Association (AMWA), and the Video Services Forum (VSF) have announced that the IP Showcase w...

30/03/2026

NAB 2026: BBright To Demonstrate Single-Stream ST 2110 Playout

At NAB Show 2026 BBright will present a demonstration of its One Stream for the World concept, showing how a single ST 2110 playout stream can simultaneously ...

30/03/2026

NAB 2026: OpenDrives To Demonstrate New Storage and Edge Products

OpenDrives will demonstrate new products at NAB Show 2026, with two locations in the West Hall: a pod (W3443-E) in the Sports Business Hub and a cabana at W1158...

30/03/2026

Behind the Mic: Amazon Prime Hosts 90th Master Tournament With Host Terry Gannon

Behind The Mic provides a roundup of recent news regarding on-air talent, including new deals, departures, and assignments compiled from press releases and repo...

30/03/2026

Op-Ed: Preparing for Agentic AI in Live Sports

The economics of live sports streaming have changed. New rights models, cloud production tools, and lower-cost distribution have made it possible for high schoo...

30/03/2026

Movimento Strings from Sonora Cinematic

MPE-capable chamber strings library announced Alongside their collection of Kontakt instruments, Sonora Cinematic have been steadily introducing a series of...

30/03/2026

UJAM release Groovemate Latigo

Latin-inspired percussion instrument announced Built on a newly developed engine and interface, UJAM's latest instrument has been designed to create Lat...

30/03/2026

Best Service launch Desert Winds

Latest Eduardo Tarilonte collaboration announced The latest library to join Best Service's ever-growing range includes four solo wind instruments that c...