Tech Papers 2022: This paper introduces a semi-supervised TTS system that trains from broadcast commentary to enable automated program production.
Abstract
Thanks to the rapid progress in deep learning technology, the text-to-speech (TTS) system we developed has achieved the same quality as human speech, enabling us to launch a fully automatic program production system known as “AI Anchor.” The TTS system needs a large amount of speech and label data, but data production costs are high and TTS speakers cannot be easily added. This paper presents a novel TTS method featuring automatic training from broadcast commentary. It uses an approach that allows for a new semi-supervised learning method using an accentual data recognition method specialized for TTS.
Exclusive Content
This article is available with a Technical Paper Pass
Dynamic streaming content packaging with C2PA
Tech Papers 2026: This paper presents an implementation of the approach adopted by C2PA for live video to dynamic packaging.
Dynamic power control for sustainable broadcast transmitter networks
Tech Papers 2026: This paper proposes an approach that uses predictive modelling in combination with real-time interference monitoring to optimise transmitter powers dynamically, with minimal impact on the consumer.
A standardised framework for C2PA provenance in media workflows
Tech Papers 2026: This paper presents the first standardised framework for implementing C2PA for media provenance across newsrooms of varying sizes and operational contexts.
Building MXL together: Progress of the multi-vendor open-source media exchange SDK
Tech Papers 2026: This paper presents the Media eXchange Layer (MXL), an open-source SDK project hosted by the Linux Foundation in collaboration with the EBU and NABA.
Search first, inspect visually when needed: A multi-agent architecture for semantic video archival retrieval
Tech Papers 2026: This paper introduces Smart Chat, a multi-agent video question-answering system that searches indexed video moments, localizes candidate evidence, and inspects a short clip only when visual verification is needed.