This paper addresses the need for a consistent evaluation framework designed specifically for video content.
The increasing applicability of facial recognition technology (FRT) in the broadcast and media industries necessitates a standardised evaluation framework specifically designed for video content. The absence of such a framework poses challenges in the decision-making process regarding the implementation of facial recognition systems, as reliance on conventional Machine Learning (ML) metrics may lead to suboptimal choices. In fact, these metrics prioritise performance optimization, which can inadvertently overlook user-centric properties essential for practical applications and result in masking critical user-centric properties, such as the relevance and accessibility of the metadata produced.
To address this gap, the EBU has developed a benchmark tailored for facial recognition in television programming, accompanied by a state-of-the-art AI model optimised for this framework. This initiative involved the extensive annotation of a video dataset guided by user-centric metrics, which prioritise the accurate retrieval of relevant personalities in accordance with the requirements of documentalists. Our strategy improves...
Exclusive Content
This article is available with a Technical Paper Pass
Next generation video compression standards
Tech Papers 2026: This paper presents an overview of the design criteria and development goals for a new video compression standardisation project.
A low-latency interactive system for real-time video understanding based on VLMs
Abstract: This paper introduces a unified edge-cloud system for real-time video vision-language model applications.
Selective multi-pass encoding for cost-effective video streaming
Tech Papers 2026: This paper presents a content-adaptive strategy, CASE, that predicts whether additional encoding passes would provide meaningful gains using a lightweight mechanism that derives spatial and temporal features from each video segment.
Scalable SSIM estimation from PSNR for per-title and context adaptive encoding workflows
Tech Papers 2026: This paper proposes ApproxSSIMate, a low-complexity method for estimating SSIM from PSNR combined with reference-sequence statistics.
Feasibility and deployment strategies for cloud-based AOIP audio consoles
Tech Papers 2026: This paper investigates the feasibility of cloud-based audio-mixing systems by re-examining existing assumptions about network conditions, multicast transport and synchronisation.



