Linguistic influences on bottom-up and top-down clustering for speaker diarizationReportar como inadecuado

Linguistic influences on bottom-up and top-down clustering for speaker diarization - Descarga este documento en PDF. Documentación en PDF para descargar gratis. Disponible también para leer online.

1 Eurecom Sophia Antipolis

Abstract : While bottom-up approaches have emerged as the standard, default approach to clustering for speaker diarization we have always found the top-down approach gives equivalent or superior performance. Our recent work shows that significant gains in performance can be obtained when cluster purification is applied to the output of topdown systems but that it can degrade performance when applied to the output of bottom-up systems. This paper demonstrates that these observations can be accounted for by factors unrelated to the speaker and that they can impact more strongly on the performance of bottom-up clustering strategies than top-down strategies. Experimental results confirm that clusters produced through top-down clustering are better normalized against phone variation than those produced through bottom-up clustering and that this accounts for the observed inconsistencies in purification performance. The work highlights the need for marginalization strategies which should encourage convergence toward different speakers rather than toward nuisance factors such as that those related to the linguistic content.

Keywords : Speaker diarization hierarchical clustering purification phone normalization

Autor: Simon Bozonnet - Dong Wang - Nicholas Evans - Raphaël Troncy -



Documentos relacionados