HAL will be down for maintenance from Friday, June 10 at 4pm through Monday, June 13 at 9am. More information
Skip to Main content Skip to Navigation
Conference papers

What Makes a Speaker Recognizable in TV Broadcast? Going Beyond Speaker Identification Error Rate

Abstract : Speaker identification approaches for TV broadcast are usually evaluated and compared based on global error rates derived from the overall duration of missed detection, false alarm and confusion. Based on the analysis of the output of the systems submitted to the final round of the French evaluation campaign REPERE, this paper highlights the fact that these average met-rics lead to the incorrect intuition that current state-of-the-art algorithms partially recognize all speakers. Setting aside incorrect diarization and adverse acoustic conditions, we show that their performance is in fact essentially bi-modal: in a given show, either all speech turns of a speaker are correctly identified or none of them are. We then proceed with trying to understand and explain this behavior, through perfomance prediction experiments. These experiments show that the most discriminant speaker characteristics are – first – their total speech duration in the current show and – then only – the amount of training data available to build their acoustic model.
Document type :
Conference papers
Complete list of metadata

Cited literature [27 references]  Display  Hide  Download

Contributor : Sylvain Meignier Connect in order to contact the contributor
Submitted on : Thursday, April 6, 2017 - 8:59:44 AM
Last modification on : Tuesday, March 15, 2022 - 3:21:33 AM
Long-term archiving on: : Friday, July 7, 2017 - 12:30:14 PM


Publisher files allowed on an open archive


  • HAL Id : hal-01433205, version 1


Delphine Charlet, Johann Poignant, Hervé Bredin, Corinne Fredouille, Sylvain Meignier. What Makes a Speaker Recognizable in TV Broadcast? Going Beyond Speaker Identification Error Rate. ERRARE Workshop, a satellite event of Interspeech 2015., 2015, Sinaia, Romania. ⟨hal-01433205⟩



Record views


Files downloads