ODESSA/PLUMCOT at Albayzin Multimodal Diarization Challenge 2018

Abstract : This paper describes ODESSA and PLUMCOT submissions to Albayzin Multimodal Diarization Challenge 2018. Given a list of people to recognize (alongside image and short video samples of those people), the task consists in jointly answering the two questions “who speaks when?” and “who appears when?”. Both consortia submitted 3 runs (1 primary and 2 contrastive) based on the same underlying mono-modal neural technologies : neural speaker segmentation, neural speaker embeddings, neural face embeddings, and neural talking-face detection. Our submissions aim at showing that face clustering and recognition can (hopefully) help to improve speaker diarization.
Type de document :
Communication dans un congrès
IberSPEECH 2018, 2018, Barcelona, Spain. pp.194--198, 2018
Liste complète des métadonnées

https://hal.archives-ouvertes.fr/hal-01987807
Contributeur : Hervé Bredin <>
Soumis le : lundi 21 janvier 2019 - 12:36:38
Dernière modification le : mardi 12 février 2019 - 01:30:17

Identifiants

  • HAL Id : hal-01987807, version 1

Citation

Benjamin Maurice, Hervé Bredin, Ruiqing Yin, Jose Patino, Héctor Delgado, et al.. ODESSA/PLUMCOT at Albayzin Multimodal Diarization Challenge 2018. IberSPEECH 2018, 2018, Barcelona, Spain. pp.194--198, 2018. 〈hal-01987807〉

Partager

Métriques

Consultations de la notice

12