Albayzín-2014 evaluation: audio segmentation and classification in broadcast news domains

TitleAlbayzín-2014 evaluation: audio segmentation and classification in broadcast news domains
Publication TypeJournal Article
Year of Publication2015
AuthorsCastán, D, Tavárez, D, López Otero, P, Franco, J, Delgado, H, Navas, E, Docío Fernández, L, Ramos, D, Serrano, J, Ortega, A, Lleida, E
JournalEURASIP Journal on Audio, Speech, and Music Processing
Volume2015
Date Published12/2015
AbstractAudio segmentation is important as a pre-processing task to improve the performance of many speech technology tasks and, therefore, it has an undoubted research interest. This paper describes the database, the metric, the systems and the results for the Albayzín-2014 audio segmentation campaign. In contrast to previous evaluations where the task was the segmentation of non-overlapping classes, Albayzín-2014 evaluation proposes the delimitation of the presence of speech, music and/or noise that can be found simultaneously. The database used in the evaluation was created by fusing different media and noises in order to increase the difficulty of the task. Seven segmentation systems from four different research groups were evaluated and combined. Their experimental results were analyzed and compared with the aim of providing a benchmark and showing up the promising directions in this field.
Citation Key568