New bilingual speech databases for audio diarization

LREC 2014 · David Tavarez, Eva Navas, Daniel Erro, Ibon Saratxaga, Inma Hernaez ·

This paper describes the process of collecting and recording two new bilingual speech databases in Spanish and Basque. They are designed primarily for speaker diarization in two different application domains: broadcast news audio and recorded meetings. First, both databases have been manually segmented. Next, several diarization experiments have been carried out in order to evaluate them. Our baseline speaker diarization system has been applied to both databases with around 30{\%} of DER for broadcast news audio and 40{\%} of DER for recorded meetings. Also, the behavior of the system when different languages are used by the same speaker has been tested.

PDF Abstract