Harmonic Cues for Number of Simultaneous Speakers Estimation

Rafi, Umer; Bardeli, Rolf

AES E-Library

Harmonic Cues for Number of Simultaneous Speakers Estimation

Overlapped speech, where several speakers are speaking simultaneously, is a common occurence in multiparty discussions such as meetings. This kind of speech presents a great challenge to automatic speech processing systems such as speech recognition systems and speaker diarisation systems. In recent speaker diarisation systems, a large portion of the remaining error comes from overlapped speech. So far little work has been done on detecting overlapped speech and the number of speakers present in overlapped speech. In this paper we first describe a model-based approach for estimating the number of simultaneous speakers. Then, we propose a new approach called Spectral Peak Clustering where instead of training statistical models we extract spectral peaks from the input data and then cluster them into components by using a similarity measure between peaks where each component represents a speaker present in the input data.

Authors: Rafi, Umer; Bardeli, Rolf
Affiliation: Fraunhofer Institute for Intellegent Analysis and Information Systems, Fraunhofer IAIS, Sankt Augustin, Germany
AES Conference: 53rd International Conference: Semantic Audio (January 2014)
Paper Number: P1-12
Publication Date: January 27, 2014 Import into BibTeX
Permalink: https://www.aes.org/e-lib/browse.cfm?elib=17088

Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member and would like to subscribe to the E-Library then Join the AES!

This paper costs $33 for non-members and is free for AES members and E-Library subscribers.

Learn more about the AES E-Library

E-Library Location: (CD 53rdPapers) /conf/53/aes53-000020.pdf

Start a discussion about this paper!

AES E-Library

Harmonic Cues for Number of Simultaneous Speakers Estimation

ABOUT AES

Contact Us