Taseska M, Lamani G, Habets E (2016)
Publication Language: English
Publication Type: Authored book
Publication year: 2016
Series: Lecture Notes in Electrical Engineering (LNEE)
Pages Range: 59-69
Edition: 387
ISBN: 9783319322124
DOI: 10.1007/978-3-319-32213-1_6
Speaker detection, localization and tracking are required in systems that involve e.g. hands-free speech acquisition, or blind source separation. Localization can be done in the (TF) domain, where location features extracted using microphone arrays are used to cluster the TF bins corresponding to the same source. The TF clustering approaches provide an alternative to the Bayesian tracking approaches that are based on Kalman and particle filters. In this work, we propose a maximum-likelihood approach where detection, localization, and tracking are achieved by online clustering of narrowband position estimates, while incorporating the speech presence probability at each TF bin in a unified manner.
APA:
Taseska, M., Lamani, G., & Habets, E. (2016). Online clustering of narrowband position estimates with application to multi-speaker detection and tracking.
MLA:
Taseska, Maja, Gleni Lamani, and Emanuël Habets. Online clustering of narrowband position estimates with application to multi-speaker detection and tracking. 2016.
BibTeX: Download