PLDA-based mean shift speakers' short segments clustering. (September 2017)
- Record Type:
- Journal Article
- Title:
- PLDA-based mean shift speakers' short segments clustering. (September 2017)
- Main Title:
- PLDA-based mean shift speakers' short segments clustering
- Authors:
- Salmun, Itay
Shapiro, Ilya
Opher, Irit
Lapidot, Itshak - Abstract:
- Highlights: Use of PLDA as a criterion to choose the best i-vectors for the new mean. Replacing the constant threshold for as the i-vector's neighborhood by the kNN which dramatically increased the stability of the system. Abstract: This paper extends upon a previous work using Mean Shift algorithm to perform speaker clustering on i-vectors generated from short speech segments. In this paper we examine the effectiveness of probabilistic linear discriminant analysis (PLDA) scoring as the metric of the mean shift clustering algorithm in the presence of different numbers of speakers. Our proposed method, combined with k-nearest neighbors (kNN) for bandwidth estimation, yields better and more robust results in comparison to the cosine similarity with fixed neighborhood bandwidth for clustering segments of large numbers of speakers. In the case of 30 speakers, we achieved significant improvement in cluster and speaker purity with the PLDA-based mean shift algorithm compared to the cosine-based baseline system.
- Is Part Of:
- Computer speech & language. Volume 45(2017)
- Journal:
- Computer speech & language
- Issue:
- Volume 45(2017)
- Issue Display:
- Volume 45, Issue 2017 (2017)
- Year:
- 2017
- Volume:
- 45
- Issue:
- 2017
- Issue Sort Value:
- 2017-0045-2017-0000
- Page Start:
- 411
- Page End:
- 436
- Publication Date:
- 2017-09
- Subjects:
- Speaker clustering -- Mean shift clustering -- Probabilistic linear discriminant analysis -- Two-covariance model -- K-nearest neighbors -- I-vectors -- Short segments
Speech processing systems -- Periodicals
Automatic speech recognition -- Periodicals
Computers -- Periodicals
Linguistics -- Periodicals
Speech-Language Pathology -- Periodicals
Traitement automatique de la parole -- Périodiques
Reconnaissance automatique de la parole -- Périodiques
Automatic speech recognition
Speech processing systems
Electronic journals
Periodicals
006.454 - Journal URLs:
- http://www.journals.elsevier.com/computer-speech-and-language/ ↗
http://www.elsevier.com/journals ↗ - DOI:
- 10.1016/j.csl.2017.04.006 ↗
- Languages:
- English
- ISSNs:
- 0885-2308
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 3394.276600
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 2060.xml