On the use of nearest feature line for speaker identification

Ke Chen, T Wu, H Zhang

Research output: Contribution to journalArticle

30 Citations (Scopus)

Abstract

As a new pattern classification method, nearest feature line (NFL) provides an effective way to tackle the sort of pattern recognition problems where only limited data are available for training. In this paper, we explore the use of NFL for speaker identification in terms of limited data and examine how the NFL performs in such a vexing problem of various mismatches between training and test. In order to speed up NFL in decision-making, we propose an alternative method for similarity measure. We have applied the improved NFL to speaker identification of different operating modes. Its text-dependent performance is better than the dynamic time warping (DTW) on the Ti46 corpus, while its computational load is much lower than that of DTW. Moreover, we propose an utterance partitioning strategy used in the NFL for better performance. For the text-independent mode, we employ the NFL to be a new similarity measure in vector quantization (VQ), which causes the VQ to perform better on the KING corpus. Some computational issues on the NFL are also discussed in this paper. (C) 2002 Elsevier Science B.V. All rights reserved.
Original languageEnglish
Pages (from-to)1735-1746
Number of pages12
JournalPattern Recognition Letters
Volume23
Issue number14
DOIs
Publication statusPublished - 1 Dec 2002

Fingerprint

Dive into the research topics of 'On the use of nearest feature line for speaker identification'. Together they form a unique fingerprint.

Cite this