On the use of nearest feature line for speaker identification

Ke Chen, Ting Yao Wu, Hong Jiang Zhang

    Research output: Contribution to journalArticlepeer-review

    Abstract

    As a new pattern classification method, nearest feature line (NFL) provides an effective way to tackle the sort of pattern recognition problems where only limited data are available for training. In this paper, we explore the use of NFL for speaker identification in terms of limited data and examine how the NFL performs in such a vexing problem of various mismatches between training and test. In order to speed up NFL in decision-making, we propose an alternative method for similarity measure. We have applied the improved NFL to speaker identification of different operating modes. Its text-dependent performance is better than the dynamic time warping (DTW) on the Ti46 corpus, while its computational load is much lower than that of DTW. Moreover, we propose an utterance partitioning strategy used in the NFL for better performance. For the text-independent mode, we employ the NFL to be a new similarity measure in vector quantization (VQ), which causes the VQ to perform better on the KING corpus. Some computational issues on the NFL are also discussed in this paper. © Elsevier Science B.V. All rights reserved.
    Original languageEnglish
    Pages (from-to)1735-1746
    Number of pages11
    JournalPattern Recognition Letters
    Volume23
    Issue number14
    DOIs
    Publication statusPublished - Dec 2002

    Keywords

    • Dynamic time warping
    • Nearest feature line
    • Nearest neighboring measure
    • Speaker identification
    • Vector quantization

    Fingerprint

    Dive into the research topics of 'On the use of nearest feature line for speaker identification'. Together they form a unique fingerprint.

    Cite this