• DocumentCode
    585678
  • Title

    Speaker identification using feature vector reduction of row mean of different transforms

  • Author

    Kekre, H.B. ; Kulkarni, Vaishali

  • Author_Institution
    MPSTME, NMIMS Univ., Mumbai, India
  • fYear
    2012
  • fDate
    19-20 Oct. 2012
  • Firstpage
    1
  • Lastpage
    5
  • Abstract
    In this paper a novel approach to text dependent speaker identification based on feature vector reduction technique of the row mean is proposed. Five different Orthogonal Transform Techniques: Discrete Fourier Transform (DFT), Discrete Cosine Transform (DCT), Discrete Sine Transform (DST), Discrete Hartley Transform (DHT) and Walsh Hadamard Transform (WHT) are applied on the framed speech signal. Feature extraction in the testing and matching phases has been done by using feature vector reduction technique applied on the row mean vector of the magnitude of the transformed speech signal. Two similarity measures Euclidean distance and Manhattan distance are used for feature matching. The results indicate that the accuracy using both the similarity measures remains steady up to certain reduction in feature vector permitting to reduce feature vector size. This algorithm is tested using two databases: a locally created database and CSLU Database. It is observed that, DFT allows maximum percentage of feature vector reduction. It out performs other Transforms with a big margin.
  • Keywords
    discrete Fourier transforms; feature extraction; speaker recognition; CSLU Database; DCT; DFT; DHT; DST; Discrete Fourier Transform; Euclidean distance; Manhattan distance; WHT; Walsh Hadamard Transform; different transforms; discrete Hartley transform; discrete cosine transform; discrete sine transform; feature extraction; feature matching; feature vector reduction; orthogonal transform techniques; row mean; speaker identification; speech signal; Accuracy; Databases; Discrete Fourier transforms; Discrete cosine transforms; Speech; Vectors; Euclidean distance; Manhattan distance; Row Mean; Speaker Recognition; Text Dependent Identification;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Communication, Information & Computing Technology (ICCICT), 2012 International Conference on
  • Conference_Location
    Mumbai
  • Print_ISBN
    978-1-4577-2077-2
  • Type

    conf

  • DOI
    10.1109/ICCICT.2012.6398100
  • Filename
    6398100