• DocumentCode
    3718252
  • Title

    Influence of simultaneous spoken sentences on the properties of spectral peaks

  • Author

    Tomasz Maka;Miroslaw Lazoryszczak

  • Author_Institution
    Faculty of Computer Science and Information Technology, West Pomeranian University of Technology, Szczecin, Zolnierska 52, 71-210, Poland
  • fYear
    2015
  • Firstpage
    87
  • Lastpage
    92
  • Abstract
    In this study, an approach to analyse the properties of spectral peaks of simultaneously talking speakers in monophonic audio signal has been described. We have proposed a technique based on spectral peaks tracking and attributes calculated from peaks histogram. Spectral peaks have been estimated using linear prediction-based spectral envelope for each frame of source signal. The features have been computed from the histogram at different frequency bands. The statistical properties of the obtained features have been used to find out the relationship with the number of speech sources. Proposed approach has been tested using a dedicated database featuring sentences with the same and mixed gender, where the number of speakers varies from two to twelve. Different configuration parameters like frame size, bin width of the histogram and linear prediction order have been used in the conducted experiments. The results show that obtained trends of statistical descriptors are directly connected with the number of voice sources. The proposed descriptors and performed regression analysis can be a basis to estimate the number of speakers in single audio stream.
  • Keywords
    "Mixers","Bandwidth"
  • Publisher
    ieee
  • Conference_Titel
    Signal Processing: Algorithms, Architectures, Arrangements, and Applications (SPA), 2015
  • ISSN
    2326-0262
  • Electronic_ISBN
    2326-0319
  • Type

    conf

  • DOI
    10.1109/SPA.2015.7365139
  • Filename
    7365139