DocumentCode :
3495006
Title :
Audio source type segmentation using a perceptually based representation
Author :
Melih, Kathy ; Gonzalez, Ruben
Author_Institution :
Griffith Univ., Gold Coast, Australia
Volume :
1
fYear :
1999
fDate :
1999
Firstpage :
51
Abstract :
Existing audio retrieval systems fall into one of two categories: systems that can accept data of only a single type (e.g. automatic speech recognition systems) or systems that report to offer content based retrieval for audio data of any type. However, systems belonging to the latter category often impose the restriction that only one type of sound can be presented at a time. This requirement is reasonable since the interpretation of various audio qualities such as pitch and rhythm depends upon the audio type. Pitch variation, for example, can be interpreted as the melody line in music while in speech it can be used as a means for detecting change of speaker. The problem, however, is that existing systems either expect segmentation to have been performed a priori or perform the segmentation in a completely separate process. This introduces unnecessary processing and file manipulation overheads. To combat this, a new perceptually based representation has been developed specifically to support content-based retrieval. This paper discusses the application of the new representation to sound source segmentation and identification
Keywords :
audio signal processing; content-based retrieval; hearing; signal representation; audio qualities; audio retrieval systems; audio source type segmentation; automatic speech recognition systems; content based retrieval; content-based retrieval; human auditory system; melody line; music; perceptually based representation; pitch variation; rhythm; sound source identification; sound source segmentation; speaker change detection; Audio recording; Automatic speech recognition; Content based retrieval; Gold; Information retrieval; Music information retrieval; Performance analysis; Rhythm; Video recording; Web sites;
fLanguage :
English
Publisher :
ieee
Conference_Titel :
Signal Processing and Its Applications, 1999. ISSPA '99. Proceedings of the Fifth International Symposium on
Conference_Location :
Brisbane, Qld.
Print_ISBN :
1-86435-451-8
Type :
conf
DOI :
10.1109/ISSPA.1999.818110
Filename :
818110
Link To Document :
بازگشت