• DocumentCode
    2574093
  • Title

    Creating visual vocabulary based on SIFT descriptor in compressed domain

  • Author

    Sui, Lei ; Zhang, Jing ; Zhuo, Li ; Yang, Yuncong

  • Author_Institution
    Signal & Inf. Process. Lab., Beijing Univ. of Technol., Beijing, China
  • fYear
    2011
  • fDate
    9-11 Nov. 2011
  • Firstpage
    1
  • Lastpage
    5
  • Abstract
    Recently bag-of-words (BoW) model having been widely used in textual information processing has been extended into many tasks in visual domain such as image classification, scene analysis, image annotation and image retrieval, namely bag-of-visual-words (BoVW) model. Therefore, it is essential to create an effective visual vocabulary. Most of existing approaches create visual vocabularies from image in pixel domain, which requires extra processing time for decompressed images, since most images are stored in compressed format. In this paper we propose to create a visual vocabulary based on Scale Invariant Feature Transform (SIFT) descriptor in compressed domain with the following three steps, (1) constructing low-resolution images in compressed domain; (2) extracting SIFT descriptor from low-resolution images; and (3) creating a visual vocabulary based on extracted SIFT descriptors. In order to evaluate the performance of the visual words, experiments have been conducted on identifying pornographic images. Experimental results indicate that the proposed method can recognize pornographic images accurately with much reduced computational time.
  • Keywords
    data compression; image coding; image resolution; text analysis; transforms; BoVW model; BoW model; bag-of-visual-words; compressed domain; extracted SIFT descriptors; image annotation; image classification; image decompression; image retrieval; low-resolution image construction; pixel domain; pornographic images; scale invariant feature transform; scene analysis; textual information processing; visual domain; visual vocabulary; Conferences; Discrete cosine transforms; Feature extraction; Image coding; Image recognition; Visualization; Vocabulary; SIFT descriptor; bag-of-words; compressed domian; image recognition; visual words;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Wireless Communications and Signal Processing (WCSP), 2011 International Conference on
  • Conference_Location
    Nanjing
  • Print_ISBN
    978-1-4577-1009-4
  • Electronic_ISBN
    978-1-4577-1008-7
  • Type

    conf

  • DOI
    10.1109/WCSP.2011.6096718
  • Filename
    6096718