Literature DB >> 22128004

Task-dependent visual-codebook compression.

Rongrong Ji1, Hongxun Yao, Wei Liu, Xiaoshuai Sun, Qi Tian.   

Abstract

A visual codebook serves as a fundamental component in many state-of-the-art computer vision systems. Most existing codebooks are built based on quantizing local feature descriptors extracted from training images. Subsequently, each image is represented as a high-dimensional bag-of-words histogram. Such highly redundant image description lacks efficiency in both storage and retrieval, in which only a few bins are nonzero and distributed sparsely. Furthermore, most existing codebooks are built based solely on the visual statistics of local descriptors, without considering the supervise labels coming from the subsequent recognition or classification tasks. In this paper, we propose a task-dependent codebook compression framework to handle the above two problems. First, we propose to learn a compression function to map an originally high-dimensional codebook into a compact codebook while maintaining its visual discriminability. This is achieved by a codeword sparse coding scheme with Lasso regression, which minimizes the descriptor distortions of training images after codebook compression. Second, we propose to adapt our codebook compression to the subsequent recognition or classification tasks. This is achieved by introducing a label constraint kernel (LCK) into our compression loss function. In particular, our LCK can model heterogeneous kinds of supervision, i.e., (partial) category labels, correlative semantic annotations, and image query logs. We validated our codebook compression in three computer vision tasks: 1) object recognition in PASCAL Visual Object Class 07; 2) near-duplicate image retrieval in UKBench; and 3) web image search in a collection of 0.5 million Flickr photographs. Our compressed codebook has shown superior performances over several state-of-the-art supervised and unsupervised codebooks.

Mesh:

Year:  2011        PMID: 22128004     DOI: 10.1109/TIP.2011.2176950

Source DB:  PubMed          Journal:  IEEE Trans Image Process        ISSN: 1057-7149            Impact factor:   10.856


  11 in total

1.  An Efficient Augmented Lagrangian Method for Statistical X-Ray CT Image Reconstruction.

Authors:  Jiaojiao Li; Shanzhou Niu; Jing Huang; Zhaoying Bian; Qianjin Feng; Gaohang Yu; Zhengrong Liang; Wufan Chen; Jianhua Ma
Journal:  PLoS One       Date:  2015-10-23       Impact factor: 3.240

2.  A time-critical adaptive approach for visualizing natural scenes on different devices.

Authors:  Tianyang Dong; Siyuan Liu; Jiajia Xia; Jing Fan; Ling Zhang
Journal:  PLoS One       Date:  2015-02-27       Impact factor: 3.240

3.  Parameter estimation of fractional-order chaotic systems by using quantum parallel particle swarm optimization algorithm.

Authors:  Yu Huang; Feng Guo; Yongling Li; Yufeng Liu
Journal:  PLoS One       Date:  2015-01-20       Impact factor: 3.240

4.  A Probabilistic Analysis of Sparse Coded Feature Pooling and Its Application for Image Retrieval.

Authors:  Yunchao Zhang; Jing Chen; Xiujie Huang; Yongtian Wang
Journal:  PLoS One       Date:  2015-07-01       Impact factor: 3.240

5.  Remote safety monitoring for elderly persons based on omni-vision analysis.

Authors:  Yun Xiang; Yi-Ping Tang; Bao-Qing Ma; Hang-Chen Yan; Jun Jiang; Xu-Yuan Tian
Journal:  PLoS One       Date:  2015-05-15       Impact factor: 3.240

6.  A lightweight distributed framework for computational offloading in mobile cloud computing.

Authors:  Muhammad Shiraz; Abdullah Gani; Raja Wasim Ahmad; Syed Adeel Ali Shah; Ahmad Karim; Zulkanain Abdul Rahman
Journal:  PLoS One       Date:  2014-08-15       Impact factor: 3.240

7.  A combined approach to cartographic displacement for buildings based on skeleton and improved elastic beam algorithm.

Authors:  Yuangang Liu; Qingsheng Guo; Yageng Sun; Xiaoya Ma
Journal:  PLoS One       Date:  2014-12-03       Impact factor: 3.240

8.  Robust Optical Recognition of Cursive Pashto Script Using Scale, Rotation and Location Invariant Approach.

Authors:  Riaz Ahmad; Saeeda Naz; Muhammad Zeshan Afzal; Sayed Hassan Amin; Thomas Breuel
Journal:  PLoS One       Date:  2015-09-14       Impact factor: 3.240

9.  A topic clustering approach to finding similar questions from large question and answer archives.

Authors:  Wei-Nan Zhang; Ting Liu; Yang Yang; Liujuan Cao; Yu Zhang; Rongrong Ji
Journal:  PLoS One       Date:  2014-03-04       Impact factor: 3.240

10.  On-device mobile visual location recognition by using panoramic images and compressed sensing based visual descriptors.

Authors:  Tao Guan; Yin Fan; Liya Duan; Junqing Yu
Journal:  PLoS One       Date:  2014-06-03       Impact factor: 3.240

View more

北京卡尤迪生物科技股份有限公司 © 2022-2023.