Literature DB >> 32149687

A Multiple-Instance Densely-Connected ConvNet for Aerial Scene Classification.

Qi Bi, Kun Qin, Zhili Li, Han Zhang, Kai Xu, Gui-Song Xia.   

Abstract

In contrast with nature scenes, aerial scenes are often composed of many objects crowdedly distributed on the surface in bird's view, the description of which usually demands more discriminative features as well as local semantics. However, when applied to scene classification, most of the existing convolution neural networks (ConvNets) tend to depict global semantics of images, and the loss of low- and mid-level features can hardly be avoided, especially when the model goes deeper. To tackle these challenges, in this paper, we propose a multiple-instance densely-connected ConvNet (MIDC-Net) for aerial scene classification. It regards aerial scene classification as a multiple-instance learning problem so that local semantics can be further investigated. Our classification model consists of an instance-level classifier, a multiple instance pooling and followed by a bag-level classification layer. In the instance-level classifier, we propose a simplified dense connection structure to effectively preserve features from different levels. The extracted convolution features are further converted into instance feature vectors. Then, we propose a trainable attention-based multiple instance pooling. It highlights the local semantics relevant to the scene label and outputs the bag-level probability directly. Finally, with our bag-level classification layer, this multiple instance learning framework is under the direct supervision of bag labels. Experiments on three widely-utilized aerial scene benchmarks demonstrate that our proposed method outperforms many state-of-the-art methods by a large margin with much fewer parameters.

Year:  2020        PMID: 32149687     DOI: 10.1109/TIP.2020.2975718

Source DB:  PubMed          Journal:  IEEE Trans Image Process        ISSN: 1057-7149            Impact factor:   10.856


  1 in total

1.  Aerial scene understanding in the wild: Multi-scene recognition via prototype-based memory networks.

Authors:  Yuansheng Hua; Lichao Mou; Jianzhe Lin; Konrad Heidler; Xiao Xiang Zhu
Journal:  ISPRS J Photogramm Remote Sens       Date:  2021-07       Impact factor: 8.979

  1 in total

北京卡尤迪生物科技股份有限公司 © 2022-2023.