arXiv Analytics

Sign in

arXiv:1902.09941 [cs.CV]AbstractReferencesReviewsResources

Unsupervised Part Mining for Fine-grained Image Classification

Jian Zhang, Runsheng Zhang, Yaping Huang, Qi Zou

Published 2019-02-26Version 1

Fine-grained image classification remains challenging due to the large intra-class variance and small inter-class variance. Since the subtle visual differences are only in local regions of discriminative parts among subcategories, part localization is a key issue for fine-grained image classification. Most existing approaches localize object or parts in an image with object or part annotations, which are expensive and labor-consuming. To tackle this issue, we propose a fully unsupervised part mining (UPM) approach to localize the discriminative parts without even image-level annotations, which largely improves the fine-grained classification performance. We first utilize pattern mining techniques to discover frequent patterns, i.e., co-occurrence highlighted regions, in the feature maps extracted from a pre-trained convolutional neural network (CNN) model. Inspired by the fact that these relevant meaningful patterns typically hold appearance and spatial consistency, we then cluster the mined regions to obtain the cluster centers and the discriminative parts surrounding the cluster centers are generated. Importantly, any annotations and sophisticated training procedures are not used in our proposed part localization approach. Finally, a multi-stream classification network is built for aggregating the original, object-level and part-level features simultaneously. Compared with other state-of-the-art approaches, our UPM approach achieves the competitive performance.

Related articles: Most relevant | Search more
arXiv:2005.10979 [cs.CV] (Published 2020-05-22)
Focus Longer to See Better:Recursively Refined Attention for Fine-Grained Image Classification
arXiv:2409.03192 [cs.CV] (Published 2024-09-05)
PEPL: Precision-Enhanced Pseudo-Labeling for Fine-Grained Image Classification in Semi-Supervised Learning
arXiv:2311.04157 [cs.CV] (Published 2023-11-07)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis