arXiv:2107.11159 Abstract | arXiv Analytics

arXiv:2107.11159 [cs.CV]Abstract References Reviews Resources

Learning Discriminative Representations for Multi-Label Image Recognition

Mohammed Hassanin, Ibrahim Radwan, Salman Khan, Murat Tahtali

Published 2021-07-23Version 1

Multi-label recognition is a fundamental, and yet is a challenging task in computer vision. Recently, deep learning models have achieved great progress towards learning discriminative features from input images. However, conventional approaches are unable to model the inter-class discrepancies among features in multi-label images, since they are designed to work for image-level feature discrimination. In this paper, we propose a unified deep network to learn discriminative features for the multi-label task. Given a multi-label image, the proposed method first disentangles features corresponding to different classes. Then, it discriminates between these classes via increasing the inter-class distance while decreasing the intra-class differences in the output space. By regularizing the whole network with the proposed loss, the performance of applying the wellknown ResNet-101 is improved significantly. Extensive experiments have been performed on COCO-2014, VOC2007 and VOC2012 datasets, which demonstrate that the proposed method outperforms state-of-the-art approaches by a significant margin of 3:5% on large-scale COCO dataset. Moreover, analysis of the discriminative feature learning approach shows that it can be plugged into various types of multi-label methods as a general module.

Categories: cs.CV

Keywords: multi-label image recognition, learning discriminative representations, discriminative feature, method outperforms state-of-the-art approaches, method first disentangles features corresponding

Related articles: Most relevant | Search more

arXiv:2407.20920 [cs.CV] (Published 2024-07-30)

SSPA: Split-and-Synthesize Prompting with Gated Alignments for Multi-Label Image Recognition

Hao Tan, Zichang Tan, Jun Li, Jun Wan, Zhen Lei, Stan Z. Li

arXiv:2205.13092 [cs.CV] (Published 2022-05-26)

Semantic-Aware Representation Blending for Multi-Label Image Recognition with Partial Labels

Tao Pu, Tianshui Chen, Hefeng Wu, Yongyi Lu, Liang Lin

arXiv:2204.03795 [cs.CV] (Published 2022-04-08)

Semantic Representation and Dependency Learning for Multi-Label Image Recognition

Tao Pu, Lixian Yuan, Hefeng Wu, Tianshui Chen, Ling Tian, Liang Lin