Multi-label Image Classification via Coarse-to-Fine Attention

LYU Fan; LI Linyan; Victor S. Sheng; FU Qiming; HU Fuyuan

doi:10.1049/cje.2019.07.015

LYU Fan, LI Linyan, Victor S. Sheng, FU Qiming, HU Fuyuan. Multi-label Image Classification via Coarse-to-Fine Attention[J]. Chinese Journal of Electronics, 2019, 28(6): 1118-1126. DOI: 10.1049/cje.2019.07.015

Citation:

Multi-label Image Classification via Coarse-to-Fine Attention

Abstract

Abstract

Great efforts have been made by using deep neural networks to recognize multi-label images. Since multi-label image classification is very complicated, many studies seek to use the attention mechanism as a kind of guidance. Conventional attention-based methods always analyzed images directly and aggressively, which is difficult to well understand complicated scenes. We propose a global/local attention method that can recognize a multi-label image from coarse to fine by mimicking how human-beings observe images. Our global/local attention method first concentrates on the whole image, and then focuses on its local specific objects. We also propose a joint max-margin objective function, which enforces that the minimum score of positive labels should be larger than the maximum score of negative labels horizontally and vertically. This function further improve our multi-label image classification method. We evaluate the effectiveness of our method on two popular multi-label image datasets (i.e., Pascal VOC and MS-COCO). Our experimental results show that our method outperforms state-of-the-art methods.

FullText(HTML)

References (55)

Cited By

Multi-label Image Classification via Coarse-to-Fine Attention

Abstract

Catalog

Links

Chinese Journal of Electronics

Multi-label Image Classification via Coarse-to-Fine Attention

Abstract

Catalog

Links

Chinese Journal of Electronics

Export File

Citation

Format

Content