Fine Grained Image Classification Combining Dynamic Adaptive Modulation and Structural Relationship Learning

doi:10.15888/j.cnki.csa.009571

AIPUB归智期刊联盟

WeChat

Mobile website

2025-8-3- 0

Home > Archive>Volume 33, Issue 8, 2024 >166-175. DOI:10.15888/j.cnki.csa.009571

PDF HTML XML Export Cite reminder

Fine Grained Image Classification Combining Dynamic Adaptive Modulation and Structural Relationship Learning
DOI:
                        10.15888/j.cnki.csa.009571
                    
CSTR:
                        32024.14.csa.009571
                    
Author:
                        WANG Yan-GenWANG Yan-Gen
College of Computer and Data Science, Fuzhou University, Fuzhou 35108, China
Find this author on All Journals
Find this author on BaiDu
Search for this author on this site
CHEN FeiCHEN Fei
College of Computer and Data Science, Fuzhou University, Fuzhou 35108, China
Find this author on All Journals
Find this author on BaiDu
Search for this author on this site
CHEN QuanCHEN Quan
College of Computer and Data Science, Fuzhou University, Fuzhou 35108, China
Find this author on All Journals
Find this author on BaiDu
Search for this author on this site

                    
Affiliation:
Clc Number:
Fund Project:

Article

Figures

Metrics

Reference

Cited by

Materials

Comments

Abstract:

Due to the small inter-class differences and large intra-class differences of fine-grained images, the key to fine-grained image classification tasks is to find subtle differences between categories. Recently, Vision Transformer-based networks mostly focus on mining the most prominent discriminative region features in images. There are two problems with this. Firstly, the network ignores mining classification clues from other discriminative regions, which can easily confuse similar categories. secondly, the structural relationships of images are ignored, resulting in inaccurate extraction of category features. To solve the above problems, this study proposes two modules: dynamic adaptive modulation and structural relationship learning. The dynamic adaptive modulation module forces the network to search for multiple discriminative regions, and then the structural relationship learning module is used to construct structural relationships between discriminative regions. Finally, the graph convolutional network is used to fuse semantic and structural information to obtain predicted classification results. The proposed method achieves testing accuracy of 92.9% and 93.0% on the CUB-200-2011 dataset and NA-Birds dataset, respectively, which is superior to existing state-of-the-art networks.

Key words:fine grained image classification;Vision Transformer (ViT);dynamic adaptive modulation;structural relationship learning;graph convolutional network (GCN)

Get Citation

王衍根,陈飞,陈权.结合动态自适应调制和结构关系学习的细粒度图像分类.计算机系统应用,2024,33(8):166-175

Copy

Article Metrics

Abstract:
PDF:
HTML:
Cited by:

History

Received:January 27,2024
Revised:February 29,2024
Adopted:
Online: June 28,2024
Published:

Article QR Code

You are the first1025910Visitors
Copyright: Institute of Software, Chinese Academy of Sciences Beijing ICP No. 05046678-3
Address：4# South Fourth Street, Zhongguancun,Haidian, Beijing,Postal Code：100190
Phone：010-62661041 Fax： Email：csa (a) iscas.ac.cn
Technical Support：Beijing Qinyun Technology Development Co., Ltd.

Beijing Public Network Security No. 11040202500063