Multimodal Fine-Grained Transformer Model for Pest Recognition

Yinshuo Zhang,Lei Chen,Yuan Yuan

doi:10.3390/electronics12122620

Abstract

Deep learning has shown great potential in smart agriculture, especially in the field of pest recognition. However, existing methods require large datasets and do not exploit the semantic associations between multimodal data. To address these problems, this paper proposes a multimodal fine-grained transformer (MMFGT) model, a novel pest recognition method that improves three aspects of transformer architecture to meet the needs of few-shot pest recognition. On the one hand, the MMFGT uses self-supervised learning to extend the transformer structure to extract target features using contrastive learning to reduce the reliance on data volume. On the other hand, fine-grained recognition is integrated into the MMFGT to focus attention on finely differentiated areas of pest images to improve recognition accuracy. In addition, the MMFGT further improves the performance in pest recognition by using the joint multimodal information from the pest’s image and natural language description. Extensive experimental results demonstrate that the MMFGT obtains more competitive results compared to other excellent models, such as ResNet, ViT, SwinT, DINO, and EsViT, in pest recognition tasks, with recognition accuracy up to 98.12% and achieving 5.92% higher accuracy compared to the state-of-the-art DINO method for the baseline.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Electronics	Publication Date: Jun 10, 2023
Citations: 3	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Multimodal Fine-Grained Transformer Model for Pest Recognition

Abstract

Talk to us

Similar Papers

More From: Electronics

Lead the way for us

Similar Papers

Fine-grained recognition algorithm of crop pests based on cross-layer bilinear aggregation and multi-task learning
Juquan Ruan ... Guangsun Yin
Journal of Agricultural Engineering | VOL. 55
Juquan Ruan, et. al.Juquan Ruan ... Guangsun Yin
30 Oct 2024
Journal of Agricultural Engineering | VOL. 55

Few-shot agricultural pest recognition based on multimodal masked autoencoder
Yinshuo Zhang ... Yuan Yuan
Crop Protection | VOL. -
Yinshuo Zhang, et. al.Yinshuo Zhang ... Yuan Yuan
01 Oct 2024
Crop Protection | VOL. -

Smart Farming Becomes Even Smarter With Deep Learning—A Bibliographical Analysis
Zeynep Unal
IEEE Access | VOL. 8
Zeynep UnalZeynep Unal
01 Jan 2020
IEEE Access | VOL. 8

EResNet-SVM: an overfitting-relieved deep learning model for recognition of plant diseases and pests.
Haitao Xiong ... Ziyang Wang
Journal of the science of food and agriculture | VOL. 104
Haitao Xiong, et. al.Haitao Xiong ... Ziyang Wang
27 Apr 2024
Journal of the science of food and agriculture | VOL. 104

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Multimodal Fine-Grained Transformer Model for Pest Recognition

Abstract

Talk to us

Similar Papers

More From: Electronics