DRFnet: Dynamic receptive field network for object detection and image recognition.

Minjie Tan,Binbin Liang,Xinyang Yuan,Songchen Han

doi:10.3389/fnbot.2022.1100697

Abstract

Biological experiments discovered that the receptive field of neurons in the primary visual cortex of an animal's visual system is dynamic and capable of being altered by the sensory context. However, in a typical convolution neural network (CNN), a unit's response only comes from a fixed receptive field, which is generally determined by the preset kernel size in each layer. In this work, we simulate the dynamic receptive field mechanism in the biological visual system (BVS) for application in object detection and image recognition. We proposed a Dynamic Receptive Field module (DRF), which can realize the global information-guided responses under the premise of a slight increase in parameters and computational cost. Specifically, we design a transformer-style DRF module, which defines the correlation coefficient between two feature points by their relative distance. For an input feature map, we first divide the relative distance corresponding to different receptive field regions between the target feature point and its surrounding feature points into N different discrete levels. Then, a vector containing N different weights is automatically learned from the dataset and assigned to each feature point, according to the calculated discrete level that this feature point belongs. In this way, we achieve a correlation matrix primarily measuring the relationship between the target feature point and its surrounding feature points. The DRF-processed responses of each feature point are computed by multiplying its corresponding correlation matrix with the input feature map, which computationally equals to accomplish a weighted sum of all feature points exploiting the global and long-range information as the weight. Finally, by superimposing the local responses calculated by a traditional convolution layer with DRF responses, our proposed approach can integrate the rich context information among neighbors and the long-range dependencies of background into the feature maps. With the proposed DRF module, we achieved significant performance improvement on four benchmark datasets for both tasks of object detection and image recognition. Furthermore, we also proposed a new matching strategy that can improve the detection results of small targets compared with the traditional IOU-max matching strategy.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

DRFnet: Dynamic receptive field network for object detection and image recognition.

Abstract

Talk to us

Similar Papers

More From: Frontiers in neurorobotics

Lead the way for us

Journal: Frontiers in neurorobotics	Publication Date: Jan 10, 2023
License type: CC BY 4.0

Similar Papers

Multifocal Attention Filters Targets from Distracters within and beyond Primate MT Neurons' Receptive Field Boundaries
Robert Niebergall ... Julio C Martinez-Trujillo
Neuron | VOL. 72
Robert Niebergall, et. al.Robert Niebergall ... Julio C Martinez-Trujillo
01 Dec 2011
Neuron | VOL. 72

The role of neural mechanisms of attention in solving the binding problem.
John H Reynolds ... Robert Desimone
Neuron | VOL. 24
John H Reynolds, et. al.John H Reynolds ... Robert Desimone
01 Sep 1999
Neuron | VOL. 24

Receptive Fields of Second‐Order Taste Neurons in Sheep
Mark B Vogt ... Charlotte M Mistretta
Annals of the New York Academy of Sciences | VOL. 510
Mark B Vogt, et. al.Mark B Vogt ... Charlotte M Mistretta
01 Nov 1987
Annals of the New York Academy of Sciences | VOL. 510

Population Receptive Field Dynamics in Human Visual Cortex
Koen V Haak ... Frans W Cornelissen
PLoS ONE | VOL. 7
Koen V Haak, et. al.Koen V Haak ... Frans W Cornelissen
23 May 2012
PLoS ONE | VOL. 7

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

DRFnet: Dynamic receptive field network for object detection and image recognition.

Abstract

Talk to us

Similar Papers

More From: Frontiers in neurorobotics