Multiscale Deep Learning for Detection and Recognition: A Comprehensive Survey.

Licheng Jiao,Mengjiao Wang,Xu Liu,Lingling Li,Fang Liu,Zhixi Feng,Shuyuan Yang,Biao Hou

doi:10.1109/tnnls.2024.3389454

Abstract

Recently, the multiscale problem in computer vision has gradually attracted people's attention. This article focuses on multiscale representation for object detection and recognition, comprehensively introduces the development of multiscale deep learning, and constructs an easy-to-understand, but powerful knowledge structure. First, we give the definition of scale, explain the multiscale mechanism of human vision, and then lead to the multiscale problem discussed in computer vision. Second, advanced multiscale representation methods are introduced, including pyramid representation, scale-space representation, and multiscale geometric representation. Third, the theory of multiscale deep learning is presented, which mainly discusses the multiscale modeling in convolutional neural networks (CNNs) and Vision Transformers (ViTs). Fourth, we compare the performance of multiple multiscale methods on different tasks, illustrating the effectiveness of different multiscale structural designs. Finally, based on the in-depth understanding of the existing methods, we point out several open issues and future directions for multiscale deep learning.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Multiscale Deep Learning for Detection and Recognition: A Comprehensive Survey.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems

Lead the way for us

Journal: IEEE Transactions on Neural Networks and Learning Systems	Publication Date: Jan 1, 2024
Citations: 1

Similar Papers

Road traffic sign recognition algorithm based on computer vision
Huiming Dai ... Xin Zhang
International Journal of Computational Vision and Robotics | VOL. 8
Huiming Dai, et. al.Huiming Dai ... Xin Zhang
01 Jan 2018
International Journal of Computational Vision and Robotics | VOL. 8

Real Time Human Emotion Recognition Using Artificial Neural Networks
...
-
, et. al. ...
30 Apr 2018
30 Apr 2018

Large-Scale Image Segmentation with Convolutional Networks
...
-
, et. al. ...
01 Jan 2017
01 Jan 2017

To detect and Recognize Object from Videos for Computer Vision by Parallel Approach using Deep Learning
G Nalinipriya ... Balamurugan Baluswarny
-
G Nalinipriya, et. al.G Nalinipriya ... Balamurugan Baluswarny
01 Jun 2018
01 Jun 2018

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Multiscale Deep Learning for Detection and Recognition: A Comprehensive Survey.

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Neural Networks and Learning Systems