Visual Tracking via Hierarchical Deep Reinforcement Learning

Dawei Zhang,Minglu Li,Riheng Jia,Zhonglong Zheng

doi:10.1609/aaai.v35i4.16443

Abstract

Visual tracking has achieved great progress due to numerous different algorithms. However, deep trackers based on classification or Siamese network still have their specific limitations. In this work, we show how to teach machines to track a generic object in videos like humans, who can use a few search steps to perform tracking. By constructing a Markov decision process in Deep Reinforcement Learning (DRL), our agents can learn to determine hierarchical decisions on tracking mode and motion estimation. To be specific, our Hierarchical DRL framework is composed of a Siamese-based observation network which models the motion information of an arbitrary target, a policy network for mode switch and an actor-critic network for box regression. This tracking strategy is more in line with human behavior paradigm, and is effective and efficient to cope with fast motion, background clutter and large deformations. Extensive experiments on the GOT-10k, OTB-100, UAV-123, VOT and LaSOT tracking benchmarks, demonstrate that the proposed tracker achieves state-of-the-art performance while running in real-time.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Visual Tracking via Hierarchical Deep Reinforcement Learning

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Journal: Proceedings of the AAAI Conference on Artificial Intelligence	Publication Date: May 18, 2021
Citations: 15

Similar Papers

TrackDQN: Visual Tracking via Deep Reinforcement Learning
Pei Yang ... Jiyue Huang
-
Pei Yang, et. al.Pei Yang ... Jiyue Huang
01 Oct 2019
01 Oct 2019

Siamese network with a depthwise over-parameterized convolutional layer for visual tracking.
Yuanyun Wang ... Jun Wang
PloS one | VOL. 17
Yuanyun Wang, et. al.Yuanyun Wang ... Jun Wang
31 Aug 2022
PloS one | VOL. 17

Discovering Primary Objects in Videos by Saliency Fusion and Iterative Appearance Estimation
Jiong Yang ... Xiaohui Shen
IEEE Transactions on Circuits and Systems for Video Technology | VOL. 26
Jiong Yang, et. al.Jiong Yang ... Xiaohui Shen
01 Jun 2016
IEEE Transactions on Circuits and Systems for Video Technology | VOL. 26

A Cooperative Tracker by Fusing Correlation Filter and Siamese Network
Bin Zhou ... Xin Liu
-
Bin Zhou, et. al.Bin Zhou ... Xin Liu
01 Jan 2020
01 Jan 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Visual Tracking via Hierarchical Deep Reinforcement Learning

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence