RTDOD: A large-scale RGB-thermal domain-incremental object detection dataset for UAVs

Hangtao Feng,Lu Zhang,Siqi Zhang,Dong Wang,Xu Yang,Zhiyong Liu

doi:10.1016/j.imavis.2023.104856

Abstract

Recently, visual understanding using unmanned aerial vehicles (UAVs) has gained significant attention due to its wide range of applications, including delivery, security investigation and surveillance. However, most existing UAV-based datasets only capture color images under ideal illumination and weather conditions, typically sunny days. This limitation fails to account for the complexity of real-world scenarios, such as cloudy or foggy weather, and nighttime conditions. Deep learning methods trained on color images with good lighting and weather conditions struggle to adapt to the complex visual scenes in these scenarios. Moreover, color images may not provide sufficient visual information under the complex visual scenes. To bridge this gap and meet the demands of real-world applications, we propose a large-scale RGB-Thermal Domain-incremental Object Detection (RTDOD) dataset in this paper. Our dataset includes RGB and thermal videos synchronously captured using calibrated color thermal cameras mounted on UAVs. It covers various weather conditions, from sunny to foggy to rainy, and spans from day to night. We sample and obtain approximately 16,200 pairs of images, and manually label dense annotations, including object bounding boxes and object categories. With the proposed dataset, we introduce a challenging domain-incremental object detection task. We also present a baseline approach that uses task-related gates to filter features for knowledge distillation to reduce forgetting. Experimental results on the RTDOD dataset demonstrate the effectiveness of our proposed method in domain-incremental object detection. To facilitate future research and development in domain-incremental object detection tasks on aerial images, the RTDOD dataset and our baseline model are made available at https://github.com/fenght96/RTDOD.ARTICLE INFO.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

RTDOD: A large-scale RGB-thermal domain-incremental object detection dataset for UAVs

Abstract

Talk to us

Similar Papers

More From: Image and Vision Computing

Lead the way for us

Journal: Image and Vision Computing	Publication Date: Oct 31, 2023
Citations: 3

Similar Papers

ANALYSIS OF THE FEATURES OF THE USE UNMANNED AERIAL VEHICLES BY THE ARMED FORCES OF THE RUSSIAN FEDERATION DURING A FULL-SCALE ARMED INVASION
M.M Synyshyn ... V.S Demchyshyn
Collection of scientific works of the Military Institute of Kyiv National Taras Shevchenko University | VOL. -
M.M Synyshyn, et. al.M.M Synyshyn ... V.S Demchyshyn
01 Jan 2023
Collection of scientific works of the Military Institute of Kyiv National Taras Shevchenko University | VOL. -

ОБҐРУНТУВАННЯ НЕОБХІДНОСТІ СТВОРЕННЯ МОБІЛЬНОГО КОМПЛЕКСУ УПРАВЛІННЯ БЕЗПІЛОТНИМИ ЛІТАЛЬНИМИ АПАРАТАМИ НА БАЗІ ШАСІ ВАНТАЖНОГО АВТОМОБІЛЯ
K Sporyshev ... K Gynbin
The collection of scientific works of the National Academy of the National Guard of Ukraine | VOL. 2
K Sporyshev, et. al.K Sporyshev ... K Gynbin
01 Jan 2023
The collection of scientific works of the National Academy of the National Guard of Ukraine | VOL. 2

Use of Unmanned Aerial Assault Vehicles (UAAV) as an Asymmetric Factor
Alper Alpaslan Eker ... Eray Sallar
Journal of Military and Information Science | VOL. 2
Alper Alpaslan Eker, et. al.Alper Alpaslan Eker ... Eray Sallar
29 Nov 2014
Journal of Military and Information Science | VOL. 2

Application of UAV Video Communication Systems During Investigation of Emergency Situations
Ihor Maladyka ... Oleksandr Dzhulay
-
Ihor Maladyka, et. al.Ihor Maladyka ... Oleksandr Dzhulay
29 Jul 2022
29 Jul 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

RTDOD: A large-scale RGB-thermal domain-incremental object detection dataset for UAVs

Abstract

Talk to us

Similar Papers

More From: Image and Vision Computing