Dealing with uncertainty: Balancing exploration and exploitation in deep recurrent reinforcement learning

Valentina Zangirolami,Matteo Borrotti

doi:10.1016/j.knosys.2024.111663

Abstract

Incomplete knowledge of the environment leads an agent to make decisions under uncertainty. One of the major dilemmas in Reinforcement Learning (RL) where an autonomous agent has to balance two contrasting needs in making its decisions is: exploiting the current knowledge of the environment to maximize the cumulative reward as well as exploring actions that allow improving the knowledge of the environment, hopefully leading to higher reward values (exploration–exploitation trade-off). Concurrently, another relevant issue regards the full observability of the states, which may not be assumed in all applications. For instance, when 2D images are considered as input in an RL approach used for finding the best actions within a 3D simulation environment. In this work, we address these issues by deploying and testing several techniques to balance exploration and exploitation trade-off on partially observable systems for predicting steering wheels in autonomous driving scenarios. More precisely, the final aim is to investigate the effects of using both adaptive and deterministic exploration strategies coupled with a Deep Recurrent Q-Network. Additionally, we adapted and evaluated the impact of a modified quadratic loss function to improve the learning phase of the underlying Convolutional Recurrent Neural Network. We show that adaptive methods better approximate the trade-off between exploration and exploitation and, in general, Softmax and Max-Boltzmann strategies outperform ϵ-greedy techniques.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Dealing with uncertainty: Balancing exploration and exploitation in deep recurrent reinforcement learning

Abstract

Talk to us

Similar Papers

More From: Knowledge-Based Systems

Lead the way for us

Similar Papers

Deep Recurrent Q-Network with Truncated History
Hyunwoo Oh ... Tomoyuki Kaneko
-
Hyunwoo Oh, et. al.Hyunwoo Oh ... Tomoyuki Kaneko
01 Nov 2018
01 Nov 2018

A Validation Approach for Deep Reinforcement Learning of a Robotic Arm in a 3D Simulated Environment
Monica Gruosso ... Nicola Capece
-
Monica Gruosso, et. al.Monica Gruosso ... Nicola Capece
21 Jan 2021
21 Jan 2021

Model-Free Deep Recurrent Q-Network Reinforcement Learning for Quantum Circuit Architectures Design
Tomah Sogabe ... Katsuyoshi Sakamoto
Quantum Reports | VOL. 4
Tomah Sogabe, et. al.Tomah Sogabe ... Katsuyoshi Sakamoto
21 Sep 2022
Quantum Reports | VOL. 4

Deep imitation learning for 3D navigation tasks
Ahmed Hussein ... Eyad Elyan
Neural Computing and Applications | VOL. 29
Ahmed Hussein, et. al.Ahmed Hussein ... Eyad Elyan
04 Dec 2017
Neural Computing and Applications | VOL. 29

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Dealing with uncertainty: Balancing exploration and exploitation in deep recurrent reinforcement learning

Abstract

Talk to us

Similar Papers

More From: Knowledge-Based Systems