Augmented Memory Replay in Reinforcement Learning With Continuous Control

Mirza Ramicic,Andrea Bonarini

doi:10.1109/tcds.2021.3050723

Abstract

Online reinforcement learning agents are currently able to process an increasing amount of data by converting it into a higher order value functions. This expansion of the information collected from the environment increases the agent's state space enabling it to scale up to a more complex problems but also increases the risk of forgetting by learning on redundant or conflicting data. To improve the approximation of a large amount of data, a random mini-batch of the past experiences that are stored in the replay memory buffer is often replayed at each learning step. The proposed work takes inspiration from a biological mechanism which act as a protective layer of human brain higher cognitive functions: active memory consolidation mitigates the effect of forgetting of previous memories by dynamically processing the new ones. The similar dynamics are implemented by a proposed augmented memory replay AMR capable of optimizing the replay of the experiences from the agent's memory structure by altering or augmenting their relevance. Experimental results show that an evolved AMR augmentation function capable of increasing the significance of the specific memories is able to further increase the stability and convergence speed of the learning algorithms dealing with the complexity of continuous action domains.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Augmented Memory Replay in Reinforcement Learning With Continuous Control

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Cognitive and Developmental Systems

Lead the way for us

Journal: IEEE Transactions on Cognitive and Developmental Systems	Publication Date: Jan 13, 2021
Citations: 4

Similar Papers

Continuous Control in Car Simulator with Deep Reinforcement Learning
Fan Yang ... Ping Wang
-
Fan Yang, et. al.Fan Yang ... Ping Wang
08 Dec 2018
08 Dec 2018

Sample Efficient Reinforcement Learning Using Graph-Based Memory Reconstruction
Yongxin Kang ... Yifan Zang
IEEE Transactions on Artificial Intelligence | VOL. 5
Yongxin Kang, et. al.Yongxin Kang ... Yifan Zang
01 Feb 2024
IEEE Transactions on Artificial Intelligence | VOL. 5

A preliminary investigation of 3D preconditioned conjugate gradient reconstruction for cone-beam CT
Lin Fu ... Robert M Nishikawa
-
Lin Fu, et. al.Lin Fu ... Robert M Nishikawa
23 Feb 2012
23 Feb 2012

Robot Path Planning of Improved Adaptive Ant Colony System Algorithm Based on Dijkstra
Chonglin Gu ... Ansong Feng
Journal of Robotics | VOL. 2022
Chonglin Gu, et. al.Chonglin Gu ... Ansong Feng
15 Dec 2022
Journal of Robotics | VOL. 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Augmented Memory Replay in Reinforcement Learning With Continuous Control

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Cognitive and Developmental Systems