Augmented Replay Memory in Reinforcement Learning With Continuous Control

by   Mirza Ramicic, et al.

Online reinforcement learning agents are currently able to process an increasing amount of data by converting it into a higher order value functions. This expansion of the information collected from the environment increases the agent's state space enabling it to scale up to a more complex problems but also increases the risk of forgetting by learning on redundant or conflicting data. To improve the approximation of a large amount of data, a random mini-batch of the past experiences that are stored in the replay memory buffer is often replayed at each learning step. The proposed work takes inspiration from a biological mechanism which act as a protective layer of human brain higher cognitive functions: active memory consolidation mitigates the effect of forgetting of previous memories by dynamically processing the new ones. The similar dynamics are implemented by a proposed augmented memory replay AMR capable of optimizing the replay of the experiences from the agent's memory structure by altering or augmenting their relevance. Experimental results show that an evolved AMR augmentation function capable of increasing the significance of the specific memories is able to further increase the stability and convergence speed of the learning algorithms dealing with the complexity of continuous action domains.


page 1

page 2

page 3

page 5

page 6

page 7


Memory-efficient Reinforcement Learning with Knowledge Consolidation

Artificial neural networks are promising as general function approximato...

Active Inference in Hebbian Learning Networks

This work studies how brain-inspired neural ensembles equipped with loca...

A Dual Memory Structure for Efficient Use of Replay Memory in Deep Reinforcement Learning

In this paper, we propose a dual memory structure for reinforcement lear...

Reinforcement Learning for Robust Missile Autopilot Design

Designing missiles' autopilot controllers has been a complex task, given...

The Tensor Brain: A Unified Theory of Perception, Memory and Semantic Decoding

We present a unified computational theory of perception and memory. In o...

Generative Memory for Lifelong Reinforcement Learning

Our research is focused on understanding and applying biological memory ...

Paused Agent Replay Refresh

Reinforcement learning algorithms have become more complex since the inv...

Please sign up or login with your details

Forgot password? Click here to reset