BSAC: Bayesian Strategy Network Based Soft Actor-Critic in Deep Reinforcement Learning

08/11/2022
by   Qin Yang, et al.
6

Adopting reasonable strategies is challenging but crucial for an intelligent agent with limited resources working in hazardous, unstructured, and dynamic environments to improve the system utility, decrease the overall cost, and increase mission success probability. Deep Reinforcement Learning (DRL) helps organize agents' behaviors and actions based on their state and represents complex strategies (composition of actions). This paper proposes a novel hierarchical strategy decomposition approach based on Bayesian chaining to separate an intricate policy into several simple sub-policies and organize their relationships as Bayesian strategy networks (BSN). We integrate this approach into the state-of-the-art DRL method, soft actor-critic (SAC), and build the corresponding Bayesian soft actor-critic (BSAC) model by organizing several sub-policies as a joint policy. We compare the proposed BSAC method with the SAC and other state-of-the-art approaches such as TD3, DDPG, and PPO on the standard continuous control benchmarks – Hopper-v2, Walker2d-v2, and Humanoid-v2 – in MuJoCo with the OpenAI Gym environment. The results demonstrate that the promising potential of the BSAC method significantly improves training efficiency. The open sourced codes for BSAC can be accessed at https://github.com/herolab-uga/bsac.

READ FULL TEXT

page 2

page 5

page 10

page 11

research
03/07/2023

A Strategy-Oriented Bayesian Soft Actor-Critic Model

Adopting reasonable strategies is challenging but crucial for an intelli...
research
01/30/2021

Stay Alive with Many Options: A Reinforcement Learning Approach for Autonomous Navigation

Hierarchical reinforcement learning approaches learn policies based on h...
research
07/27/2022

SAC-AP: Soft Actor Critic based Deep Reinforcement Learning for Alert Prioritization

Intrusion detection systems (IDS) generate a large number of false alert...
research
04/25/2020

A State Aggregation Approach for Solving Knapsack Problem with Deep Reinforcement Learning

This paper proposes a Deep Reinforcement Learning (DRL) approach for sol...
research
03/24/2023

Multi-Task Reinforcement Learning in Continuous Control with Successor Feature-Based Concurrent Composition

Deep reinforcement learning (DRL) frameworks are increasingly used to so...
research
01/10/2023

Deep Reinforcement Learning for Autonomous Ground Vehicle Exploration Without A-Priori Maps

Autonomous Ground Vehicles (AGVs) are essential tools for a wide range o...
research
10/14/2022

Just Round: Quantized Observation Spaces Enable Memory Efficient Learning of Dynamic Locomotion

Deep reinforcement learning (DRL) is one of the most powerful tools for ...

Please sign up or login with your details

Forgot password? Click here to reset