Actor-critic Learning Research Articles

This paper addresses a learning-based path following control scheme for a biomimetic underwater vehicle (BUV) driven by undulatory fins. A dynamic line-of-sight (DLOS) guidance system is designed, which uses a virtual ball with a dynamic radius to detect the reference path. This DLOS system guides our BUV in the path following control and extracts essential information for the Markov decision process (MDP) of the control task. A deep reinforcement learning (DRL) algorithm, sample-observed soft actor-critic (SOSAC) is proposed. The can train out control policy with greater cumulative reward and higher success rate by using two tricks: sample observation and sample diversification. Based on the DLOS system, the MDP of the control task, and a multilayer perceptron (MLP) trained by the SOSAC, our control scheme is established. Experiments show that our BUV can successfully achieve path following control in an indoor pool environment by using this control scheme. <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">Note to Practitioners</i> —The motivation of this paper is to design a practical end-to-end path following control scheme for the BUV driven by undulatory fins, and verify this scheme in a real-world environment. Unlike common autonomous underwater vehicles (AUVs) using axial propellers, the BUVs apply biomimetic propellers such as the undulatory fin. Multimodel wave patterns can be implemented by the undulatory fin, which generates nonlinear thrust and lateral force simultaneously. This propulsive feature makes the driving force on different directions of the BUV to be strong coupled, and it is complicated to convert the outputs of a common controller into waveform parameters of the undulatory fins to control the BUV. Therefore, in this paper, we proposed an end-to-end learning-based path following controller, which observes environmental information and directly generates waveform parameters to control our BUV. Experiments suggest that our control scheme is practical and valid.

Read full abstract

Actor-critic Learning Research Articles

Related Topics

Articles published on Actor-critic Learning

Multiagent Soft Actor–Critic Learning for Distributed ESS Enabled Robust Voltage Regulation of Active Distribution Grids

Distribution network voltage control considering virtual power plants cooperative optimization with transactive energy

GRU-integrated constrained soft actor-critic learning enabled fully distributed scheduling strategy for residential virtual power plant

Bayesian Strategy Networks Based Soft Actor-Critic Learning

Joint Computing, Pushing, and Caching Optimization for Mobile-Edge Computing Networks via Soft Actor–Critic Learning

Expert System-Based Multiagent Deep Deterministic Policy Gradient for Swarm Robot Decision Making.

Cooperative Finitely Excited Learning for Dynamical Games.

Reinforcement Learning-Based Adaptive Optimal Control for Nonlinear Systems With Asymmetric Hysteresis.

Attention-Enhanced Actor–Critic Learning for Household Nonintrusive Load Monitoring

Sample-Observed Soft Actor-Critic Learning for Path Following of a Biomimetic Underwater Vehicle

Kernel-based Actor-Critic Learning Framework for Autonomous Brain Control on Trajectory

Reinforcement actor-critic learning as a rehearsal in MicroRTS

Multi-Timescale Actor-Critic Learning for Computing Resource Management With Semi-Markov Renewal Process Mobility

Actor–critic learning based PID control for robotic manipulators

Adaptive fault-tolerant control for non-minimum phase hypersonic vehicles based on adaptive dynamic programming

Hierarchical Multiagent Formation Control Scheme via Actor-Critic Learning.

Reinforcement learning-based saturated adaptive robust output-feedback funnel control of surface vessels in different weather conditions

Koopman-operator-based learning control of air-breathing hypersonic vehicles with nonminimum phase properties

Receding Horizon Actor–Critic Learning Control for Nonlinear Time-Delay Systems With Unknown Dynamics

Federated Multiagent Actor–Critic Learning Task Offloading in Intelligent Logistics

Lead the way for us

Editage

Paperpal

R Discovery

Mind the Graph

Actor-critic Learning Research Articles

Related Topics

Articles published on Actor-critic Learning

Multiagent Soft Actor–Critic Learning for Distributed ESS Enabled Robust Voltage Regulation of Active Distribution Grids

Distribution network voltage control considering virtual power plants cooperative optimization with transactive energy

GRU-integrated constrained soft actor-critic learning enabled fully distributed scheduling strategy for residential virtual power plant

Bayesian Strategy Networks Based Soft Actor-Critic Learning

Joint Computing, Pushing, and Caching Optimization for Mobile-Edge Computing Networks via Soft Actor–Critic Learning

Expert System-Based Multiagent Deep Deterministic Policy Gradient for Swarm Robot Decision Making.

Cooperative Finitely Excited Learning for Dynamical Games.

Reinforcement Learning-Based Adaptive Optimal Control for Nonlinear Systems With Asymmetric Hysteresis.

Attention-Enhanced Actor–Critic Learning for Household Nonintrusive Load Monitoring

Sample-Observed Soft Actor-Critic Learning for Path Following of a Biomimetic Underwater Vehicle

Kernel-based Actor-Critic Learning Framework for Autonomous Brain Control on Trajectory

Reinforcement actor-critic learning as a rehearsal in MicroRTS

Multi-Timescale Actor-Critic Learning for Computing Resource Management With Semi-Markov Renewal Process Mobility

Actor–critic learning based PID control for robotic manipulators

Adaptive fault-tolerant control for non-minimum phase hypersonic vehicles based on adaptive dynamic programming

Hierarchical Multiagent Formation Control Scheme via Actor-Critic Learning.

Reinforcement learning-based saturated adaptive robust output-feedback funnel control of surface vessels in different weather conditions

Koopman-operator-based learning control of air-breathing hypersonic vehicles with nonminimum phase properties

Receding Horizon Actor–Critic Learning Control for Nonlinear Time-Delay Systems With Unknown Dynamics

Federated Multiagent Actor–Critic Learning Task Offloading in Intelligent Logistics