LSTM-DPPO based deep reinforcement learning controller for path following optimization of unmanned surface vehicle

JOURNAL OF SYSTEMS ENGINEERING AND ELECTRONICS(2023)

引用 0|浏览1
暂无评分
摘要
To solve the path following control problem for unmanned surface vehicles (USVs), a control method based on deep reinforcement learning (DRL) with long short-term memory (LSTM) networks is proposed. A distributed proximal policy optimization (DPPO) algorithm, which is a modified actorcritic-based type of reinforcement learning algorithm, is adapted to improve the controller performance in repeated trials. The LSTM network structure is introduced to solve the strong temporal correlation USV control problem. In addition, a specially designed path dataset, including straight and curved paths, is established to simulate various sailing scenarios so that the reinforcement learning controller can obtain as much handling experience as possible. Extensive numerical simulation results demonstrate that the proposed method has better control performance under missions involving complex maneuvers than trained with limited scenarios and can potentially be applied in practice.
更多
查看译文
关键词
unmanned surface vehicle (USV),deep reinforcement learning (DRL),path following,path dataset,proximal policy optimization,long short-term memory (LSTM)
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要