Reinforcement Learning for Robot Navigation with Adaptive Forward Simulation Time (AFST) in a Semi-Markov Model

Autor:	Chen, Yu'an, Ye, Ruosong, Tao, Ziyang, Liu, Hongjian, Chen, Guangda, Peng, Jie, Ma, Jun, Zhang, Yu, Ji, Jianmin, Zhang, Yanyong
Rok vydání:	2021
Předmět:	Computer Science - Robotics Computer Science - Artificial Intelligence
Druh dokumentu:	Working Paper
Popis:	Deep reinforcement learning (DRL) algorithms have proven effective in robot navigation, especially in unknown environments, by directly mapping perception inputs into robot control commands. However, most existing methods ignore the local minimum problem in navigation and thereby cannot handle complex unknown environments. In this paper, we propose the first DRL-based navigation method modeled by a semi-Markov decision process (SMDP) with continuous action space, named Adaptive Forward Simulation Time (AFST), to overcome this problem. Specifically, we reduce the dimensions of the action space and improve the distributed proximal policy optimization (DPPO) algorithm for the specified SMDP problem by modifying its GAE to better estimate the policy gradient in SMDPs. Experiments in various unknown environments demonstrate the effectiveness of AFST.
Databáze:	arXiv
Externí odkaz:	http://arxiv.org/abs/2108.06161 Zobrazit plný text záznamu View this record from Arxiv