GLDAP: Global Dynamic Action Persistence Adaptation for Deep Reinforcement Learning

Autor:	Junbo Tong, Daming Shi, Yi Liu, Wenhui Fan
Rok vydání:	2023
Předmět:	Control and Systems Engineering Computer Science (miscellaneous) Software
Zdroj:	ACM Transactions on Autonomous and Adaptive Systems. 18:1-18
ISSN:	1556-4703 1556-4665
DOI:	10.1145/3590154
Popis:	In the implementation of deep reinforcement learning (DRL), action persistence strategies are often adopted so agents maintain their actions for a fixed or variable number of steps. The choice of the persistent duration for agent actions usually has notable effects on the performance of reinforcement learning algorithms. Aiming at the research gap of global dynamic optimal action persistence and its application in multi-agent systems, we propose a novel framework: global dynamic action persistence (GLDAP), which achieves global action persistence adaptation for deep reinforcement learning. We introduce a closed-loop method that is used to learn the estimated value and the corresponding policy of each candidate action persistence. Our experiment shows that GLDAP achieves an average of 2.5%~90.7% performance improvement and 3~20 times higher sampling efficiency over several baselines across various single-agent and multi-agent domains. We also validate the ability of GLDAP to determine the optimal action persistence through multiple experiments.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_________::5a99374555fe8c50f224ae9c3b2b9721 https://doi.org/10.1145/3590154 Zobrazit plný text záznamu