Zobrazeno 1 - 1
of 1
pro vyhledávání: '"Furuyama, Ryoma"'
Imitation learning is often used in addition to reinforcement learning in environments where reward design is difficult or where the reward is sparse, but it is difficult to be able to imitate well in unknown states from a small amount of expert data
Externí odkaz:
http://arxiv.org/abs/2401.16772