Yayınlanmış 1 Ocak 2023 | Sürüm v1
Dergi makalesi Açık

Learning what to memorize: Using intrinsic motivation to form useful memory in partially observable reinforcement learning

Oluşturanlar

  • 1. Izmir Univ Econ, Dept Comp Engn, TR-35330 Izmir, Turkiye

Açıklama

Reinforcement Learning faces an important challenge in partially observable environments with long-term dependencies. In order to learn in an ambiguous environment, an agent has to keep previous perceptions in a memory. Earlier memory-based approaches use a fixed method to determine what to keep in the memory, which limits them to certain problems. In this study, we follow the idea of giving the control of the memory to the agent by allowing it to take memory-changing actions. Thus, the agent becomes more adaptive to the dynamics of an environment. Further, we formalize an intrinsic motivation to support this learning mechanism, which guides the agent to memorize distinctive events and enable it to disambiguate its state in the environment. Our overall approach is tested and analyzed on several partial observable tasks with long-term dependencies. The experiments show a clear improvement in terms of learning performance compared to other memory based methods.

Dosyalar

bib-6a63c3c8-d2a2-4de4-92b3-c56afba8e1fa.txt

Dosyalar (183 Bytes)

Ad Boyut Hepisini indir
md5:c22721a86f5026d49afe7acc78f39626
183 Bytes Ön İzleme İndir