References & Citations
Computer Science > Robotics
Title: Guided Dyna-Q for Mobile Robot Exploration and Navigation
(Submitted on 23 Apr 2020 (this version), latest version 16 Mar 2021 (v2))
Abstract: Model-based reinforcement learning (RL) enables an agent to learn world models from trial-and-error experiences toward achieving long-term goals. Automated planning, on the other hand, can be used for accomplishing tasks through reasoning with declarative action knowledge. Despite their shared goal of completing complex tasks, the development of RL and automated planning has mainly been isolated due to their different modalities of computation. Focusing on improving model-based RL agent's exploration strategy and sample efficiency, we develop Guided Dyna-Q (GDQ) to enable RL agents to reason with action knowledge to avoid exploring less-relevant states toward more efficient task accomplishment. GDQ has been evaluated in simulation and using a mobile robot conducting navigation tasks in an office environment. Results show that GDQ reduces the effort in exploration while improving the quality of learned policies.
Submission history
From: Yohei Hayamizu [view email][v1] Thu, 23 Apr 2020 21:03:30 GMT (1292kb,D)
[v2] Tue, 16 Mar 2021 14:47:46 GMT (4148kb,D)
Link back to: arXiv, form interface, contact.