Title | Navigation with memory in a partially observable environment |
Publication Type | Journal Article |
Year of Publication | 2006 |
Authors | Montesanto A., Tascini G, Puliti P, Baldassarri P |
Journal | Robotics and Autonomous Systems |
Volume | 54 |
Pagination | 84-94 |
Abstract | The paper presents an architecture that allows the reactive visual navigation via an unsupervised reinforcement learning. This objective is reached using Q-learning and a hierarchical approach to the developed architecture. Using these techniques requires a deviation from the Partially Observable Markov Decision Processes (POMDP) and some innovations: heuristic techniques for generalizing the experience and for treating the partial observability; a technique for the speed adjournment of the Q function; the definition of a special reinforcement policy adequate for learning a complex task without supervision. The result is a satisfactory learning of the navigation assignment in a simulated environment. © 2005 Elsevier B.V. All rights reserved. |
URL | http://www.scopus.com/inward/record.url?eid=2-s2.0-29344460057&partnerID=40&md5=a3f0d1229db6d5c8658cd509ca5acf81 |
DOI | 10.1016/j.robot.2005.09.015 |
Navigation with memory in a partially observable environment
0