返回
Developing navigation behavior through self-organizing distinctive-state abstraction
DOI:10.1080/09540090600768609.png)
摘要
En 中文
A major challenge in reinforcement learning research is to extend the methods that have worked well on discrete, short-range, low-dimensional problems to continuous, high-diameter, high-dimensional problems, such as robot navigation using high-resolution sensors. Self-organizing distinctive-state abstraction (SODA) is a new, generic method by which a robot in a continuous world can better learn to navigate, by learning a set of high-level features and building temporally extended actions to carry it between distinctive states based on those features. A SODA agent first uses a self-organizing feature map to develop a set of high-level perceptual features while exploring the environment with primitive, local actions. The agent then builds a set of high-level actions composed of generic trajectory-following and hill-climbing control laws that carry it between the states at local maxima of feature activations. In an experiment on a simulated robot navigation task, the SODA agent learns to perform a task requiring 300 small-scale, local actions using as few as nine new, temporally extended actions, significantly improving learning time over navigating with the local actions.
Keyword:
developmental robotics
self-organization
reinforcement learning
robot navigation
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.4
论文数:
849
被引数:
1.5K
机构
暂无机构信息
引用论文
Towards autonomous sensor and actuator model induction on a mobile robot移动机器人上的自主传感器和执行器模型感应
CONNECTION SCIENCE
IF3.4
Between MDPs and semi-MDPs: A framework for temporal abstraction in reinforcement learning在mdp和半mdp之间: 强化学习中的时间抽象框架

