返回
Automatic collective motion tuning using actor-critic deep reinforcement learning
DOI:10.1016/j.swevo.2022.101085.png)
摘要
En 中文
Collective behaviours such as swarm formation of autonomous agents offer the advantages of efficient movement, redundancy, and potential for human guidance of a single swarm organism. However, tuning the behaviour of a group of agents so that they swarm, is difficult. Behaviour-bootstrapping algorithms permit agents to self-tune behaviour adapted for their physical form and associated movement constraints. This paper proposes a reinforce-ment learning framework to tune collective motion behaviours from random behaviours. The learning process is guided by a novel reward function capable of autonomously detecting generic collective motion behaviours from sensor data about the relative velocity and position of neighbouring agents. Our reward function is de-signed using a meta-learner trained on a human-labelled collective motion dataset. We demonstrate that our reinforcement learner can tune the behaviour of randomly moving groups so that structured collective motion emerges. We compare our framework to an existing developmental evolutionary framework for this purpose. Our results demonstrate that the proposed learning framework can generate behaviours with different collective motion characteristics more quickly than existing approaches. In addition, the trained reinforcement learner can tune the behaviour of robots with movement characteristics that it has not been trained on.
Keyword:
Collective behaviour
Swarm intelligence
Reinforcement learning
Collective motion tuning
Rule engine
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
8.5
论文数:
2.2K
被引数:
1.0W
机构
暂无机构信息
引用论文
Time-Varying Formation Control for Unmanned Aerial Vehicles: Theories and Applications无人机时变编队控制: 理论与应用

