返回
Efficient Insertion Control for Precision Assembly Based on Demonstration Learning and Reinforcement Learning
DOI:10.1109/TII.2020.3020065.png)
摘要
En 中文
Multiple peg-in-hole insertion control is one of the challenging tasks in precision assembly for its complex contact dynamics. In this article, an insertion policy learning method is proposed for multiple peg-in-hole precision assembly. The insertion policy learning process is separated into two phases: initial policy learning and residual policy learning. In initial policy learning, a state-to-action policy mapping model based on the Gaussian mixture model (GMM) is established. And Gaussian mixture regression (GMR) is used to generalize the policy reuse. In residual policy learning, a reinforcement learning method named normalized advantage function (NAF) is employed to refine the insertion policy via agent's exploration in the insertion environment. Moreover, an adaptive action exploration (AAE) strategy is designed to improve the performance of exploration, and the prioritized experience replay strategy is introduced to make the residual policy learning from historical experience more efficient. Besides, the hierarchical reward function is designed considering the contact dynamics as well as the efficiency and safety of precision insertion. Finally, comprehensive experiments are conducted to validate the effectiveness of the proposed insertion policy learning method.
Keyword:
Learning (artificial intelligence)
Task analysis
Learning systems
Gaussian distribution
Informatics
Automation
Data models
Demonstration learning
insertion policy learning
multiple peg-in-hole insertion
precision assembly
reinforcement learning
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
9.9
论文数:
8.6K
被引数:
6.0W
机构
引用论文
Multiple Targets for Suppression of RNA Interference by Tomato Aspermy Virus Protein 2B
Biochemistry
IF0
Sentiment Analysis of Text Reviews Using Lexicon-Enhanced Bert Embedding (LeBERT) Model with Convolutional Neural Network使用词库增强的Bert嵌入 (LeBERT) 模型和卷积神经网络对文本评论进行情感分析

