arrow
Return

A stable method for task priority adaptation in quadratic programming via reinforcement learning

delete2025-02-01
delete0
delete
OA
AI
A
Andrea Testa *
M
Marco Laghi
E
Edoardo Del Bianco
G
Gennaro Raiola
E
Enrico Mingo Hoffman
A
Arash Ajoudani
DOI:10.1016/j.rcim.2024.102857delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
In emerging manufacturing facilities, robots must enhance their flexibility. They are expected to perform complex jobs, showing different behaviors on the need, all within unstructured environments, and without requiring reprogramming or setup adjustments. To address this challenge, we introduce the A3CQP, a non- strict hierarchical Quadratic Programming (QP) controller. It seamlessly combines both motion and interaction functionalities, with priorities dynamically and autonomously adapted through a Reinforcement Learning-based adaptation module. This module utilizes the Asynchronous Advantage Actor-Critic algorithm (A3C) to ensure rapid convergence and stable training within continuous action and observation spaces. The experimental validation, involving a collaborative peg-in-hole assembly and the polishing of a wooden plate, demonstrates the effectiveness of the proposed solution in terms of its automatic adaptability, responsiveness, flexibility, and safety.
Keywords:
Optimization and optimal control
Reinforcement learning
Machine learning for robot control
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

R
Robotics and Computer-Integrated Manufacturing
IF:
11.4
Papers:
3.3K
Citations:
1.3W

Organization

U
University of Trento
Scholars:
8.8K
Papers: 9.0K
Citations: 1.2W
I
istituto italiano di tecnologia - iit
Scholars:
9.1K
Papers: 6.9K
Citations: 8