arrow
Return

Deterministic delay-aware reinforcement learning

delete2025-12-02
delete0
PRE
AI
S
Sathira Dilshan Bataduwaarachchi *
Z
Zoran Najdovski
H
Hieu Trinh
C
Chee Peng Lim
V
Van Thanh Huynh
DOI:10.1016/j.robot.2025.105271delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Reinforcement Learning (RL) has effectively paved the way in achieving robotic control during the past decade. As a result, the avenue of integrating RL-powered robotic control and teleoperation has caught the attention of researchers. Every RL framework involves the basis of suitable observation and action communication between the environment and the agent, and the involvement of teleoperation can introduce random time delays within the said communication process. Achieving robotic control under such constraints remains an untapped area in the domain of reinforcement learning. We take the initiative to achieve the goal of robotic control while handling delays in the RL setting based on a fitting Markov Decision Process (MDP) structure. Our algorithm will learn a deterministic policy and can tackle control environments, especially robotic manipulation environments, using observations with proprioceptive information. We methodically present the theoretical adjustments based on an existing dominant off-policy algorithm to express the algorithm’s competency with proof of convergence. We perform experimentations with DeepMind Control Suite, illustrating significant results showing the algorithm’s capabilities in learning complex environments powered by delay-aware RL.

Journal

Robotics and Autonomous Systems cover
Robotics and Autonomous Systems
IF:
5.2
Papers:
633
Citations:
1.0W

Organization

S
Swinburne University of Technology
Scholars:
9.3K
Papers: 1.2W
Citations: 2.0W
D
Deakin University
Scholars:
2.0W
Papers: 2.1W
Citations: 2.8W
researcher View more organizations