A memory and attention-based reinforcement learning for musculoskeletal robots with prior knowledge of muscle synergies-Reference-Cited by-同舟云学术

A memory and attention-based reinforcement learning for musculoskeletal robots with prior knowledge of muscle synergies

Published:2024-04-15 Issue:2 Volume:44 Page:316-333
ISSN:2754-6969
Container-title:Robotic Intelligence and Automation
language:en
Short-container-title:RIA

Author:

Wang Xiaona,Chen Jiahao,Qiao Hong

Abstract

Purpose Limited by the types of sensors, the state information available for musculoskeletal robots with highly redundant, nonlinear muscles is often incomplete, which makes the control face a bottleneck problem. The aim of this paper is to design a method to improve the motion performance of musculoskeletal robots in partially observable scenarios, and to leverage the ontology knowledge to enhance the algorithm’s adaptability to musculoskeletal robots that have undergone changes. Design/methodology/approach A memory and attention-based reinforcement learning method is proposed for musculoskeletal robots with prior knowledge of muscle synergies. First, to deal with partially observed states available to musculoskeletal robots, a memory and attention-based network architecture is proposed for inferring more sufficient and intrinsic states. Second, inspired by muscle synergy hypothesis in neuroscience, prior knowledge of a musculoskeletal robot’s muscle synergies is embedded in network structure and reward shaping. Findings Based on systematic validation, it is found that the proposed method demonstrates superiority over the traditional twin delayed deep deterministic policy gradients (TD3) algorithm. A musculoskeletal robot with highly redundant, nonlinear muscles is adopted to implement goal-directed tasks. In the case of 21-dimensional states, the learning efficiency and accuracy are significantly improved compared with the traditional TD3 algorithm; in the case of 13-dimensional states without velocities and information from the end effector, the traditional TD3 is unable to complete the reaching tasks, while the proposed method breaks through this bottleneck problem. Originality/value In this paper, a novel memory and attention-based reinforcement learning method with prior knowledge of muscle synergies is proposed for musculoskeletal robots to deal with partially observable scenarios. Compared with the existing methods, the proposed method effectively improves the performance. Furthermore, this paper promotes the fusion of neuroscience and robotics.

Publisher

Emerald

Reference43 articles.

1. Huxley-type cross-bridge models in Largeish-scale musculoskeletal models; an evaluation of computational cost;Journal of Biomechanics,2019

2. Neural manifold modulated continual reinforcement learning for musculoskeletal robots,2022

3. Twin-delayed DDPG: a deep reinforcement learning technique to model a continuous movement of an intelligent robot agent,2019

4. Opensim: open-source software to create and analyze dynamic simulations of movement;IEEE Transactions on Biomedical Engineering,2007

5. Gate-variants of gated recurrent unit (GRU) neural networks, in ‘2017 IEEE 60th international midwest symposium on circuits and systems (MWSCAS)’, IEEE,2017