SinoTechIntel Academic Portal
🏛️ Indexed Academic Journal

Chinese Journal of Mechanical Engineering

Access authentic peer-reviewed engineering methodologies, experimental datasets, and scientific literature published in this journal on SinoTechIntel.

Total Research Papers: 101
Access: 100% Free Open Access
Browse by Publication Year & VolumeReset All Filters ✕

Published Research PapersFiltered: Year 2025 • 38 • 51

Showing 1 of 101 peer-reviewed papers with full Graphical Abstracts.

Original ResearchVol. 38, Issue 51 • pp. 100-112DOI: 10.1186/s10033-025-01204-yJan 15, 2025

Learning Manipulation from Expert Demonstrations Based on Multiple Data Associations and Physical Constraints

Authors: Yangqing Ye, Yaojie Mao, Shiming Qiu, Chuan’guo Tang, Zhirui Pan, Weiwei Wan, Shibo Cai, Guanjun Bao

Learning from demonstration is widely regarded as a promising paradigm for robots to acquire diverse skills. Other than the artificial learning from observation-action pairs for machines, humans can learn to imitate in a more versatile and effective manner: acquiring skills through mere “observation”. Video to Command task is widely perceived as a promising approach for task-based learning, which yet faces two key challenges: (1) High redundancy and low frame rate of fine-grained action sequences make it difficult to manipulate objects robustly and accurately. (2) Video to Command models often prioritize accuracy and richness of output commands over physical capabilities, leading to impractical or unsafe instructions for robots. This article presents a novel Video to Command framework that employs multiple data associations and physical constraints. First, we introduce an object-level appearance-contrasting multiple data association strategy to effectively associate manipulated objects in visually complex environments, capturing dynamic changes in video content. Then, we propose a multi-task Video to Command model that utilizes object-level video content changes to compile expert demonstrations into manipulation commands. Finally, a multi-task hybrid loss function is proposed to train a Video to Command model that adheres to the constraints of the physical world and manipulation tasks. Our method achieved over 10% on BLEU_N, METEOR, ROUGE_L, and CIDEr compared to the up-to-date methods. The dual-arm robot prototype was established to demonstrate the whole process of learning from an expert demonstration of multiple skills and then executing the tasks by a robot.

Learning Manipulation from Expert Demonstrations Based on Multiple Data Associations and Physical Constraints
Graphical Abstract