Learning Robot Manipulation from Cross-Morphology Demonstration

被引：0

作者：

Salhotra, Gautam ^{[1
]}

Liu, I-Chun Arthur ^{[1
]}

Sukhatme, Gaurav S. ^{[1
]}

机构：

[1] Univ Southern Calif, Robot Embedded Syst Lab, Los Angeles, CA 90089 USA

来源：

CONFERENCE ON ROBOT LEARNING, VOL 229 | 2023年 / 229卷

关键词：

Imitation from Observation; Learning from Demonstration; TRAJECTORY OPTIMIZATION;

D O I：

暂无

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Some Learning from Demonstrations (LfD) methods handle small mismatches in the action spaces of the teacher and student. Here we address the case where the teacher's morphology is substantially different from that of the student. Our framework, Morphological Adaptation in Imitation Learning (MAIL), bridges this gap allowing us to train an agent from demonstrations by other agents with significantly different morphologies. MAIL learns from suboptimal demonstrations, so long as they provide some guidance towards a desired solution. We demonstrate MAIL on manipulation tasks with rigid and deformable objects including 3D cloth manipulation interacting with rigid obstacles. We train a visual control policy for a robot with one end-effector using demonstrations from a simulated agent with two end-effectors. MAIL shows up to 24% improvement in a normalized performance metric over LfD and non-LfD baselines. It is deployed to a real Franka Panda robot, handles multiple variations in properties for objects (size, rotation, translation), and cloth-specific properties (color, thickness, size, material). An overview is on this website.

引用

页数：21

共 56 条

[1]

Al-Hafez Firas, 2023, 11 INT C LEARN REPR

[2]

[Anonymous], 2017, P 31 INT C NEUR INF

[3]

[Anonymous], 2002 IEEERSJ INT C

[4]

[Anonymous], INT JOINT C ARTIFICI

[5]

[Anonymous], 2022, Machines, DOI [DOI 10.3390/SUN22KEYPOINTEXTRACTION, 10.3390/sun22KeypointExtraction]

[6]

Bern JM, 2019, ROBOTICS: SCIENCE AND SYSTEMS XV

[7]

Brohan A., 2022, arXiv

[8] A tutorial on the cross-entropy method [J].

De Boer, PT ;

Kroese, DP ;

Mannor, S ;

Rubinstein, RY .

ANNALS OF OPERATIONS RESEARCH, 2005, 134 (01) :19-67

[9]

Finn C., 2017, Personality development across the lifespan, P357, DOI DOI 10.1016/B978-0-12-804674-6.00022-3

[10] Super-Human Performance in Gran Turismo Sport Using Deep Reinforcement Learning [J].

Fuchs, Florian ;

Song, Yunlong ;

Kaufmann, Elia ;

Scaramuzza, Davide ;

Durr, Peter .

IEEE ROBOTICS AND AUTOMATION LETTERS, 2021, 6 (03) :4257-4264

← 1 2 3 4 5 6 →