首页 /研究 /On the Effectiveness of Retrieval, Alignment, and Replay in Manipulation
MANIPULATION

On the Effectiveness of Retrieval, Alignment, and Replay in Manipulation

Norman Di Palo, Edward Johns

发表年份
2024
引用次数
10

摘要

Imitation learning with visual observations is notoriously inefficient when addressed with end-to-end behavioural cloning methods. In this letter, we explore an alternative paradigm which decomposes reasoning into three phases. First, a <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">retrieval</i> phase, which informs the robot <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">what</i> it can do with an object. Second, an <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">alignment</i> phase, which informs the robot <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">where</i> to interact with the object. And third, a <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">replay</i> phase, which informs the robot <italic xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xlink="http://www.w3.org/1999/xlink">how</i> to interact with the object. Through a series of real-world experiments on everyday tasks, such as grasping, pouring, and inserting objects, we show that this decomposition brings unprecedented learning efficiency, and effective inter- and intra-class generalisation.

关键词

Computer scienceArtificial intelligenceObject (grammar)Class (philosophy)Information retrieval

相关论文

查看 MANIPULATION 分类全部论文