跳到主要导航 跳到搜索 跳到主要内容

Learning Rearrangement Manipulation via Scene Prediction in Point Cloud

  • Beijing Institute of Technology

科研成果: 期刊稿件文章同行评审

摘要

Predicting scene evolution conditioned on robotic actions is a vital technique in modeling robot manipulations. Previous studies have primarily focused on learning spatiotemporally continuous actions like Cartesian displacements, thus applying them to planar-pushing tasks. In this letter, we propose a scene prediction model that learns higher-level robot actions, such as grasping and pick-and-place, and is applicable to planning such actions in rearrangement manipulations. The model takes partially observed point clouds (e.g., from a single camera) and robot pick-and-place actions to predict point clouds of future scenes. The model directly learns scene prediction using point cloud observation and representation, without requiring prior knowledge of object properties like cad models, pose, or instance segmentation. We train the model only on a synthetic dataset acquired entirely automatically with minimal human intervention. Our experiments validate that our model can substantially learn robot grasping and pick-and-place actions and show that, when integrated into a sample-based planning framework, predicting scenes in point clouds outperforms the image-based baseline in grasping and rearrangement manipulation. Moreover, results show that our method can be directly transferred to real-world environments without fine-tuning and show promising performance on a collection of 18 household objects.

源语言英语
页(从-至)11090-11097
页数8
期刊IEEE Robotics and Automation Letters
9
12
DOI
出版状态已出版 - 2024

学术指纹

探究 'Learning Rearrangement Manipulation via Scene Prediction in Point Cloud' 的科研主题。它们共同构成独一无二的学术指纹。

引用此