FoundationGrasp: Generalizable Task-Oriented Grasping with Foundation   Models

Chao Tang; Dehao Huang; Wenlong Dong; Ruinian Xu; Hong Zhang

arXiv:2404.10399·cs.RO·October 10, 2024·2 cites

FoundationGrasp: Generalizable Task-Oriented Grasping with Foundation Models

Chao Tang, Dehao Huang, Wenlong Dong, Ruinian Xu, Hong Zhang

PDF

Open Access

TL;DR

FoundationGrasp introduces a foundation model-based framework for task-oriented grasping that generalizes well to new objects and tasks, validated through extensive experiments and real-robot tests.

Contribution

It leverages foundation models to enable open-ended, generalizable task-oriented grasping, surpassing prior methods limited to closed-set knowledge.

Findings

01

Outperforms existing methods on LaViA-TaskGrasp dataset

02

Successfully generalizes to novel objects and tasks

03

Validated in real-robot grasping experiments

Abstract

Task-oriented grasping (TOG), which refers to synthesizing grasps on an object that are configurationally compatible with the downstream manipulation task, is the first milestone towards tool manipulation. Analogous to the activation of two brain regions responsible for semantic and geometric reasoning during cognitive processes, modeling the intricate relationship between objects, tasks, and grasps necessitates rich semantic and geometric prior knowledge about these elements. Existing methods typically restrict the prior knowledge to a closed-set scope, limiting their generalization to novel objects and tasks out of the training set. To address such a limitation, we propose FoundationGrasp, a foundation model-based TOG framework that leverages the open-ended knowledge from foundation models to learn generalizable TOG skills. Extensive experiments are conducted on the contributed…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Robot Manipulation and Learning · Fuzzy Logic and Control Systems