Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Task-Oriented Grasp Prediction with Visual-Language Inputs

Feb 28, 2023

Chao Tang, Dehao Huang, Lingxiao Meng, Weiyu Liu, Hong Zhang

Figure 1 for Task-Oriented Grasp Prediction with Visual-Language Inputs

Figure 2 for Task-Oriented Grasp Prediction with Visual-Language Inputs

Figure 3 for Task-Oriented Grasp Prediction with Visual-Language Inputs

Figure 4 for Task-Oriented Grasp Prediction with Visual-Language Inputs

Share this with someone who'll enjoy it:

Abstract:To perform household tasks, assistive robots receive commands in the form of user language instructions for tool manipulation. The initial stage involves selecting the intended tool (i.e., object grounding) and grasping it in a task-oriented manner (i.e., task grounding). Nevertheless, prior researches on visual-language grasping (VLG) focus on object grounding, while disregarding the fine-grained impact of tasks on object grasping. Task-incompatible grasping of a tool will inevitably limit the success of subsequent manipulation steps. Motivated by this problem, this paper proposes GraspCLIP, which addresses the challenge of task grounding in addition to object grounding to enable task-oriented grasp prediction with visual-language inputs. Evaluation on a custom dataset demonstrates that GraspCLIP achieves superior performance over established baselines with object grounding only. The effectiveness of the proposed method is further validated on an assistive robotic arm platform for grasping previously unseen kitchen tools given the task specification. Our presentation video is available at: https://www.youtube.com/watch?v=e1wfYQPeAXU.

* 8 pages, 8 figures, submitted to IROS 2023

View paper on

Share this with someone who'll enjoy it:

Title:Task-Oriented Grasp Prediction with Visual-Language Inputs

Paper and Code