Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension

Feb 17, 2023

Zhi Zhang, Helen Yannakoudakis, Xiantong Zhen, Ekaterina Shutova

Figure 1 for CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension

Figure 2 for CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension

Figure 3 for CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension

Figure 4 for CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension

Share this with someone who'll enjoy it:

Abstract:The task of multimodal referring expression comprehension (REC), aiming at localizing an image region described by a natural language expression, has recently received increasing attention within the research comminity. In this paper, we specifically focus on referring expression comprehension with commonsense knowledge (KB-Ref), a task which typically requires reasoning beyond spatial, visual or semantic information. We propose a novel framework for Commonsense Knowledge Enhanced Transformers (CK-Transformer) which effectively integrates commonsense knowledge into the representations of objects in an image, facilitating identification of the target objects referred to by the expressions. We conduct extensive experiments on several benchmarks for the task of KB-Ref. Our results show that the proposed CK-Transformer achieves a new state of the art, with an absolute improvement of 3.14% accuracy over the existing state of the art.

View paper on

Share this with someone who'll enjoy it:

Title:CK-Transformer: Commonsense Knowledge Enhanced Transformers for Referring Expression Comprehension

Paper and Code