Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Juyeon Kim

KGMEL: Knowledge Graph-Enhanced Multimodal Entity Linking

Apr 21, 2025

Juyeon Kim, Geon Lee, Taeuk Kim, Kijung Shin

Abstract:Entity linking (EL) aligns textual mentions with their corresponding entities in a knowledge base, facilitating various applications such as semantic search and question answering. Recent advances in multimodal entity linking (MEL) have shown that combining text and images can reduce ambiguity and improve alignment accuracy. However, most existing MEL methods overlook the rich structural information available in the form of knowledge-graph (KG) triples. In this paper, we propose KGMEL, a novel framework that leverages KG triples to enhance MEL. Specifically, it operates in three stages: (1) Generation: Produces high-quality triples for each mention by employing vision-language models based on its text and images. (2) Retrieval: Learns joint mention-entity representations, via contrastive learning, that integrate text, images, and (generated or KG) triples to retrieve candidate entities for each mention. (3) Reranking: Refines the KG triples of the candidate entities and employs large language models to identify the best-matching entity for the mention. Extensive experiments on benchmark datasets demonstrate that KGMEL outperforms existing methods. Our code and datasets are available at: https://github.com/juyeonnn/KGMEL.

* SIGIR 2025 (Short)

Via

Access Paper or Ask Questions

Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback

Oct 05, 2024

Fatemeh Pesaran Zadeh, Juyeon Kim, Jin-Hwa Kim, Gunhee Kim

Figure 1 for Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback

Figure 2 for Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback

Figure 3 for Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback

Figure 4 for Text2Chart31: Instruction Tuning for Chart Generation with Automatic Feedback

Abstract:Large language models (LLMs) have demonstrated strong capabilities across various language tasks, notably through instruction-tuning methods. However, LLMs face challenges in visualizing complex, real-world data through charts and plots. Firstly, existing datasets rarely cover a full range of chart types, such as 3D, volumetric, and gridded charts. Secondly, supervised fine-tuning methods do not fully leverage the intricate relationships within rich datasets, including text, code, and figures. To address these challenges, we propose a hierarchical pipeline and a new dataset for chart generation. Our dataset, Text2Chart31, includes 31 unique plot types referring to the Matplotlib library, with 11.1K tuples of descriptions, code, data tables, and plots. Moreover, we introduce a reinforcement learning-based instruction tuning technique for chart generation tasks without requiring human feedback. Our experiments show that this approach significantly enhances the model performance, enabling smaller models to outperform larger open-source models and be comparable to state-of-the-art proprietary models in data visualization tasks. We make the code and dataset available at https://github.com/fatemehpesaran310/Text2Chart31.

* EMNLP 2024 Main. Code and dataset are released at https://github.com/fatemehpesaran310/Text2Chart31

Via

Access Paper or Ask Questions

Re-Ex: Revising after Explanation Reduces the Factual Errors in LLM Responses

Feb 27, 2024

Juyeon Kim, Jeongeun Lee, Yoonho Chang, Chanyeol Choi, Junseong Kim, Jy-yong Sohn

Figure 1 for Re-Ex: Revising after Explanation Reduces the Factual Errors in LLM Responses

Figure 2 for Re-Ex: Revising after Explanation Reduces the Factual Errors in LLM Responses

Figure 3 for Re-Ex: Revising after Explanation Reduces the Factual Errors in LLM Responses

Figure 4 for Re-Ex: Revising after Explanation Reduces the Factual Errors in LLM Responses

Abstract:Mitigating hallucination issues is one of the main challenges of LLMs we need to overcome, in order to reliably use them in real-world scenarios. Recently, various methods are proposed to check the factual errors in the LLM-generated texts and revise them accordingly, to reduce the hallucination issue. In this paper, we propose Re-Ex, a method of revising LLM-generated texts, which introduces a novel step dubbed as the factual error explanation step. Re-Ex revises the initial response of LLMs using 3-steps: first, external tools are used to get the evidences on the factual errors in the response; second, LLMs are instructed to explain the problematic parts of the response based on the evidences gathered in the first step; finally, LLMs revise the response using the explanation obtained in the second step. In addition to the explanation step, we propose new prompting techniques to reduce the amount of tokens and wall-clock time required for the response revision process. Compared with existing methods including Factool, CoVE, and RARR, Re-Ex provides better revision performance with less time and fewer tokens in multiple benchmarks.

* Preprint

Via

Access Paper or Ask Questions