Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Yiruo Cheng

CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation

Oct 30, 2024

Yiruo Cheng, Kelong Mao, Ziliang Zhao, Guanting Dong, Hongjin Qian, Yongkang Wu, Tetsuya Sakai, Ji-Rong Wen, Zhicheng Dou

Figure 1 for CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation

Figure 2 for CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation

Figure 3 for CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation

Figure 4 for CORAL: Benchmarking Multi-turn Conversational Retrieval-Augmentation Generation

Abstract:Retrieval-Augmented Generation (RAG) has become a powerful paradigm for enhancing large language models (LLMs) through external knowledge retrieval. Despite its widespread attention, existing academic research predominantly focuses on single-turn RAG, leaving a significant gap in addressing the complexities of multi-turn conversations found in real-world applications. To bridge this gap, we introduce CORAL, a large-scale benchmark designed to assess RAG systems in realistic multi-turn conversational settings. CORAL includes diverse information-seeking conversations automatically derived from Wikipedia and tackles key challenges such as open-domain coverage, knowledge intensity, free-form responses, and topic shifts. It supports three core tasks of conversational RAG: passage retrieval, response generation, and citation labeling. We propose a unified framework to standardize various conversational RAG methods and conduct a comprehensive evaluation of these methods on CORAL, demonstrating substantial opportunities for improving existing approaches.

Via

Access Paper or Ask Questions

A Survey of Conversational Search

Oct 21, 2024

Fengran Mo, Kelong Mao, Ziliang Zhao, Hongjin Qian, Haonan Chen, Yiruo Cheng, Xiaoxi Li, Yutao Zhu, Zhicheng Dou, Jian-Yun Nie

Figure 1 for A Survey of Conversational Search

Figure 2 for A Survey of Conversational Search

Figure 3 for A Survey of Conversational Search

Figure 4 for A Survey of Conversational Search

Abstract:As a cornerstone of modern information access, search engines have become indispensable in everyday life. With the rapid advancements in AI and natural language processing (NLP) technologies, particularly large language models (LLMs), search engines have evolved to support more intuitive and intelligent interactions between users and systems. Conversational search, an emerging paradigm for next-generation search engines, leverages natural language dialogue to facilitate complex and precise information retrieval, thus attracting significant attention. Unlike traditional keyword-based search engines, conversational search systems enhance user experience by supporting intricate queries, maintaining context over multi-turn interactions, and providing robust information integration and processing capabilities. Key components such as query reformulation, search clarification, conversational retrieval, and response generation work in unison to enable these sophisticated interactions. In this survey, we explore the recent advancements and potential future directions in conversational search, examining the critical modules that constitute a conversational search system. We highlight the integration of LLMs in enhancing these systems and discuss the challenges and opportunities that lie ahead in this dynamic field. Additionally, we provide insights into real-world applications and robust evaluations of current conversational search systems, aiming to guide future research and development in conversational search.

* 35 pages, 8 figures, continue to update

Via

Access Paper or Ask Questions

Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding

Feb 20, 2024

Yiruo Cheng, Kelong Mao, Zhicheng Dou

Figure 1 for Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding

Figure 2 for Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding

Figure 3 for Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding

Figure 4 for Interpreting Conversational Dense Retrieval by Rewriting-Enhanced Inversion of Session Embedding

Abstract:Conversational dense retrieval has shown to be effective in conversational search. However, a major limitation of conversational dense retrieval is their lack of interpretability, hindering intuitive understanding of model behaviors for targeted improvements. This paper presents CONVINV, a simple yet effective approach to shed light on interpretable conversational dense retrieval models. CONVINV transforms opaque conversational session embeddings into explicitly interpretable text while faithfully maintaining their original retrieval performance as much as possible. Such transformation is achieved by training a recently proposed Vec2Text model based on the ad-hoc query encoder, leveraging the fact that the session and query embeddings share the same space in existing conversational dense retrieval. To further enhance interpretability, we propose to incorporate external interpretable query rewrites into the transformation process. Extensive evaluations on three conversational search benchmarks demonstrate that CONVINV can yield more interpretable text and faithfully preserve original retrieval performance than baselines. Our work connects opaque session embeddings with transparent query rewriting, paving the way toward trustworthy conversational search.

Via

Access Paper or Ask Questions