Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches

Aug 30, 2024

Chao Wang, Neo Wu, Lin Ning, Luyang Liu, Jun Xie, Shawn O'Banion, Bradley Green

Figure 1 for UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches

Figure 2 for UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches

Figure 3 for UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches

Figure 4 for UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches

Share this with someone who'll enjoy it:

Abstract:Large language models (LLMs) have shown remarkable capabilities in generating user summaries from a long list of raw user activity data. These summaries capture essential user information such as preferences and interests, and therefore are invaluable for LLM-based personalization applications, such as explainable recommender systems. However, the development of new summarization techniques is hindered by the lack of ground-truth labels, the inherent subjectivity of user summaries, and human evaluation which is often costly and time-consuming. To address these challenges, we introduce \UserSumBench, a benchmark framework designed to facilitate iterative development of LLM-based summarization approaches. This framework offers two key components: (1) A reference-free summary quality metric. We show that this metric is effective and aligned with human preferences across three diverse datasets (MovieLens, Yelp and Amazon Review). (2) A novel robust summarization method that leverages time-hierarchical summarizer and self-critique verifier to produce high-quality summaries while eliminating hallucination. This method serves as a strong baseline for further innovation in summarization techniques.

View paper on

Share this with someone who'll enjoy it:

Title:UserSumBench: A Benchmark Framework for Evaluating User Summarization Approaches

Paper and Code