Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization

Aug 19, 2024

Ran Liu, Ming Liu, Min Yu, Jianguo Jiang, Gang Li, Dan Zhang, Jingyuan Li, Xiang Meng, Weiqing Huang

Figure 1 for GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization

Figure 2 for GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization

Figure 3 for GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization

Figure 4 for GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization

Share this with someone who'll enjoy it:

Abstract:Pre-trained language models are increasingly being used in multi-document summarization tasks. However, these models need large-scale corpora for pre-training and are domain-dependent. Other non-neural unsupervised summarization approaches mostly rely on key sentence extraction, which can lead to information loss. To address these challenges, we propose a lightweight yet effective unsupervised approach called GLIMMER: a Graph and LexIcal features based unsupervised Multi-docuMEnt summaRization approach. It first constructs a sentence graph from the source documents, then automatically identifies semantic clusters by mining low-level features from raw texts, thereby improving intra-cluster correlation and the fluency of generated sentences. Finally, it summarizes clusters into natural sentences. Experiments conducted on Multi-News, Multi-XScience and DUC-2004 demonstrate that our approach outperforms existing unsupervised approaches. Furthermore, it surpasses state-of-the-art pre-trained multi-document summarization models (e.g. PEGASUS and PRIMERA) under zero-shot settings in terms of ROUGE scores. Additionally, human evaluations indicate that summaries generated by GLIMMER achieve high readability and informativeness scores. Our code is available at https://github.com/Oswald1997/GLIMMER.

* 19 pages, 7 figures. Accepted by ECAI 2024

View paper on

Share this with someone who'll enjoy it:

Title:GLIMMER: Incorporating Graph and Lexical Features in Unsupervised Multi-Document Summarization

Paper and Code