Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Jun 24, 2024

Yuting Mei, Linli Yao, Qin Jin

Figure 1 for UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Figure 2 for UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Figure 3 for UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Figure 4 for UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Share this with someone who'll enjoy it:

Abstract:With the surge in the amount of video data, video summarization techniques, including visual-modal(VM) and textual-modal(TM) summarization, are attracting more and more attention. However, unimodal summarization inevitably loses the rich semantics of the video. In this paper, we focus on a more comprehensive video summarization task named Bimodal Semantic Summarization of Videos (BiSSV). Specifically, we first construct a large-scale dataset, BIDS, in (video, VM-Summary, TM-Summary) triplet format. Unlike traditional processing methods, our construction procedure contains a VM-Summary extraction algorithm aiming to preserve the most salient content within long videos. Based on BIDS, we propose a Unified framework UBiSS for the BiSSV task, which models the saliency information in the video and generates a TM-summary and VM-summary simultaneously. We further optimize our model with a list-wise ranking-based objective to improve its capacity to capture highlights. Lastly, we propose a metric, $NDCG_{MS}$, to provide a joint evaluation of the bimodal summary. Experiments show that our unified framework achieves better performance than multi-stage summarization pipelines. Code and data are available at https://github.com/MeiYutingg/UBiSS.

* Proceedings of the 2024 International Conference on Multimedia Retrieval, May 2024, Pages 1034-1042 * Accepted by ACM International Conference on Multimedia Retrieval (ICMR'24)

View paper on

Share this with someone who'll enjoy it:

Title:UBiSS: A Unified Framework for Bimodal Semantic Summarization of Videos

Paper and Code