Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:DebateSum: A large-scale argument mining and summarization dataset

Nov 14, 2020

Allen Roush, Arvind Balaji

Figure 1 for DebateSum: A large-scale argument mining and summarization dataset

Figure 2 for DebateSum: A large-scale argument mining and summarization dataset

Figure 3 for DebateSum: A large-scale argument mining and summarization dataset

Share this with someone who'll enjoy it:

Abstract:Prior work in Argument Mining frequently alludes to its potential applications in automatic debating systems. Despite this focus, almost no datasets or models exist which apply natural language processing techniques to problems found within competitive formal debate. To remedy this, we present the DebateSum dataset. DebateSum consists of 187,386 unique pieces of evidence with corresponding argument and extractive summaries. DebateSum was made using data compiled by competitors within the National Speech and Debate Association over a 7-year period. We train several transformer summarization models to benchmark summarization performance on DebateSum. We also introduce a set of fasttext word-vectors trained on DebateSum called debate2vec. Finally, we present a search engine for this dataset which is utilized extensively by members of the National Speech and Debate Association today. The DebateSum search engine is available to the public here: http://www.debate.cards

* Accepted for oral presentation at the 7th Workshop on Argument Mining (ARGMIN 2020) held at The 28th International Conference on Computational Linguistics (COLING 2020)

View paper on

Share this with someone who'll enjoy it:

Title:DebateSum: A large-scale argument mining and summarization dataset

Paper and Code