Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:NumHG: A Dataset for Number-Focused Headline Generation

Sep 04, 2023

Jian-Tao Huang, Chung-Chi Chen, Hen-Hsen Huang, Hsin-Hsi Chen

Figure 1 for NumHG: A Dataset for Number-Focused Headline Generation

Figure 2 for NumHG: A Dataset for Number-Focused Headline Generation

Figure 3 for NumHG: A Dataset for Number-Focused Headline Generation

Figure 4 for NumHG: A Dataset for Number-Focused Headline Generation

Share this with someone who'll enjoy it:

Abstract:Headline generation, a key task in abstractive summarization, strives to condense a full-length article into a succinct, single line of text. Notably, while contemporary encoder-decoder models excel based on the ROUGE metric, they often falter when it comes to the precise generation of numerals in headlines. We identify the lack of datasets providing fine-grained annotations for accurate numeral generation as a major roadblock. To address this, we introduce a new dataset, the NumHG, and provide over 27,000 annotated numeral-rich news articles for detailed investigation. Further, we evaluate five well-performing models from previous headline generation tasks using human evaluation in terms of numerical accuracy, reasonableness, and readability. Our study reveals a need for improvement in numerical accuracy, demonstrating the potential of the NumHG dataset to drive progress in number-focused headline generation and stimulate further discussions in numeral-focused text generation.

* NumEval@SemEval-2024 Dataset

View paper on

Share this with someone who'll enjoy it:

Title:NumHG: A Dataset for Number-Focused Headline Generation

Paper and Code