Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:TNT-KID: Transformer-based Neural Tagger for Keyword Identification

Mar 20, 2020

Matej Martinc, Blaž Škrlj, Senja Pollak

Figure 1 for TNT-KID: Transformer-based Neural Tagger for Keyword Identification

Figure 2 for TNT-KID: Transformer-based Neural Tagger for Keyword Identification

Figure 3 for TNT-KID: Transformer-based Neural Tagger for Keyword Identification

Figure 4 for TNT-KID: Transformer-based Neural Tagger for Keyword Identification

Share this with someone who'll enjoy it:

Abstract:With growing amounts of available textual data, development of algorithms capable of automatic analysis, categorization and summarization of these data has become a necessity. In this research we present a novel algorithm for keyword identification, i.e., an extraction of one or multi-word phrases representing key aspects of a given document, called Transformer-based Neural Tagger for Keyword IDentification (TNT-KID). By adapting the transformer architecture for a specific task at hand and leveraging language model pretraining on a small domain specific corpus, the model is capable of overcoming deficiencies of both supervised and unsupervised state-of-the-art approaches to keyword extraction by offering competitive and robust performance on a variety of different datasets while requiring only a fraction of manually labeled data required by the best performing systems. This study also offers thorough error analysis with valuable insights into inner workings of the model and an ablation study measuring the influence of specific components of the keyword identification workflow on the overall performance.

* Submitted to Natural Language Engineering journal

View paper on

Share this with someone who'll enjoy it:

Title:TNT-KID: Transformer-based Neural Tagger for Keyword Identification

Paper and Code