Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:On the Use of Linguistic Features for the Evaluation of Generative Dialogue Systems

Apr 13, 2021

Ian Berlot-Attwell, Frank Rudzicz

Figure 1 for On the Use of Linguistic Features for the Evaluation of Generative Dialogue Systems

Figure 2 for On the Use of Linguistic Features for the Evaluation of Generative Dialogue Systems

Figure 3 for On the Use of Linguistic Features for the Evaluation of Generative Dialogue Systems

Figure 4 for On the Use of Linguistic Features for the Evaluation of Generative Dialogue Systems

Share this with someone who'll enjoy it:

Abstract:Automatically evaluating text-based, non-task-oriented dialogue systems (i.e., `chatbots') remains an open problem. Previous approaches have suffered challenges ranging from poor correlation with human judgment to poor generalization and have often required a gold standard reference for comparison or human-annotated data. Extending existing evaluation methods, we propose that a metric based on linguistic features may be able to maintain good correlation with human judgment and be interpretable, without requiring a gold-standard reference or human-annotated data. To support this proposition, we measure and analyze various linguistic features on dialogues produced by multiple dialogue models. We find that the features' behaviour is consistent with the known properties of the models tested, and is similar across domains. We also demonstrate that this approach exhibits promising properties such as zero-shot generalization to new domains on the related task of evaluating response relevance.

View paper on

Share this with someone who'll enjoy it:

Title:On the Use of Linguistic Features for the Evaluation of Generative Dialogue Systems

Paper and Code