Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Yosuke Kishinami

Bipartite-play Dialogue Collection for Practical Automatic Evaluation of Dialogue Systems

Nov 19, 2022

Shiki Sato, Yosuke Kishinami, Hiroaki Sugiyama, Reina Akama, Ryoko Tokuhisa, Jun Suzuki

Abstract:Automation of dialogue system evaluation is a driving force for the efficient development of dialogue systems. This paper introduces the bipartite-play method, a dialogue collection method for automating dialogue system evaluation. It addresses the limitations of existing dialogue collection methods: (i) inability to compare with systems that are not publicly available, and (ii) vulnerability to cheating by intentionally selecting systems to be compared. Experimental results show that the automatic evaluation using the bipartite-play method mitigates these two drawbacks and correlates as strongly with human subjectivity as existing methods.

* 9 pages, Accepted to The AACL-IJCNLP 2022 Student Research Workshop (SRW)

Via

Access Paper or Ask Questions

Target-Guided Open-Domain Conversation Planning

Sep 20, 2022

Yosuke Kishinami, Reina Akama, Shiki Sato, Ryoko Tokuhisa, Jun Suzuki, Kentaro Inui

Figure 1 for Target-Guided Open-Domain Conversation Planning

Figure 2 for Target-Guided Open-Domain Conversation Planning

Figure 3 for Target-Guided Open-Domain Conversation Planning

Figure 4 for Target-Guided Open-Domain Conversation Planning

Abstract:Prior studies addressing target-oriented conversational tasks lack a crucial notion that has been intensively studied in the context of goal-oriented artificial intelligence agents, namely, planning. In this study, we propose the task of Target-Guided Open-Domain Conversation Planning (TGCP) task to evaluate whether neural conversational agents have goal-oriented conversation planning abilities. Using the TGCP task, we investigate the conversation planning abilities of existing retrieval models and recent strong generative models. The experimental results reveal the challenges facing current technology.

* 9 pages, Accepted to The 29th International Conference on Computational Linguistics (COLING 2022)

Via

Access Paper or Ask Questions