How “open” are the conversations with open-domain chatbots? A proposal for Speech Event based evaluation

A. Seza Doğruöz; Gabriel Skantze

doi:10.18653/v1/2021.sigdial-1.41

How “open” are the conversations with open-domain chatbots? A proposal for Speech Event based evaluation

Abstract

Open-domain chatbots are supposed to converse freely with humans without being restricted to a topic, task or domain. However, the boundaries and/or contents of open-domain conversations are not clear. To clarify the boundaries of “openness”, we conduct two studies: First, we classify the types of “speech events” encountered in a chatbot evaluation data set (i.e., Meena by Google) and find that these conversations mainly cover the “small talk” category and exclude the other speech event categories encountered in real life human-human communication. Second, we conduct a small-scale pilot study to generate online conversations covering a wider range of speech event categories between two humans vs. a human and a state-of-the-art chatbot (i.e., Blender by Facebook). A human evaluation of these generated conversations indicates a preference for human-human conversations, since the human-chatbot conversations lack coherence in most speech event categories. Based on these results, we suggest (a) using the term “small talk” instead of “open-domain” for the current chatbots which are not that “open” in terms of conversational abilities yet, and (b) revising the evaluation methods to test the chatbot conversations against other speech events.

Anthology ID:: 2021.sigdial-1.41
Volume:: Proceedings of the 22nd Annual Meeting of the Special Interest Group on Discourse and Dialogue
Month:: July
Year:: 2021
Address:: Singapore and Online
Editors:: Haizhou Li, Gina-Anne Levow, Zhou Yu, Chitralekha Gupta, Berrak Sisman, Siqi Cai, David Vandyke, Nina Dethlefs, Yan Wu, Junyi Jessy Li
Venue:: SIGDIAL
SIG:: SIGDIAL
Publisher:: Association for Computational Linguistics
Note:
Pages:: 392–402
Language:
URL:: https://preview.aclanthology.org/jlcl-multiple-ingestion/2021.sigdial-1.41/
DOI:: 10.18653/v1/2021.sigdial-1.41
Bibkey:
Cite (ACL):: A. Seza Doğruöz and Gabriel Skantze. 2021. How “open” are the conversations with open-domain chatbots? A proposal for Speech Event based evaluation. In Proceedings of the 22nd Annual Meeting of the Special Interest Group on Discourse and Dialogue, pages 392–402, Singapore and Online. Association for Computational Linguistics.
Cite (Informal):: How “open” are the conversations with open-domain chatbots? A proposal for Speech Event based evaluation (Doğruöz & Skantze, SIGDIAL 2021)
Copy Citation:
PDF:: https://preview.aclanthology.org/jlcl-multiple-ingestion/2021.sigdial-1.41.pdf
Video:: https://www.youtube.com/watch?v=bYXcZg_VWiE

PDF Cite Search Video Fix data