arXiv · 1809.06641
Talking to myself: self-dialogues as data for conversational agents
Abstract
Conversational agents are gaining popularity with the increasing ubiquity of smart devices. However, training agents in a data driven manner is challenging due to a lack of suitable corpora. This paper presents a novel method for gathering topical, unstructured conversational data in an efficient way: self-dialogues through crowd-sourcing. Alongside this paper, we include a corpus of 3.6 million words across 23 topics. We argue the utility of the corpus by comparing self-dialogues with standard two-party conversations as well as data from other corpora.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Joachim Fainberg, Ben Krause, Mihai Dobre, Marco Damonte, Emmanuel Kahembwe, Daniel Duma, Bonnie Webber, Federico Fancellu. 2018-09-19. Talking to myself: self-dialogues as data for conversational agents. https://arxiv.org/abs/1809.06641
Cite the original work for its findings. Save a collection to share your selection of sources.