NaturalConv: A Chinese Dialogue Dataset Towards Multi-turn Topic-driven Conversation

03/03/2021
by   Xiaoyang Wang, et al.
1

In this paper, we propose a Chinese multi-turn topic-driven conversation dataset, NaturalConv, which allows the participants to chat anything they want as long as any element from the topic is mentioned and the topic shift is smooth. Our corpus contains 19.9K conversations from six domains, and 400K utterances with an average turn number of 20.1. These conversations contain in-depth discussions on related topics or widely natural transition between multiple topics. We believe either way is normal for human conversation. To facilitate the research on this corpus, we provide results of several benchmark models. Comparative results show that for this dataset, our current models are not able to provide significant improvement by introducing background knowledge/topic. Therefore, the proposed dataset should be a good benchmark for further research to evaluate the validity and naturalness of multi-turn conversation systems. Our dataset is available at https://ai.tencent.com/ailab/nlp/dialogue/#datasets.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/08/2020

KdConv: A Chinese Multi-domain Dialogue Dataset Towards Multi-turn Knowledge-driven Conversation

The research of knowledge-driven conversational systems is largely limit...
research
06/13/2019

Proactive Human-Machine Conversation with Explicit Conversation Goals

Though great progress has been made for human-machine conversation, curr...
research
05/23/2023

Multi-Granularity Prompts for Topic Shift Detection in Dialogue

The goal of dialogue topic shift detection is to identify whether the cu...
research
10/15/2020

Response Selection for Multi-Party Conversations with Dynamic Topic Tracking

While participants in a multi-party multi-turn conversation simultaneous...
research
09/01/2022

Exploring Effective Information Utilization in Multi-Turn Topic-Driven Conversations

Conversations are always related to certain topics. However, it is chall...
research
10/16/2022

CDConv: A Benchmark for Contradiction Detection in Chinese Conversations

Dialogue contradiction is a critical issue in open-domain dialogue syste...
research
05/02/2023

Topic Shift Detection in Chinese Dialogues: Corpus and Benchmark

Dialogue topic shift detection is to detect whether an ongoing topic has...

Please sign up or login with your details

Forgot password? Click here to reset