What makes us curious? analysis of a corpus of open-domain questions

10/28/2021
by   Zhaozhen Xu, et al.
14

Every day people ask short questions through smart devices or online forums to seek answers to all kinds of queries. With the increasing number of questions collected it becomes difficult to provide answers to each of them, which is one of the reasons behind the growing interest in automated question answering. Some questions are similar to existing ones that have already been answered, while others could be answered by an external knowledge source such as Wikipedia. An important question is what can be revealed by analysing a large set of questions. In 2017, "We the Curious" science centre in Bristol started a project to capture the curiosity of Bristolians: the project collected more than 10,000 questions on various topics. As no rules were given during collection, the questions are truly open-domain, and ranged across a variety of topics. One important aim for the science centre was to understand what concerns its visitors had beyond science, particularly on societal and cultural issues. We addressed this question by developing an Artificial Intelligence tool that can be used to perform various processing tasks: detection of equivalence between questions; detection of topic and type; and answering of the question. As we focused on the creation of a "generalist" tool, we trained it with labelled data from different datasets. We called the resulting model QBERT. This paper describes what information we extracted from the automated analysis of the WTC corpus of open-domain questions.

READ FULL TEXT
research
05/25/2022

QAMPARI: : An Open-domain Question Answering Benchmark for Questions with Many Answers from Multiple Paragraphs

Existing benchmarks for open-domain question answering (ODQA) typically ...
research
10/13/2021

Open-Domain Question-Answering for COVID-19 and Other Emergent Domains

Since late 2019, COVID-19 has quickly emerged as the newest biomedical d...
research
06/02/2020

Open-Domain Question Answering with Pre-Constructed Question Spaces

Open-domain question answering aims at solving the task of locating the ...
research
04/01/2023

What Does the Indian Parliament Discuss? An Exploratory Analysis of the Question Hour in the Lok Sabha

The TCPD-IPD dataset is a collection of questions and answers discussed ...
research
12/05/2022

QBERT: Generalist Model for Processing Questions

Using a single model across various tasks is beneficial for training and...
research
12/09/2019

Why I killed my copper – Highlights about the FTTO in the ESR

FTTO means Fiber To The Office, in reference to FTTH (Fibre To The Home)...
research
02/10/2019

Engaging Audiences in Virtual Museums by Interactively Prompting Guiding Questions

Virtual museums aim to promote access to cultural artifacts. However, th...

Please sign up or login with your details

Forgot password? Click here to reset