On the Feasibility of Predicting Questions being Forgotten in Stack Overflow

10/29/2021
by   Thi Huyen Nguyen, et al.
5

For their attractiveness, comprehensiveness and dynamic coverage of relevant topics, community-based question answering sites such as Stack Overflow heavily rely on the engagement of their communities: Questions on new technologies, technology features as well as technology versions come up and have to be answered as technology evolves (and as community members gather experience with it). At the same time, other questions cease in importance over time, finally becoming irrelevant to users. Beyond filtering low-quality questions, "forgetting" questions, which have become redundant, is an important step for keeping the Stack Overflow content concise and useful. In this work, we study this managed forgetting task for Stack Overflow. Our work is based on data from more than a decade (2008 - 2019) - covering 18.1M questions, that are made publicly available by the site itself. For establishing a deeper understanding, we first analyze and characterize the set of questions about to be forgotten, i.e., questions that get a considerable number of views in the current period but become unattractive in the near future. Subsequently, we examine the capability of a wide range of features in predicting such forgotten questions in different categories. We find some categories in which those questions are more predictable. We also discover that the text-based features are surprisingly not helpful in this prediction task, while the meta information is much more predictive.

READ FULL TEXT

page 4

page 6

page 9

research
10/04/2022

Mining Duplicate Questions of Stack Overflow

There has a been a significant rise in the use of Community Question Ans...
research
12/15/2022

Best-Answer Prediction in Q A Sites Using User Information

Community Question Answering (CQA) sites have spread and multiplied sign...
research
05/03/2019

Question Relatedness on Stack Overflow: The Task, Dataset, and Corpus-inspired Models

Domain-specific community question answering is becoming an integral par...
research
10/12/2017

How to Ask for Technical Help? Evidence-based Guidelines for Writing Questions on Stack Overflow

Context: The success of Stack Overflow and other community-based questio...
research
11/17/2018

Deep Dive into Anonymity: A Large Scale Analysis of Quora Questions

Anonymity forms an integral and important part of our digital life. It e...
research
07/02/2018

ColdRoute: Effective Routing of Cold Questions in Stack Exchange Sites

Routing questions in Community Question Answer services (CQAs) such as S...
research
10/18/2020

Querent Intent in Multi-Sentence Questions

Multi-sentence questions (MSQs) are sequences of questions connected by ...

Please sign up or login with your details

Forgot password? Click here to reset