Bringing Structure into Summaries: Crowdsourcing a Benchmark Corpus of Concept Maps

04/14/2017
by   Tobias Falke, et al.
0

Concept maps can be used to concisely represent important information and bring structure into large document collections. Therefore, we study a variant of multi-document summarization that produces summaries in the form of concept maps. However, suitable evaluation datasets for this task are currently missing. To close this gap, we present a newly created corpus of concept maps that summarize heterogeneous collections of web documents on educational topics. It was created using a novel crowdsourcing approach that allows us to efficiently determine important elements in large document collections. We release the corpus along with a baseline system and proposed evaluation protocol to enable further research on this variant of summarization.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/23/2020

AQuaMuSe: Automatically Generating Datasets for Query-Based Multi-Document Summarization

Summarization is the task of compressing source document(s) into coheren...
research
07/09/2023

A Personalized Reinforcement Learning Summarization Service for Learning Structure from Unstructured Data

The exponential growth of textual data has created a crucial need for to...
research
06/22/2022

Multi-LexSum: Real-World Summaries of Civil Rights Lawsuits at Multiple Granularities

With the advent of large language models, methods for abstractive summar...
research
07/30/2019

Abstractive Document Summarization without Parallel Data

Abstractive summarization typically relies on large collections of paire...
research
06/02/2022

TSTR: Too Short to Represent, Summarize with Details! Intro-Guided Extended Summary Generation

Many scientific papers such as those in arXiv and PubMed data collection...
research
04/29/2019

Semantic Matching of Documents from Heterogeneous Collections: A Simple and Transparent Method for Practical Applications

We present a very simple, unsupervised method for the pairwise matching ...
research
01/31/2023

Archive TimeLine Summarization (ATLS): Conceptual Framework for Timeline Generation over Historical Document Collections

Archive collections are nowadays mostly available through search engines...

Please sign up or login with your details

Forgot password? Click here to reset