Towards NLP-supported Semantic Data Management

05/14/2020
by   Andreas Burgdorf, et al.
0

The heterogeneity of data poses a great challenge when data from different sources is to be merged for one application. Solutions for this are offered, for example, by ontology-based data management (OBDM). A challenge of OBDM is the automatic creation of semantic models from datasets. This process is typically performed either data- or label-driven and always involves manual human intervention. We identified textual descriptions of data, a form of metadata, quickly to be produced and consumed by humans, as third possible basis for automatic semantic modelling. In this paper, we present, how we plan to use textual descriptions to enhance semantic data management. We will use state of the art NLP technologies to identify concepts within textual descriptions and build semantic models from this in combination with an evolving ontology. We will use automatically identified models in combination with the human data provider to automatically extend the ontology so that it learns new verified concepts over time. Finally, we will use the created ontology and automatically identified semantic models to either rate descriptions for new data sources or even to automatically generate descriptive texts that are easier to understand by the human user than formal models. We present the procedure which we plan for the ongoing research, as well as expected outcomes.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/16/2016

Learning the Semantics of Structured Data Sources

Information sources such as relational databases, spreadsheets, XML, JSO...
research
07/14/2021

The I-ADOPT Interoperability Framework for FAIRer data descriptions of biodiversity

Biodiversity, the variation within and between species and ecosystems, i...
research
12/10/2013

OntoVerbal: a Generic Tool and Practical Application to SNOMED CT

Ontology development is a non-trivial task requiring expertise in the ch...
research
03/07/2019

Automatic Ontology Learning from Domain-Specific Short Unstructured Text Data

Ontology learning is a critical task in industry, dealing with identifyi...
research
08/16/2018

Sememe Prediction: Learning Semantic Knowledge from Unstructured Textual Wiki Descriptions

Huge numbers of new words emerge every day, leading to a great need for ...
research
05/30/2017

Preliminary results on Ontology-based Open Data Publishing

Despite the current interest in Open Data publishing, a formal and compr...
research
08/26/2022

Need for Design Patterns: Interoperability Issues and Modelling Challenges for Observational Data

Interoperability issues concerning observational data have gained attent...

Please sign up or login with your details

Forgot password? Click here to reset