The Materials Science Procedural Text Corpus: Annotating Materials Synthesis Procedures with Shallow Semantic Structures

05/16/2019
by   Sheshera Mysore, et al.
0

Materials science literature contains millions of materials synthesis procedures described in unstructured natural language text. Large-scale analysis of these synthesis procedures would facilitate deeper scientific understanding of materials synthesis and enable automated synthesis planning. Such analysis requires extracting structured representations of synthesis procedures from the raw text as a first step. To facilitate the training and evaluation of synthesis extraction models, we introduce a dataset of 230 synthesis procedures annotated by domain experts with labeled graphs that express the semantics of the synthesis sentences. The nodes in this graph are synthesis operations and their typed arguments, and labeled edges specify relations between the nodes. We describe this new resource in detail and highlight some specific challenges to annotating scientific text with shallow semantic structure. We make the corpus available to the community to promote further research and development of scientific information extraction systems.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/22/2022

PcMSP: A Dataset for Scientific Action Graphs Extraction from Polycrystalline Materials Synthesis Procedure Text

Scientific action graphs extraction from materials synthesis procedures ...
research
11/18/2017

Automatically Extracting Action Graphs from Materials Science Synthesis Procedures

Computational synthesis planning approaches have achieved recent success...
research
01/23/2022

ULSA: Unified Language of Synthesis Actions for Representation of Synthesis Protocols

Applying AI power to predict syntheses of novel materials requires high-...
research
02/05/2023

Inorganic synthesis recommendation by machine learning materials similarity from scientific literature

Synthesis prediction is a key accelerator for the rapid design of advanc...
research
01/08/2019

Computational Register Analysis and Synthesis

The study of register in computational language research has historicall...
research
04/26/2023

Extracting Structured Seed-Mediated Gold Nanorod Growth Procedures from Literature with GPT-3

Although gold nanorods have been the subject of much research, the pathw...
research
06/20/2023

ChatGPT Chemistry Assistant for Text Mining and Prediction of MOF Synthesis

We use prompt engineering to guide ChatGPT in the automation of text min...

Please sign up or login with your details

Forgot password? Click here to reset