ChemGrapher: Optical Graph Recognition of Chemical Compounds by Deep Learning

02/23/2020
by   Martijn Oldenhof, et al.
0

In drug discovery, knowledge of the graph structure of chemical compounds is essential. Many thousands of scientific articles in chemistry and pharmaceutical sciences have investigated chemical compounds, but in cases the details of the structure of these chemical compounds is published only as an images. A tool to analyze these images automatically and convert them into a chemical graph structure would be useful for many applications, such drug discovery. A few such tools are available and they are mostly derived from optical character recognition. However, our evaluation of the performance of those tools reveals that they make often mistakes in detecting the correct bond multiplicity and stereochemical information. In addition, errors sometimes even lead to missing atoms in the resulting graph. In our work, we address these issues by developing a compound recognition method based on machine learning. More specifically, we develop a deep neural network model for optical compound recognition. The deep learning solution presented here consists of a segmentation model, followed by three classification models that predict atom locations, bonds and charges. Furthermore, this model not only predicts the graph structure of the molecule but also produces all information necessary to relate each component of the resulting graph to the source image. This solution is scalable and could rapidly process thousands of images. Finally, we compare empirically the proposed method to a well-established tool and observe significant error reductions.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
09/22/2021

Differentiable Scaffolding Tree for Molecular Optimization

The structural design of functional molecules, also called molecular opt...
research
08/23/2023

MolGrapher: Graph-based Visual Recognition of Chemical Structures

The automatic analysis of chemical literature has immense potential to a...
research
11/30/2022

A Deep Learning Approach to the Prediction of Drug Side-Effects on Molecular Graphs

Predicting drug side-effects before they occur is a key task in keeping ...
research
07/25/2019

Graph Informer Networks for Molecules

In machine learning, chemical molecules are often represented by sparse ...
research
02/19/2022

Image-to-Graph Transformers for Chemical Structure Recognition

For several decades, chemical knowledge has been published in written te...
research
03/25/2021

Self-Labeling of Fully Mediating Representations by Graph Alignment

To be able to predict a molecular graph structure (W) given a 2D image o...
research
05/23/2022

MolMiner: You only look once for chemical structure recognition

Molecular structures are always depicted as 2D printed form in scientifi...

Please sign up or login with your details

Forgot password? Click here to reset