Chemical-Reaction-Aware Molecule Representation Learning

09/21/2021
by   Hongwei Wang, et al.
16

Molecule representation learning (MRL) methods aim to embed molecules into a real vector space. However, existing SMILES-based (Simplified Molecular-Input Line-Entry System) or GNN-based (Graph Neural Networks) MRL methods either take SMILES strings as input that have difficulty in encoding molecule structure information, or over-emphasize the importance of GNN architectures but neglect their generalization ability. Here we propose using chemical reactions to assist learning molecule representation. The key idea of our approach is to preserve the equivalence of molecules with respect to chemical reactions in the embedding space, i.e., forcing the sum of reactant embeddings and the sum of product embeddings to be equal for each chemical equation. This constraint is proven effective to 1) keep the embedding space well-organized and 2) improve the generalization ability of molecule embeddings. Moreover, our model can use any GNN as the molecule encoder and is thus agnostic to GNN architectures. Experimental results demonstrate that our method achieves state-of-the-art performance in a variety of downstream tasks, e.g., 17.4 in chemical reaction prediction, 2.3 prediction, and 18.5 respectively, over the best baseline method. The code is available at https://github.com/hwwang55/MolR.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/12/2020

Weisfeiler-Lehman Embedding for Molecular Graph Neural Networks

A graph neural network (GNN) is a good choice for predicting the chemica...
research
07/08/2020

Graph Neural Networks for the Prediction of Substrate-Specific Organic Reaction Conditions

We present a systematic investigation using graph neural networks (GNNs)...
research
03/18/2021

MARS: Markov Molecular Sampling for Multi-objective Drug Discovery

Searching for novel molecules with desired chemical properties is crucia...
research
11/22/2019

Approaching Small Molecule Prioritization as a Cross-Modal Information Retrieval Task through Coordinated Representation Learning

Modeling the relationship between chemical structure and molecular activ...
research
09/11/2023

Learning the Geodesic Embedding with Graph Neural Networks

We present GeGnn, a learning-based method for computing the approximate ...
research
04/27/2022

Multimodal Transformer-based Model for Buchwald-Hartwig and Suzuki-Miyaura Reaction Yield Prediction

Predicting the yield percentage of a chemical reaction is useful in many...
research
08/02/2022

AI-driven Hypernetwork of Organic Chemistry: Network Statistics and Applications in Reaction Classification

Rapid discovery of new reactions and molecules in recent years has been ...

Please sign up or login with your details

Forgot password? Click here to reset