A Deep Generative Model for Fragment-Based Molecule Generation

02/28/2020
by   Marco Podda, et al.
0

Molecule generation is a challenging open problem in cheminformatics. Currently, deep generative approaches addressing the challenge belong to two broad categories, differing in how molecules are represented. One approach encodes molecular graphs as strings of text, and learns their corresponding character-based language model. Another, more expressive, approach operates directly on the molecular graph. In this work, we address two limitations of the former: generation of invalid and duplicate molecules. To improve validity rates, we develop a language model for small molecular substructures called fragments, loosely inspired by the well-known paradigm of Fragment-Based Drug Design. In other words, we generate molecules fragment by fragment, instead of atom by atom. To improve uniqueness rates, we present a frequency-based masking strategy that helps generate molecules with infrequent fragments. We show experimentally that our model largely outperforms other language model-based competitors, reaching state-of-the-art performances typical of graph-based approaches. Moreover, generated molecules display molecular properties similar to those in the training sample, even in absence of explicit task-specific supervision.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/31/2019

Scaffold-based molecular design using graph generative model

Searching new molecules in areas like drug discovery often starts from t...
research
09/24/2019

Deep Generative Model for Sparse Graphs using Text-Based Learning with Augmentation in Generative Examination Networks

Graphs and networks are a key research tool for a variety of science fie...
research
05/17/2023

Lingo3DMol: Generation of a Pocket-based 3D Molecule using a Language Model

Structure-based drug design powered by deep generative models have attra...
research
03/05/2021

Learning to Extend Molecular Scaffolds with Structural Motifs

Recent advancements in deep learning-based modeling of molecules promise...
research
03/31/2019

Molecular geometry prediction using a deep generative graph neural network

A molecule's geometry, also known as conformation, is one of a molecule'...
research
09/18/2021

MM-Deacon: Multimodal molecular domain embedding analysis via contrastive learning

Molecular representation learning plays an essential role in cheminforma...
research
06/02/2023

Balancing Exploration and Exploitation: Disentangled β-CVAE in De Novo Drug Design

Deep generative models have recently emerged as a promising de novo drug...

Please sign up or login with your details

Forgot password? Click here to reset