Giving Attention to Generative VAE Models for De Novo Molecular Design

Orion Dollar; Nisarg Joshi; David A. C. Beck; Jim Pfaendtner

doi:10.26434/chemrxiv.13724629.v1

Theoretical and Computational Chemistry

Search within Theoretical and Computational Chemistry

Giving Attention to Generative VAE Models for De Novo Molecular Design

08 February 2021, Version 1

Working Paper

Show author details

This content is a preprint and has not undergone peer review at the time of posting.

Abstract

We explore the impact of adding attention to generative VAE models for molecular design. Four model types are compared: a simple recurrent VAE (RNN), a recurrent VAE with an added attention layer (RNNAttn), a transformer VAE (TransVAE) and the previous state-of-the-art (MosesVAE). The models are assessed based on their effect on the organization of the latent space (i.e. latent memory) and their ability to generate samples that are valid and novel. Additionally, the Shannon information entropy is used to measure the complexity of the latent memory in an information bottleneck theoretical framework and we define a novel metric to assess the extent to which models explore chemical phase space. All three models are trained on millions of molecules from either the ZINC or PubChem datasets. We find that both RNNAttn and TransVAE models perform substantially better when tasked with accurately reconstructing input SMILES strings than the MosesVAE or RNN models, particularly for larger molecules up to ~700 Da. The TransVAE learns a complex “molecular grammar” that includes detailed molecular substructures and high-level structural and atomic relationships. The RNNAttn models learn the most efficient compression of the input data while still maintaining good performance. The complexity of the compressed representation learned by each model type increases in the order of MosesVAE < RNNAttn < RNN < TransVAE. We find that there is an unavoidable tradeoff between model exploration and validity that is a function of the complexity of the latent memory. However, novel sampling schemes may be used that optimize this tradeoff and allow us to utilize the information-dense representations learned by the transformer in spite of their complexity.

Keywords

VAE

Variational Autoencoder

Computational Materials Discovery

Molecular Data Science

Molecular Design

Artificial Intelligence

Machine Learning

Supplementary materials

Title

Description

Actions

Title

SI ChemRxiv

Description

Actions

Supplementary weblinks

Title

Description

Actions

Title

Description

Actions

View

Comments

Comments are not moderated before they are posted, but they can be removed by the site moderators if they are found to be in contravention of our Commenting Policy - please read this policy before you post. Comments should be used for scholarly discussion of the content in question. You can find more information about how to use the commenting feature here .

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Now Published

Attention-based generative models for de novo molecular design

Orion Dollar, Nisarg Joshi, David A. C. Beck, Jim Pfaendtner journal article

Chemical Science , Volume 12, Issue 24

Online publication date: 2021

Version History

Feb 08, 2021 Version 1

Metrics

4,068

1,379

Views

Downloads

Citations

License

The content is available under CC BY NC ND 4.0

DOI

10.26434/chemrxiv.13724629.v1

Author’s competing interest statement

There is no conflict of interest

Giving Attention to Generative VAE Models for De Novo Molecular Design

Authors

Abstract

Keywords

Supplementary materials

Supplementary weblinks

Comments

Now Published

Version History

Metrics

License

DOI

Author’s competing interest statement

Share