Augmented Memory: Capitalizing on Experience Replay to Accelerate De Novo Molecular Design

Jeff Guo; Philippe Schwaller

doi:10.26434/chemrxiv-2023-qmqmq-v3

Theoretical and Computational Chemistry

Search within Theoretical and Computational Chemistry

Augmented Memory: Capitalizing on Experience Replay to Accelerate De Novo Molecular Design

22 May 2023, Version 3

Working Paper

Show author details

This content is a preprint and has not undergone peer review at the time of posting.

Abstract

Sample efficiency is a fundamental challenge in de novo molecular design. Ideally, molecular generative models should learn to satisfy desired objectives under minimal oracle evaluations (computational prediction or wet-lab experiment). This problem becomes more apparent when using oracles that can provide increased predictive accuracy but impose a significant cost. Molecular generative models have shown remarkable sample efficiency when coupled with reinforcement learn- ing, as demonstrated in the Practical Molecular Optimization (PMO) benchmark. Here, we propose a novel algorithm called Augmented Memory that combines data augmentation with experience replay. We show that scores obtained from oracle calls can be reused to update the model multiple times. We compare Augmented Memory to previously proposed algorithms and show significantly enhanced sample efficiency in an exploitation task and a drug discovery case study requiring both exploration and exploitation. Our method achieves a new state-of-the-art in the PMO benchmark which enforces a computational budget, and outperforms the previous best performing method on 19/23 tasks.

Keywords

Molecular generative model

Reinforcement learning

Deep learning

Drug discovery

Supplementary weblinks

Title

Description

Actions

Title

Augmented Memory

Description

Codebase with prepared files and instructions to reproduce all results

Actions

View

Comments

Comments are not moderated before they are posted, but they can be removed by the site moderators if they are found to be in contravention of our Commenting Policy - please read this policy before you post. Comments should be used for scholarly discussion of the content in question. You can find more information about how to use the commenting feature here .

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Version History

May 22, 2023 Version 3

May 18, 2023 Version 2

May 17, 2023 Version 1

Version Notes

The changes here accompany the code release to reproduce the benchmarking results. Following removal of some redundancy, the benchmark values for Best Agent Reminder changed very slightly (changes are minimal and do not affect the conclusions). However, to promote maximum reproducibility, the benchmark table has been updated.

Metrics

1,991

1,124

Views

Downloads

Citations

License

The content is available under CC BY 4.0

DOI

10.26434/chemrxiv-2023-qmqmq-v3

Funding

National Centre of Competence in Research (NCCR) Catalysis

180544

Author’s competing interest statement

The author(s) have declared they have no conflict of interest with regard to this content

Ethics

The author(s) have declared ethics committee/IRB approval is not relevant to this content

Augmented Memory: Capitalizing on Experience Replay to Accelerate De Novo Molecular Design

Authors

Abstract

Keywords

Supplementary weblinks

Comments

Version History

Version Notes

Metrics

License

DOI

Funding

Author’s competing interest statement

Ethics

Share