ChemRxiv
These are preliminary reports that have not been peer-reviewed. They should not be regarded as conclusive, guide clinical practice/health-related behavior, or be reported in news media as established information. For more information, please see our FAQs.
druchok_design_correction_control.pdf (1.65 MB)

Towards Efficient Generation, Correction and Properties Control of Unique Drug-like Structures

preprint
submitted on 04.10.2019, 21:42 and posted on 09.10.2019, 20:02 by Maksym Druchok, Dzvenymyra Yarish, Oleksandr Gurbych, Mykola Maksymenko

Efficient design and screening of the novel molecules is a major challenge in drug and material design. This report focuses on a multi-stage pipeline in which several deep neural network (DNN) models are combined to map discrete molecular representations into continuous vector space to later generate from it new molecular structures with desired properties. Here the Attention-based Sequence-to-Sequence model is added to “spellcheck” and correct generated structures while the oversampling in the continuous space allows generating candidate structures with desired distribution for properties and molecular descriptors even for small reference datasets. We further use computer simulation to validate the desired properties in the numerical experiment. With the focus on the drug design, such pipeline allows generating novel structures with control of SAS (Synthetic Accessibility Score) and a series of ADME metrics that assess the drug-likeliness.

History

Email Address of Submitting Author

mmaks@softserveinc.com

Institution

SoftServe

Country

Ukraine

ORCID For Submitting Author

0000-0003-0685-5790

Declaration of Conflict of Interest

No conflicts of interest

Exports