Materials discovery with extreme properties via reinforcement learning-guided combinatorial chemistry

Hyunseung Kim; Haeyeon Choi; Dongju Kang; Won Bo Lee; Jonggeol Na

doi:10.1039/D3SC05281H

Materials discovery with extreme properties via reinforcement learning-guided combinatorial chemistry†

Hyunseung Kim,

‡^a Haeyeon Choi,‡^bc Dongju Kang,‡^a Won Bo Lee

*^a and Jonggeol Na

*^bc

Author affiliations

* Corresponding authors

^a School of Chemical and Biological Engineering, Seoul National University, Republic of Korea
E-mail: wblee@snu.ac.kr

^b Department of Chemical Engineering and Materials Science, Ewha Womans University, Republic of Korea
E-mail: jgna@ewha.ac.kr

^c Graduate Program in System Health Science and Engineering, Ewha Womans University, Republic of Korea

Abstract

The goal of most materials discovery is to discover materials that are superior to those currently known. Fundamentally, this is close to extrapolation, which is a weak point for most machine learning models that learn the probability distribution of data. Herein, we develop reinforcement learning-guided combinatorial chemistry, which is a rule-based molecular designer driven by trained policy for selecting subsequent molecular fragments to get a target molecule. Since our model has the potential to generate all possible molecular structures that can be obtained from combinations of molecular fragments, unknown molecules with superior properties can be discovered. We theoretically and empirically demonstrate that our model is more suitable for discovering better compounds than probability distribution-learning models. In an experiment aimed at discovering molecules that hit seven extreme target properties, our model discovered 1315 of all target-hitting molecules and 7629 of five target-hitting molecules out of 100 000 trials, whereas the probability distribution-learning models failed. Moreover, it has been confirmed that every molecule generated under the binding rules of molecular fragments is 100% chemically valid. To illustrate the performance in actual problems, we also demonstrate that our models work well on two practical applications: discovering protein docking molecules and HIV inhibitors.

This article is part of the themed collection: 2024 Chemical Science Covers

Supplementary files

Article information

DOI: https://doi.org/10.1039/D3SC05281H
Article type: Edge Article
Submitted: 05 Oct 2023
Accepted: 23 Apr 2024
First published: 24 Apr 2024
This article is Open Access

All publication charges for this article have been paid for by the Royal Society of Chemistry

Download Citation

Chem. Sci., 2024,15, 7908-7925

Permissions

Request permissions

Materials discovery with extreme properties via reinforcement learning-guided combinatorial chemistry

H. Kim, H. Choi, D. Kang, W. B. Lee and J. Na, Chem. Sci., 2024, 15, 7908 DOI: 10.1039/D3SC05281H

This article is licensed under a Creative Commons Attribution-NonCommercial 3.0 Unported Licence. You can use material from this article in other publications, without requesting further permission from the RSC, provided that the correct acknowledgement is given and it is not used for commercial purposes.

To request permission to reproduce material from this article in a commercial publication, please go to the Copyright Clearance Center request page.

If you are an author contributing to an RSC publication, you do not need to request permission provided correct acknowledgement is given.

If you are the author of this article, you do not need to request permission to reproduce figures and diagrams provided correct acknowledgement is given. If you want to reproduce the whole article in a third-party commercial publication (excluding your thesis/dissertation for which permission is not required) please go to the Copyright Clearance Center request page.

Chemical Science

Materials discovery with extreme properties via reinforcement learning-guided combinatorial chemistry†

Abstract

Supplementary files

Article information

Download Citation

Permissions

Materials discovery with extreme properties via reinforcement learning-guided combinatorial chemistry

Social activity

Search articles by author

Spotlight

Advertisements