Predicting reaction conditions from limited data through active transfer learning

Shim, Eunjae; Kammeraad, Joshua A.; Xu, Ziping; Tewari, Ambuj; Cernak, Tim; Zimmerman, Paul M.

doi:10.1039/D1SC06932B

Citation Details

Predicting reaction conditions from limited data through active transfer learning

Transfer and active learning have the potential to accelerate the development of new chemical reactions, using prior data and new experiments to inform models that adapt to the target area of interest. This article shows how specifically tuned machine learning models, based on random forest classifiers, can expand the applicability of Pd-catalyzed cross-coupling reactions to types of nucleophiles unknown to the model. First, model transfer is shown to be effective when reaction mechanisms and substrates are closely related, even when models are trained on relatively small numbers of data points. Then, a model simplification scheme is tested and found to provide comparative predictivity on reactions of new nucleophiles that include unseen reagent combinations. Lastly, for a challenging target where model transfer only provides a modest benefit over random selection, an active transfer learning strategy is introduced to improve model predictions. Simple models, composed of a small number of decision trees with limited depths, are crucial for securing generalizability, interpretability, and performance of active transfer learning. more »

Award ID(s):: 2007055

PAR ID:: 10350296

Author(s) / Creator(s):: Shim, Eunjae; Kammeraad, Joshua A.; Xu, Ziping; Tewari, Ambuj; Cernak, Tim; Zimmerman, Paul M.

Date Published:: 2022-06-07

Journal Name:: Chemical Science

Volume:: 13

Issue:: 22

ISSN:: 2041-6520

Page Range / eLocation ID:: 6655 to 6668

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Journal Article:
https://doi.org/10.1039/D1SC06932B

More Like this