MultiDK: A Multiple Descriptor Multiple Kernel Approach for Molecular Discovery and Its Application to Organic Flow Battery Electrolytes.

Sungjin Kim,Adrián Jinich,Alán Aspuru-Guzik

doi:10.1021/acs.jcim.6b00332

Abstract

We propose a multiple descriptor multiple kernel (MultiDK) method for efficient molecular discovery using machine learning. We show that the MultiDK method improves both the speed and accuracy of molecular property prediction. We apply the method to the discovery of electrolyte molecules for aqueous redox flow batteries. Using multiple-type-as opposed to single-type-descriptors, we obtain more relevant features for machine learning. Following the principle of "wisdom of the crowds", the combination of multiple-type descriptors significantly boosts prediction performance. Moreover, by employing multiple kernels-more than one kernel function for a set of the input descriptors-MultiDK exploits nonlinear relations between molecular structure and properties better than a linear regression approach. The multiple kernels consist of a Tanimoto similarity kernel and a linear kernel for a set of binary descriptors and a set of nonbinary descriptors, respectively. Using MultiDK, we achieve an average performance of r2 = 0.92 with a test set of molecules for solubility prediction. We also extend MultiDK to predict pH-dependent solubility and apply it to a set of quinone molecules with different ionizable functional groups to assess their performance as flow battery electrolytes.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

MultiDK: A Multiple Descriptor Multiple Kernel Approach for Molecular Discovery and Its Application to Organic Flow Battery Electrolytes.

Abstract

Talk to us

Similar Papers

More From: Journal of Chemical Information and Modeling

Lead the way for us

Journal: Journal of Chemical Information and Modeling	Publication Date: Apr 10, 2017
Citations: 27

Similar Papers

Use of DFT to Evaluate Properties of Viologen Derivatives for Redox Flow Batteries
Alizée Debiais ... Hélène Lebel
Electrochemical Society Meeting Abstracts | VOL. MA2024-01
Alizée Debiais, et. al.Alizée Debiais ... Hélène Lebel
09 Aug 2024
Electrochemical Society Meeting Abstracts | VOL. MA2024-01

Synthesis of Electrolytes in Flow Batteries
Yan Jing ... Roy G Gordon
Electrochemical Society Meeting Abstracts | VOL. MA2020-02
Yan Jing, et. al.Yan Jing ... Roy G Gordon
23 Nov 2020
Electrochemical Society Meeting Abstracts | VOL. MA2020-02

Synthesis-Driven Computational Discovery of Organic Redoxmers for Non-Aqueous Redox Flow Batteries
Akash Jain ... Rajeev S Assary
Electrochemical Society Meeting Abstracts | VOL. MA2023-01
Akash Jain, et. al.Akash Jain ... Rajeev S Assary
28 Aug 2023
Electrochemical Society Meeting Abstracts | VOL. MA2023-01

Chemoinformatics in the United Kingdom.
Nathan Brown
Molecular informatics | VOL. 34
Nathan BrownNathan Brown
26 Aug 2015
Molecular informatics | VOL. 34

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

MultiDK: A Multiple Descriptor Multiple Kernel Approach for Molecular Discovery and Its Application to Organic Flow Battery Electrolytes.

Abstract

Talk to us

Similar Papers

More From: Journal of Chemical Information and Modeling