Predicting volume of distribution with decision tree-based regression methods using predicted tissue:plasma partition coefficients.
Predicting volume of distribution with decision tree-based regression methods using predicted tissue:plasma partition coefficients.
复制标题
使用预测的组织使用基于决策树的回归方法来预测分布的体积:等离子体分配系数。
DOI:
10.1186/s13321-015-0054-x
复制
发表时间:
2015
影响因子:
8.6
通讯作者:
Ghafourian T
中科院分区:
文献类型:
--
作者:
Freitas AA;Limbu K;Ghafourian T
Volume of distribution is an important pharmacokinetic property that indicates the extent of a drug’s distribution in the body tissues. This paper addresses the problem of how to estimate the apparent volume of distribution at steady state (Vss) of chemical compounds in the human body using decision tree-based regression methods from the area of data mining (or machine learning). Hence, the pros and cons of several different types of decision tree-based regression methods have been discussed. The regression methods predict Vss using, as predictive features, both the compounds’ molecular descriptors and the compounds’ tissue:plasma partition coefficients (Kt:p) – often used in physiologically-based pharmacokinetics. Therefore, this work has assessed whether the data mining-based prediction of Vss can be made more accurate by using as input not only the compounds’ molecular descriptors but also (a subset of) their predicted Kt:p values. Comparison of the models that used only molecular descriptors, in particular, the Bagging decision tree (mean fold error of 2.33), with those employing predicted Kt:p values in addition to the molecular descriptors, such as the Bagging decision tree using adipose Kt:p (mean fold error of 2.29), indicated that the use of predicted Kt:p values as descriptors may be beneficial for accurate prediction of Vss using decision trees if prior feature selection is applied. Decision tree based models presented in this work have an accuracy that is reasonable and similar to the accuracy of reported Vss inter-species extrapolations in the literature. The estimation of Vss for new compounds in drug discovery will benefit from methods that are able to integrate large and varied sources of data and flexible non-linear data mining methods such as decision trees, which can produce interpretable models. Decision trees for the prediction of tissue partition coefficient and volume of distribution of drugs. The online version of this article (doi:10.1186/s13321-015-0054-x) contains supplementary material, which is available to authorized users.
登录
查看更多内容
影响因子:
5.8
作者:
Ghafourian, Taravat;Barzegar-Jalali, Mohammad;Nokhodchi, Ali
通讯作者:
Nokhodchi, Ali
影响因子:
3.7
作者:
Gong, Yuping;Zhao, Zhiyang;Krise, Jeffrey P.
通讯作者:
Krise, Jeffrey P.
影响因子:
2.1
作者:
GASTEIGER, J;MARSILI, M
通讯作者:
MARSILI, M
影响因子:
3.5
作者:
Demir-Kavuk, Ozgur;Bentzien, Joerg;Knapp, Ernst-Walter
通讯作者:
Knapp, Ernst-Walter
影响因子:
3.3
作者:
Graham, Helen;Walker, Mike;Aarons, Leon
通讯作者:
Aarons, Leon