five

Half-life prediction of central nervous system (CNS) small molecules in humans using gradient tree boosting

收藏
Figshare2025-09-08 更新2026-04-28 收录
下载链接:
https://figshare.com/articles/dataset/Half-life_prediction_of_central_nervous_system_CNS_small_molecules_in_humans_using_gradient_tree_boosting/30073277
下载链接
链接失效反馈
官方服务:
资源简介:
To develop a machine learning (ML) model for early-stage prediction of human half-life of oral central nervous system (CNS) drugs and to establish a curated dataset, including key in vitro and in vivo data, to support future modeling efforts. Human and rat half-life, plasma protein binding (PPB), and liver microsomal clearance (LM) data for 76 diverse CNS drugs and candidates were obtained from public sources or evaluated at WuXi AppTec. Gradient tree boosting (GTB) models were constructed using ChemAxon’s Trainer Engine. Feature importance was assessed, and model performance was evaluated on an external validation set. The best-performing model achieved 82.4% of predictions within two-fold of observed values, with a coefficient of determination (R2) of 0.75 and a root mean square error (RMSE) of 0.25. Good generalizability was confirmed using similarity-based data splitting and Y-randomization. Integration of in vitro features, preclinical in vivo data, and physicochemical properties substantially improved predictive performance. Key features driving accurate human half-life prediction were identified. This model demonstrates practical applications for early-stage prediction of human half-life and prioritization of CNS drug candidates. The curated dataset offers a valuable resource to enhance internal databases and advance more robust predictive models.
创建时间:
2025-09-08
5,000+
优质数据集
54 个
任务类型
进入经典数据集
二维码
社区交流群

面向社区/商业的数据集话题

二维码
科研交流群

面向高校/科研机构的开源数据集话题

数据驱动未来

携手共赢发展

商业合作