five

A MALDI-TOF mass spectrometry-based haemoglobin chain quantification method for rapid screen of thalassaemia

收藏
Figshare2022-01-31 更新2026-04-28 收录
下载链接:
https://figshare.com/articles/dataset/A_MALDI-TOF_mass_spectrometry-based_haemoglobin_chain_quantification_method_for_rapid_screen_of_thalassaemia/19096208
下载链接
链接失效反馈
官方服务:
资源简介:
Thalassaemia is one of the most common inherited monogenic diseases worldwide with a heavy global health burden. Considering its high prevalence in low and middle-income countries, a cheap, accurate and high-throughput screening test of thalassaemia prior to a more expensive confirmatory diagnostic test is urgently needed. In this study, we constructed a machine learning model based on MALDI-TOF mass spectrometry quantification of haemoglobin chains in blood, and for the first time, evaluated its diagnostic efficacy in 674 thalassaemia (including both asymptomatic carriers and symptomatic patients) and control samples collected in three hospitals. Parameters related to haemoglobin imbalance (α-globin, β-globin, γ-globin, α/β and α-β) were used for feature selection before classification model construction with 8 machine learning methods in cohort 1 and further model efficiency validation in cohort 2. The logistic regression model with 5 haemoglobin peak features achieved good classification performance in validation cohort 2 (AUC 0.99, 95% CI 0.98–1, sensitivity 98.7%, specificity 95.5%). Furthermore, the logistic regression model with 6 haemoglobin peak features was also constructed to specifically identify β-thalassaemia (AUC 0.94, 95% CI 0.91–0.97, sensitivity 96.5%, specificity 87.8% in validation cohort 2). For the first time, we constructed an inexpensive, accurate and high-throughput classification model based on MALDI-TOF mass spectrometry quantification of haemoglobin chains and demonstrated its great potential in rapid screening of thalassaemia in large populations.Key messagesThalassaemia is one of the most common inherited monogenic diseases worldwide with a heavy global health burden.We constructed a machine learning model based on MALDI-TOF mass spectrometry quantification of haemoglobin chains to screen for thalassaemia. Thalassaemia is one of the most common inherited monogenic diseases worldwide with a heavy global health burden. We constructed a machine learning model based on MALDI-TOF mass spectrometry quantification of haemoglobin chains to screen for thalassaemia.
创建时间:
2022-01-31
5,000+
优质数据集
54 个
任务类型
进入经典数据集
二维码
社区交流群

面向社区/商业的数据集话题

二维码
科研交流群

面向高校/科研机构的开源数据集话题

数据驱动未来

携手共赢发展

商业合作