遇见数据集

Clinical Text Imbalance Benchmark—Results (336 configurations), v2.

收藏
Zenodo2025-10-10 更新2026-05-26 收录
官方服务:

资源简介:

This dataset contains per-configuration test metrics for a large-scale benchmark of clinical text classification under extreme class imbalance (49,035 French breast radiology reports; minority prevalence ≈0.33%). A factorial design varied two vectorisers (BoW, TF–IDF), 12 resampling methods plus a baseline, and 15 classifiers. The file ml_experiment_results.csv reports one row per executed configuration (n=336) with: Vectorizer, Sampler, Classifier, Accuracy, Balanced_Accuracy, ROC_AUC, PR_AUC, Precision_male, Recall_male, F1_male, Precision_female, Recall_female, F1_female, F1_macro, F1_weighted, TP, FP, TN, FN.

提供机构:
Zenodo
创建时间:
2025-08-24
二维码
社区交流群
二维码
科研交流群
商业服务