Supplementary Materials and Datasets for SNP Selection Using Pearson Correlation for Human Population Classification
收藏资源简介:
Overview. This record provides the supplementary datasets and materials referenced in the manuscript. Files are organized as ZIPs per subsection of Section 4, with a root-level README.md, a MANIFEST.tsv mapping each file to its section, and a DATA_DICTIONARY.tsv describing all tabular columns.Contents. 0_* packages contain inputs used in the study (pre-/post-QC genotype matrices, block assignments, correlation ranges). 1_*–6_* packages mirror §§4.1–4.6 (Dataset A/B tables & panels, significance tests, PCA-120 vs Pearson-120 metrics, runtime traces, and consensus SNP panels). Each ZIP includes a short mini-README.Reproducibility. Per-iteration results (1,000 MCCV runs) and aggregated summaries are provided; timing traces are in milliseconds; confusion matrices are row-normalized (%). PCA is fitted on training folds only; Pearson-120 uses absolute point-biserial correlation with the label.How to cite. Please cite this record as the repository [55] in the manuscript and via the DOI provided by Zenodo. Citation (IEEE):P. N. Basuki, J.P. Sri Yulianto, A. Setiawan, "Supplementary materials and datasets for SNP Selection Using Pearson Correlation for Human Population Classification". Dataset. Zenodo, v2.0.0, 2025. doi:10.5281/zenodo.17271719. v2.0.0 — Repository reorganization- Removed: legacy single-file consensus panel.- Added: 11 ZIP packages organized by Section 4 (0_* inputs; 1–6 materials).- Added: root-level README.md, MANIFEST.tsv, DATA_DICTIONARY.tsv.- No changes to analyses or conclusions; repository layout only.



