Stroke data for patients living with coronary heart disease
收藏资源简介:
This dataset contains individual-level health, behavioral, and demographic variables used to develop and evaluate machine learning models for predicting stroke risk among adults with and without coronary heart disease (CHD). The data consist of 22 variables, including clinical indicators (high blood pressure, high cholesterol, diabetes, BMI), lifestyle behaviors (smoking, physical activity, alcohol consumption, fruit and vegetable intake), access to healthcare, self-reported health status, functional limitations, mental and physical health days, as well as demographic factors (sex, age, education, income). The dataset includes both individuals who have experienced a stroke (Stroke = 1) and those without stroke (Stroke = 0), enabling development of supervised classification models. All variables are encoded numerically to support statistical modelling and machine learning. No personally identifiable information is included. This dataset was prepared as part of a study on explainable machine learning for stroke risk prediction and can be used for benchmarking, algorithm comparison, reproducibility studies, and model interpretability research.



