Dataset
303 records and 14 original variables from the Cleveland subset of UCI Heart Disease.
PUBLISHED ON GITHUB
An exploratory analysis of demographic and clinical variables associated with heart disease presence. The published repository documents data cleaning, visualizations and analytical limitations.
303 records and 14 original variables from the Cleveland subset of UCI Heart Disease.
Missing-value markers were normalized. The ca and thal variables were converted to numeric types and missing values imputed using the mode. Duplicate checks found no duplicates.
Pandas and Matplotlib were used to examine sex, age, chest pain type and maximum heart rate. No predictive machine learning model was developed.
The repository reports heart disease presence in approximately 55.3% of male patients and 25.8% of female patients in this sample. These are descriptive associations within the dataset.