An improved stacking model for predicting myocardial infarction risk in imbalanced data
摘要
Early diagnosis and treatment of myocardial infarction (MI) can significantly reduce the severity of the disease. Disease data are often imbalanced, which can lead to poor prediction outcomes when using conventional models. Therefore, developing a risk prediction model for MI with imbalanced datasets has become challenging. This paper presents a novel model called 2GDNN-FL-Stacked, which aims to address the issue of predicting the risk of MI in imbalanced data. Our group mitigates the impact of data imbalance on the model by employing random under-sampling and cost-sensitive techniques. We improve the model’s identification capabilities by stacking and combining 2GDNN-FL, CatBoost, RandomForest, and LightGBM. Our model’s Matthews Correlation Coefficient(MCC), F1-score, and Area Under the ROC Curve(AUC) scores increased by 0.87%