Cardiovascular disease remains the leading global cause of death, emphasizing the need for improved risk stratification beyond traditional tools such as Framingham, ASCVD, QRISK, and SCORE, which show limitations in diverse modern populations. Machine learning methods applied to electronic health records can enhance prediction by capturing complex, high-dimensional, and nonlinear relationships. This systematic review (2017–2022) evaluated machine learning models for cardiovascular risk prediction using EHR data, focusing on discrimination (AUROC, AUPRC), calibration, external validation, and reporting quality including TRIPOD adherence. A PRISMA-compliant search identified peer-reviewed studies applying machine learning to EHR-based cardiovascular risk prediction. Risk of bias was assessed using PROBAST, and narrative synthesis was conducted due to heterogeneity. Twenty-nine studies were included. XGBoost, random forest, and neural networks were the most common models and generally outperformed logistic regression and traditional risk scores in discrimination. However, calibration was infrequently reported, and external validation was limited, often showing reduced performance. Machine learning models demonstrate improved predictive discrimination over conventional risk scores, but limited calibration assessment and weak external validation constrain clinical applicability. Stronger validation frameworks are needed for clinical translation.
Suicidality and depression are major global health burdens, with over 700,000 suicide deaths annually and ~280 million people affected by major depressive disorder. Early risk prediction could support prevention, but traditional methods show limited accuracy. This PRISMA-compliant systematic review evaluated machine learning models for predicting suicidality and depression across electronic health records, social media, and wearable sensor data, focusing on performance, unimodal vs multimodal approaches, and ethical reporting. Searches of PubMed, PsycINFO, IEEE Xplore, arXiv, and ACM Digital Library identified eligible studies. EHR-based models showed AUROC 0.70–0.85 for suicide attempt prediction, social media models 0.70–0.80 for suicidal ideation, and wearable sensor models lower performance (0.65–0.75). Multimodal approaches improved performance by 5–10% over unimodal models. However, fewer than 20% of studies reported ethical considerations such as privacy, bias, or deployment safeguards. Overall, machine learning shows moderate-to-good predictive performance, with multimodal models performing best, but ethical reporting remains critically insufficient for clinical translation.