PofoliaShared via Pofolia

PLoS ONE· 2026Q1

Machine learning and deep learning–based prediction of hypertension and analysis of its major risk factors in Bangladesh

Shawrab Chandra, M. M. Imran Molla, Samiul Islam, Md. Matiur Rahaman et al.

Short summary

Random Forest (RF) model achieved the highest recall (68.7%) and F1-score (0.460) for predicting hypertension in Bangladesh, outperforming other ML/DL models in identifying affected individuals, despite Weighted Logistic Regression showing higher overall accuracy (81.7%).

AI-generated from the title and abstract; the full text is not read.

Abstract

Background Hypertension is a leading cause of cardiovascular morbidity and mortality in Bangladesh. This study examined its prevalence, risk factors, and predictive modeling using machine learning (ML) and deep learning (DL) approaches. Method We analyzed cross-sectional data from the 2022 Bangladesh Demographic and Health Survey, which included 14,283 adults (≥18 years). Prevalence was estimated, chi-square tests assessed associations, and four ML models (weighted logistic regression, random forest, extreme gradient boosting, light gradient boosting machine) and two DL models (TabNet, and multi-layer perceptron) were applied to predict hypertension risk. Model performance was evaluated using accuracy, precision, recall, specificity, F1 score, and area under the receiver operating characteristics curve and precision-recall curve. Results Overall prevalence was 18.04% (95% CI: 17.2%–18.9%), higher among women (18.87%) than men (16.97%). The chi-square test suggests that hypertension was significantly associated with age, BMI, diabetes, wealth index, education, household size, and region (p < 0.05). Among the machine learning and deep learning models, weighted logistic regression (WLR) achieved the highest accuracy (0.817), precision (0.444), specificity (0.981), AUC-ROC (0.751), and AUC-PR (0.357). However, WLR exhibited low recall (0.070). In contrast, the random forest (RF) model achieved the highest recall (0.687) and F1-score (0.460) on the test data, indicating greater sensitivity in identifying individuals with hypertension. Additionally, age, BMI, sex, family size, and educational level were identified as the most important predictors among the variables included in the study. Conclusion Hypertension is common in Bangladesh, with higher prevalence in women and significant association with socio-demographic determinants. Although WLR demonstrated the highest accuracy, precision, specificity, and AUC-PR, its low recall limits its utility for identifying individuals with hypertension. RF may be more suitable for public health applications because of its higher recall and F1-score; however, further external validation and assessment of its clinical utility are required before implementation.

The authors' abstract, as published at the source. PLoS ONE, 2026 · DOI ↗

TakeawaysIn the app
Key pointsIn the app
Ask the paperIn the app

The rest is in the Pofolia app

Takeaways, key points and questions to the paper; new summaries every day for your field. Free.

Sign in on the web to open

Field: Cardiology and Cardiovascular Medicine

Cardiology and Cardiovascular MedicineMedicine