What Is Missing in Surgical Risk Stratification and Outcome Prediction: A Scoping Review of End-to-End Machine Learning Approaches

📅 2026-07-31
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
Current machine learning research in surgical risk prediction is often hindered by methodological fragmentation, poor reproducibility, and limited clinical applicability. This study conducts a scoping review of 190 end-to-end machine learning pipelines based on electronic health records, systematically examining critical components including data preprocessing, model selection, evaluation strategies, and interpretability. It presents the first structured synthesis of the entire workflow for surgical risk stratification, uncovering systemic gaps in the use of open datasets, standardized evaluation benchmarks, and deep learning methodologies. The majority of studies rely on single-center proprietary data, and only about one-third incorporate interpretability techniques—factors that severely constrain model generalizability and clinical translation. This work establishes a methodological framework and practical guidance for developing reproducible, generalizable, and clinically viable surgical prediction models.
📝 Abstract
Postoperative adverse events, including mortality and morbidity, remain a major global burden, many of which are preventable through early identification of high-risk patients and targeted perioperative care. Accurate risk stratification is therefore essential. With the growing availability of large-scale electronic health records (EHRs), machine learning (ML) provides a data-driven approach to model complex clinical patterns. However, existing studies vary widely in design, and methodological practices remain fragmented. This scoping review characterizes ML pipelines for surgical risk stratification and outcome prediction using EHR data. We reviewed 190 studies covering the ML workflow, including data preprocessing, algorithm selection, model evaluation, and explainability. Most studies relied on single-center private datasets with limited data modalities, while the scarcity of open-access surgical datasets constrained reproducibility and generalizability. Reporting of key preprocessing steps, including missing data handling, feature selection, and class imbalance, was often incomplete. Conventional ML models and simple neural networks predominated, whereas deep learning and multimodal approaches remained uncommon. Benchmark datasets and standardized evaluation protocols were largely absent, hindering cross-study comparisons. Only about one-third of studies incorporated explainability methods. This review identifies methodological gaps limiting clinically robust postoperative ML tools and provides a structured reference to support more rigorous, reproducible, and clinically meaningful ML development for perioperative care.
Problem

Research questions and friction points this paper is trying to address.

surgical risk stratification
outcome prediction
machine learning
electronic health records
methodological gaps
Innovation

Methods, ideas, or system contributions that make the work stand out.

machine learning
surgical risk stratification
electronic health records
model explainability
reproducibility
🔎 Similar Papers
2024-04-17Annual International Conference of the IEEE Engineering in Medicine and Biology SocietyCitations: 0
Y
Yizhi Dong
Saw Swee Hock School of Public Health, National University of Singapore, Singapore 117549
Y
Yuhe Ke
Data Science and Artificial Intelligence Lab, Singapore General Hospital, Singapore 169608; Duke-NUS Medical School, Singapore 169857
H
Hairil Rizal Abdullah
Department of Anaesthesiology, Singapore General Hospital, Singapore 169608; Data Science and Artificial Intelligence Lab, Singapore General Hospital, Singapore 169608; Duke-NUS Medical School, Singapore 169857
Y
Yucheng Xing
Saw Swee Hock School of Public Health, National University of Singapore, Singapore 117549; National University of Singapore Guangzhou Research Translation and Innovation Institute, Guangzhou 510700, China
K
Kevan Kai Bing Teo
School of Science, Loughborough University, Loughborough LE11 3TU, UK
Ling Huang
Ling Huang
Imperial, NUS, UTC, CNRS, Sorbonne alliance
Uncertainty quantificationTrustworthy AIMedical data analysisCardiovascular computing
M
Mengling Feng
Saw Swee Hock School of Public Health, National University of Singapore, Singapore 117549; Institute of Data Science, National University of Singapore, Singapore 117602