Skip to content
Review

What Is Missing in Surgical Risk Stratification and Outcome Prediction: A Scoping Review of End-to-End Machine Learning Approaches

Jul 2026 · arXiv.org · Vol abs/2607.29090 · 0 citations · 106 references
Computer Science

TL;DR

This scoping review characterizes ML pipelines for surgical risk stratification and outcome prediction using EHR data and identifies methodological gaps limiting clinically robust postoperative ML tools and provides a structured reference to support more rigorous, reproducible, and clinically meaningful ML development for perioperative care.

Abstract

Postoperative adverse events, including mortality and morbidity, remain a major global burden, many of which are preventable through early identification of high-risk patients and targeted perioperative care. Accurate risk stratification is therefore essential. With the growing availability of large-scale electronic health records (EHRs), machine learning (ML) provides a data-driven approach to model complex clinical patterns. However, existing studies vary widely in design, and methodological practices remain fragmented. This scoping review characterizes ML pipelines for surgical risk stratification and outcome prediction using EHR data. We reviewed 190 studies covering the ML workflow, including data preprocessing, algorithm selection, model evaluation, and explainability. Most studies relied on single-center private datasets with limited data modalities, while the scarcity of open-access surgical datasets constrained reproducibility and generalizability. Reporting of key preprocessing steps, including missing data handling, feature selection, and class imbalance, was often incomplete. Conventional ML models and simple neural networks predominated, whereas deep learning and multimodal approaches remained uncommon. Benchmark datasets and standardized evaluation protocols were largely absent, hindering cross-study comparisons. Only about one-third of studies incorporated explainability methods. This review identifies methodological gaps limiting clinically robust postoperative ML tools and provides a structured reference to support more rigorous, reproducible, and clinically meaningful ML development for perioperative care.

View source

Similar papers

Review Aug 2026

P1.139. Advancing Risk Prediction After Esophagectomy: A Systematic Review of Machine Learning Models

Machine learning models for postoperative risk prediction after esophagectomy demonstrate promising discrimination, particularly for anastomotic leak, and has the potential to enhance individualized risk stratification, support shared decision-making, guide perioperative planning, and improve allocation of postoperativ...

T. Wang, Otari Beldishevski-Shotadze, N. Evennett · 0 citations
Review Aug 2026

Machine learning-based prediction of cardiovascular adverse events in patients with cancer: a systematic review.

AI/ML models show promise for predicting CV adverse events in patients with cancer; however, clinical applicability is constrained by insufficient preprocessing transparency, limited external validation, and inadequate calibration reporting.

Li-Wei Wu, Minh-Anh Le-Dang, B. Okoye et al. · 0 citations
Open access Aug 2026

Predictive Modeling of 30-Day Readmission Risk: A Machine Learning Approach for Health Management

Aim: This study aims to develop and evaluate predictive models capable of identifying patients at risk of 30-day readmission using structured inpatient data.Material and Method: The analysis was conducted on a fully synthetic dataset designed to reflect the complexity of real-world clinical data while ensuring the prot...

Alican Doğan · 0 citations
#generative ai Review Open access Oct 2026

A Critical Methodological Review of Clinical Scores, Machine Learning, and Large Language Models for the Diagnosis and Management of Pediatric Appendicitis

Background/Objectives: Acute appendicitis is the most common surgical emergency of childhood, yet its diagnosis remains difficult because presentations are atypical, inflammatory markers are nonspecific, and the consequences of error run in both directions, from negative appendectomy to missed perforation. Over the pas...

M. Bašković, Z. Pogorelić · 0 citations
Jul 2026

Are Limited Electronic Medical Record Follow-Up Data Sufficiently Useful for Validating the Performance of Survival Prediction Models?

PURPOSE Machine learning models that predict survival time for patients are increasingly used for clinical decision support. Validating model performance in deployment is important but challenging because the only timely source of follow-up/death data is the electronic medical record (EMR), which is known to undercaptu...

Vigneshwari Subramanian, Tyler Raclin, B. Narasimhan et al. · 0 citations
Review Open access Aug 2026

Machine Learning on the Pediatric Intensive Care (PIC) Database: PIC-Powered Prediction

This narrative review synthesizes ten PIC-based ML studies identified through forward citation tracking on the original PIC publication and PhysioNet dataset record in PubMed, Web of Science, and Google Scholar and concludes that PIC should be viewed as the seed for a collaborative pediatric ICU data ecosystem.

H. Ganatra, Daniah Shamim, Shawn B. Sood et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.