Previously submitted to: Journal of Medical Internet Research (no longer under consideration since Jul 24, 2026)
Date Submitted: Apr 25, 2026
Open Peer Review Period: Apr 25, 2026 - Jun 20, 2026
(closed for review but you can still tweet)
NOTE: This is an unreviewed Preprint
Warning: This is a unreviewed preprint (What is a preprint?). Readers are warned that the document has not been peer-reviewed by expert/patient reviewers or an academic editor, may contain misleading claims, and is likely to undergo changes before final publication, if accepted, or may have been rejected/withdrawn (a note "no longer under consideration" will appear above).
Peer review me: Readers with interest and expertise are encouraged to sign up as peer-reviewer, if the paper is within an open peer-review period (in this case, a "Peer Review Me" button to sign up as reviewer is displayed above). All preprints currently open for review are listed here. Outside of the formal open peer-review period we encourage you to tweet about the preprint.
Citation: Please cite this preprint only for review purposes or for grant applications and CVs (if you are the author).
Final version: If our system detects a final peer-reviewed "version of record" (VoR) published in any journal, a link to that VoR will appear below. Readers are then encourage to cite the VoR instead of this preprint.
Settings: If you are the author, you can login and change the preprint display settings, but the preprint URL/DOI is supposed to be stable and citable, so it should not be removed once posted.
Submit: To post your own preprint, simply submit to any JMIR journal, and choose the appropriate settings to expose your submitted version as preprint.
Warning: This is an author submission that is not peer-reviewed or edited. Preprints - unless they show as "accepted" - should not be relied on to guide clinical practice or health-related behavior and should not be reported in news media as established information.
Deep Learning Algorithms for Predicting Intraoperative Hypotension: A Systematic Review and Meta-Analysis
ABSTRACT
Background:
Intraoperative hypotension (IOH) is associated with myocardial injury, acute kidney injury, perioperative stroke, and 30-day mortality, yet conventional blood pressure monitoring remains reactive rather than anticipatory. Deep learning (DL) algorithms applied to continuous physiological waveforms represent a rapidly expanding paradigm for early IOH prediction, but the comparative performance of distinct DL architectures and the influence of prediction-window length, input data modality, IOH reference standard, and analysis unit on diagnostic accuracy have not been systematically synthesised.
Objective:
To quantify the pooled diagnostic accuracy of DL-based IOH prediction models and to identify methodological and clinical factors that modify their performance.
Methods:
PubMed, Embase, Web of Science, and the Cochrane Library were searched through March 2026. Methodological quality was appraised with the PROBAST+AI tool and overall certainty of evidence with the GRADE framework. A bivariate random-effects model generated pooled sensitivity, specificity, and the area under the summary receiver operating characteristic (SROC) curve, with heterogeneity quantified by τ²(Se), τ²(Sp), and the inter-study correlation ρ. Threshold effect was tested with Spearman’s correlation, publication bias with Deeks’ test, and clinical utility with Fagan’s nomogram. Prespecified subgroup analyses (prediction window, DL architecture, input modality, IOH reference standard, analysis unit) and Bayesian random-effects meta-regression explored heterogeneity sources.
Results:
Twelve studies were included; nine contributed 22 validation datasets to the quantitative synthesis. The pooled sensitivity was 0.78 (95% CI 0.73–0.81), specificity 0.88 (0.82–0.92), and SROC-AUC 0.87 (0.83–0.90); the diagnostic odds ratio was 24.7 (16.1–37.9), positive likelihood ratio 6.31, and negative likelihood ratio 0.26. Heterogeneity was τ²(Se) = 0.25, τ²(Sp) = 1.04, and ρ = −0.28; no significant threshold effect was detected (Spearman ρ = 0.29, P = 0.20). The 5-minute window achieved the highest performance (sensitivity 0.81, 95% CI 0.77–0.85; specificity 0.91, 0.84–0.95). Meta-regression identified DL architecture as the only significant moderator of specificity (P = 0.02), with hybrid CNN-RNN exceeding pure CNN (β = 1.77, 95% CI 0.45–3.09); no covariate significantly moderated sensitivity. Deeks’ test showed no statistically significant publication bias (P = 0.06). At a 10% pre-test probability, post-test probabilities were 41% (positive) and 3% (negative). GRADE certainty was Low.
Conclusions:
Deep learning models for IOH prediction achieve moderate diagnostic accuracy, with hybrid CNN-RNN architectures and 5-minute prediction windows showing the most favourable performance. The universal absence of formal calibration assessment, scarce external validation, and geographic concentration of the evidence base constrain immediate clinical translation. Prospective multinational validation with mandatory calibration reporting and patient-level evaluation is required before DL-based IOH alerts can be safely integrated into perioperative decision support. Clinical Trial: PROSPERO CRD420261377604.
Citation
Request queued. Please wait while the file is being generated. It may take some time.
Copyright
© The authors. All rights reserved. This is a privileged document currently under peer-review/community review (or an accepted/rejected manuscript). Authors have provided JMIR Publications with an exclusive license to publish this preprint on it's website for review and ahead-of-print citation purposes only. While the final peer-reviewed paper may be licensed under a cc-by license on publication, at this stage authors and publisher expressively prohibit redistribution of this draft paper other than for review purposes.