All models are local: time to replace external validation with recurrent local validation

Youssef, Alex; Pencina, Michael; Thakur, Anshul; Zhu, Tingting; Clifton, David; Shah, Nigam H.

Computer Science > Machine Learning

arXiv:2305.03219 (cs)

[Submitted on 5 May 2023 (v1), last revised 13 May 2023 (this version, v2)]

Title:All models are local: time to replace external validation with recurrent local validation

Authors:Alex Youssef, Michael Pencina, Anshul Thakur, Tingting Zhu, David Clifton, Nigam H. Shah

View PDF

Abstract:External validation is often recommended to ensure the generalizability of ML models. However, it neither guarantees generalizability nor equates to a model's clinical usefulness (the ultimate goal of any clinical decision-support tool). External validation is misaligned with current healthcare ML needs. First, patient data changes across time, geography, and facilities. These changes create significant volatility in the performance of a single fixed model (especially for deep learning models, which dominate clinical ML). Second, newer ML techniques, current market forces, and updated regulatory frameworks are enabling frequent updating and monitoring of individual deployed model instances. We submit that external validation is insufficient to establish ML models' safety or utility. Proposals to fix the external validation paradigm do not go far enough. Continued reliance on it as the ultimate test is likely to lead us astray. We propose the MLOps-inspired paradigm of recurring local validation as an alternative that ensures the validity of models while protecting against performance-disruptive data variability. This paradigm relies on site-specific reliability tests before every deployment, followed by regular and recurrent checks throughout the life cycle of the deployed algorithm. Initial and recurrent reliability tests protect against performance-disruptive distribution shifts, and concept drifts that jeopardize patient safety.

Subjects:	Machine Learning (cs.LG); Methodology (stat.ME)
Cite as:	arXiv:2305.03219 [cs.LG]
	(or arXiv:2305.03219v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2305.03219

Submission history

From: Alexey Youssef [view email]
[v1] Fri, 5 May 2023 00:48:23 UTC (158 KB)
[v2] Sat, 13 May 2023 04:20:16 UTC (158 KB)

Computer Science > Machine Learning

Title:All models are local: time to replace external validation with recurrent local validation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:All models are local: time to replace external validation with recurrent local validation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators