Operational Calibration: Debugging Confidence Errors for DNNs in the Field

Li, Zenan; Ma, Xiaoxing; Xu, Chang; Xu, Jingwei; Cao, Chun; Lü, Jian

doi:10.1145/3368089.3409696

Computer Science > Machine Learning

arXiv:1910.02352 (cs)

[Submitted on 6 Oct 2019 (v1), last revised 13 Sep 2020 (this version, v2)]

Title:Operational Calibration: Debugging Confidence Errors for DNNs in the Field

Authors:Zenan Li, Xiaoxing Ma, Chang Xu, Jingwei Xu, Chun Cao, Jian Lü

View PDF

Abstract:Trained DNN models are increasingly adopted as integral parts of software systems, but they often perform deficiently in the field. A particularly damaging problem is that DNN models often give false predictions with high confidence, due to the unavoidable slight divergences between operation data and training data. To minimize the loss caused by inaccurate confidence, operational calibration, i.e., calibrating the confidence function of a DNN classifier against its operation domain, becomes a necessary debugging step in the engineering of the whole system.
Operational calibration is difficult considering the limited budget of labeling operation data and the weak interpretability of DNN models. We propose a Bayesian approach to operational calibration that gradually corrects the confidence given by the model under calibration with a small number of labeled operation data deliberately selected from a larger set of unlabeled operation data. The approach is made effective and efficient by leveraging the locality of the learned representation of the DNN model and modeling the calibration as Gaussian Process Regression. Comprehensive experiments with various practical datasets and DNN models show that it significantly outperformed alternative methods, and in some difficult tasks it eliminated about 71% to 97% high-confidence (>0.9) errors with only about 10\% of the minimal amount of labeled operation data needed for practical learning techniques to barely work.

Comments:	Published in the Proceedings of the 28th ACM Joint European Software Engineering Conference and Symposium on the Foundations of Software Engineering (ESEC/FSE 2020)
Subjects:	Machine Learning (cs.LG); Software Engineering (cs.SE); Machine Learning (stat.ML)
Cite as:	arXiv:1910.02352 [cs.LG]
	(or arXiv:1910.02352v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1910.02352
Related DOI:	https://doi.org/10.1145/3368089.3409696

Submission history

From: Zenan Li [view email]
[v1] Sun, 6 Oct 2019 01:21:14 UTC (881 KB)
[v2] Sun, 13 Sep 2020 02:28:32 UTC (21,440 KB)

Computer Science > Machine Learning

Title:Operational Calibration: Debugging Confidence Errors for DNNs in the Field

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Operational Calibration: Debugging Confidence Errors for DNNs in the Field

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators