Papers › Uncertainty Calibration for Deep Audio Classifiers

Uncertainty Calibration for Deep Audio Classifiers

27 Jun 2022arXiv:2206.13071archive 2025-07-28

Tong Ye, Shijing Si, Jianzong Wang, Ning Cheng, Jing Xiao

Although deep Neural Networks (DNNs) have achieved tremendous success in audio classification tasks, their uncertainty calibration are still under-explored. A well-calibrated model should be accurate when it is certain about its prediction and indicate high uncertainty when it is likely to be inaccurate. In this work, we investigate the uncertainty calibration for deep audio classifiers. In particular, we empirically study the performance of popular calibration methods: (i) Monte Carlo Dropout, (ii) ensemble, (iii) focal loss, and (iv) spectral-normalized Gaussian process (SNGP), on audio classification datasets. To this end, we evaluate (i-iv) for the tasks of environment sound and music genre classification. Results indicate that uncalibrated deep audio classifiers may be over-confident, and SNGP performs the best and is very efficient on the two datasets of this paper.

PaperPDFCode

Code

shijing001/unicertainty_calibration_audio_classifiers officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Audio ClassificationClassificationGenre classificationMusic Genre Classification

Results from the paper archive 2025-07-28

No leaderboard rows for this paper in the archive.

Methods

DropoutGaussian ProcessMonte Carlo Dropout

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections