{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/arxiv-2602-22747","title":"Set-based v.s. Distribution-based Representations of Epistemic Uncertainty: A Comparative Study","arxiv_id":"2602.22747","date":"2026-02-26","proceeding":null,"authors":["Kaizheng Wang","Yunjia Wang","Fabio Cuzzolin","David Moens","Hans Hallez","Siu Lun Chau"],"abstract":"Epistemic uncertainty in neural networks is commonly modeled using two second-order paradigms: distribution-based representations, which rely on posterior parameter distributions, and set-based representations based on credal sets (convex sets of probability distributions). These frameworks are often regarded as fundamentally non-comparable due to differing semantics, assumptions, and evaluation practices, leaving their relative merits unclear. Empirical comparisons are further confounded by variations in the underlying predictive models. To clarify this issue, we present a controlled comparative study enabling principled, like-for-like evaluation of the two paradigms. Both representations are constructed from the same finite collection of predictive distributions generated by a shared neural network, isolating representational effects from predictive accuracy. Our study evaluates each representation through the lens of 3 uncertainty measures across 8 benchmarks, including selective prediction and out-of-distribution detection, spanning 6 underlying predictive models and 10 independent runs per configuration. Our results show that meaningful comparison between these seemingly non-comparable frameworks is both feasible and informative, providing insights into how second-order representation choices impact practical uncertainty-aware performance.","url_abs":"https://arxiv.org/abs/2602.22747","url_pdf":"https://arxiv.org/pdf/2602.22747","source":{"archive":null,"snapshot":"2025-07-28","note":"not in the Papers with Code archive (frozen at the snapshot)","row_kind":"graph","title_abstract_authors_date":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)"},"code_links":[],"tasks":[],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[],"syntology":{"atlas_url":"https://app.syntology.ai/?focus=2602.22747","mcp":{"get_harvested_code_for_paper":{"arxiv_id":"2602.22747"}},"developers":"https://syntology.ai/developers","read_at":"2026-09-24T18:15:14+00:00","read_at_is":"when the build read Syntology's graph, not when any sample ran","claim":"Per-sample execution status on synthesized fixtures; not a correctness claim about the paper. Samples come from repositories linked to the paper, official or community; repo_kind says which.","repos":[{"mentioned_in_github":null,"is_official":null,"provenance":"deterministic:regex_extraction","mentioned_in_paper":null,"url":"https://github.com/DBO-DKFZ/uncertainty-benchmark","reach":null}],"summary":{"unverified":15},"by_repo_kind":{"found_in_text":{"samples":15,"ran":0,"repositories":1}},"repo_kind_vocabulary":{"official":"The archive marks this repository official for the paper","named_in_paper":"The archive records that the paper mentions this repository; it is not marked official","listed":"In the archive's code links for this paper, not marked official and not recorded as mentioned in the paper","found_in_text":"Syntology found this repository in the paper's own text; whether it is the authors' implementation is not asserted","community":"Not in the archive's code links for this paper; a community repository Syntology harvested"},"n_pointer_only_for_licence":15,"samples":[{"code_sha256_prefix":"02232fdaef3e4fd6","entry":"apply_loss_heads","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/losses/EnsembleLosses.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/losses/EnsembleLosses.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"02232fdaef3e4fd6"}},{"code_sha256_prefix":"ff83216de639a533","entry":"computeMeanConfidence","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/metrics/metrics.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/metrics/metrics.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"ff83216de639a533"}},{"code_sha256_prefix":"19b084e416e0aaa0","entry":"compute_confidence","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/metrics/metrics.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/metrics/metrics.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"19b084e416e0aaa0"}},{"code_sha256_prefix":"c7e1bbea778a593e","entry":"compute_temperature","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/pl_modules/temperature_scaling.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/pl_modules/temperature_scaling.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"c7e1bbea778a593e"}},{"code_sha256_prefix":"9cc3a958a9a9ee1f","entry":"compute_temperature_from_list","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/pl_modules/temperature_scaling.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/pl_modules/temperature_scaling.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"9cc3a958a9a9ee1f"}},{"code_sha256_prefix":"e6d235606098808b","entry":"create_df","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/eval/dataset_statistics.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/eval/dataset_statistics.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"e6d235606098808b"}},{"code_sha256_prefix":"68e62d7ae9aea425","entry":"create_risk_reject_curve","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/eval/plot_functions.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/eval/plot_functions.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"68e62d7ae9aea425"}},{"code_sha256_prefix":"e42c3f06c1e0cdb1","entry":"eincheck","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/utils/einop.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/utils/einop.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"e42c3f06c1e0cdb1"}},{"code_sha256_prefix":"08c6e2c32fbaafec","entry":"einop","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/utils/einop.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/utils/einop.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"08c6e2c32fbaafec"}},{"code_sha256_prefix":"0919a8421f2b0377","entry":"normed_entropy","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/metrics/metrics.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/metrics/metrics.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"0919a8421f2b0377"}},{"code_sha256_prefix":"0d8cf62d33feee0b","entry":"pairwise_acti_diversity","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/losses/EnsembleLosses.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/losses/EnsembleLosses.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"0d8cf62d33feee0b"}},{"code_sha256_prefix":"17553b4c7ed54bba","entry":"pairwise_weightCos","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/losses/EnsembleLosses.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/losses/EnsembleLosses.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"17553b4c7ed54bba"}},{"code_sha256_prefix":"f9de4cdd8a624801","entry":"rsetattr","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/pl_modules/temperature_scaling.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/pl_modules/temperature_scaling.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"f9de4cdd8a624801"}},{"code_sha256_prefix":"3508c2a20764f9bc","entry":"tile_extractor","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/eval/visualizations_base.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/eval/visualizations_base.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"3508c2a20764f9bc"}},{"code_sha256_prefix":"cb45a2e3e06e2831","entry":"transform_batch_outputs","repo":"DBO-DKFZ/uncertainty-benchmark","repo_kind":"found_in_text","path":"src/pl_modules/basemodules.py","file_url":"https://github.com/DBO-DKFZ/uncertainty-benchmark/blob/HEAD/src/pl_modules/basemodules.py","link_basis":"first_harvest_node","language":"python","status":"unverified","verification_level":0,"contract_check":null,"metamorphic_tier":null,"behaviour_fingerprint":false,"licence":"NONE","inline_ok":false,"mcp_get_code":{"code_sha256":"cb45a2e3e06e2831"}}]},"arxiv_metadata":{"licence":"arXiv metadata, CC0 1.0 (https://info.arxiv.org/help/license)","fields":["title","abstract","authors","date"],"primary_category":"cs.LG","source":"arxiv_2026.jsonl"},"syntology_extracted_results":null}