Pith. sign in

REVIEW 2 cited by

CLUE: Neural Networks Calibration via Learning Uncertainty-Error alignment

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2505.22803 v1 pith:UFBVDYBC submitted 2025-05-28 cs.LG cs.AI

classification cs.LGcs.AI
keywords calibrationclueuncertaintylossalignmentlearningnetworksneural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Reliable uncertainty estimation is critical for deploying neural networks (NNs) in real-world applications. While existing calibration techniques often rely on post-hoc adjustments or coarse-grained binning methods, they remain limited in scalability, differentiability, and generalization across domains. In this work, we introduce CLUE (Calibration via Learning Uncertainty-Error Alignment), a novel approach that explicitly aligns predicted uncertainty with observed error during training, grounded in the principle that well-calibrated models should produce uncertainty estimates that match their empirical loss. CLUE adopts a novel loss function that jointly optimizes predictive performance and calibration, using summary statistics of uncertainty and loss as proxies. The proposed method is fully differentiable, domain-agnostic, and compatible with standard training pipelines. Through extensive experiments on vision, regression, and language modeling tasks, including out-of-distribution and domain-shift scenarios, we demonstrate that CLUE achieves superior calibration quality and competitive predictive performance with respect to state-of-the-art approaches without imposing significant computational overhead.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. FALCON-Discover: Discovering Concentrated False-Confidence Regions for Calibration

    cs.LG 2026-06 conditional novelty 6.0 of 10

    FALCON-Discover ranks predictions by disagreement between confidence, local support, and perturbation stability, recovering much more high-confidence error mass than confidence ranking on several tabular datasets.

  2. Uncertainty Estimation by Human Perception versus Neural Models

    cs.LG 2025-06 conditional novelty 5.0 of 10

    Neural network uncertainty estimates correlate only weakly with human-perceived uncertainty on three vision benchmarks, and soft-label training improves that alignment, though the claimed calibration benefit is not me...

Pith tools