REVIEW 23 cited by
OpenOOD v1.5: Enhanced Benchmark for Out-of-Distribution Detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
OpenOOD v1.5: Enhanced Benchmark for Out-of-Distribution Detection
read the original abstract
Out-of-Distribution (OOD) detection is critical for the reliable operation of open-world intelligent systems. Despite the emergence of an increasing number of OOD detection methods, the evaluation inconsistencies present challenges for tracking the progress in this field. OpenOOD v1 initiated the unification of the OOD detection evaluation but faced limitations in scalability and scope. In response, this paper presents OpenOOD v1.5, a significant improvement from its predecessor that ensures accurate and standardized evaluation of OOD detection methodologies at large scale. Notably, OpenOOD v1.5 extends its evaluation capabilities to large-scale data sets (ImageNet) and foundation models (e.g., CLIP and DINOv2), and expands its scope to investigate full-spectrum OOD detection which considers semantic and covariate distribution shifts at the same time. This work also contributes in-depth analysis and insights derived from comprehensive experimental results, thereby enriching the knowledge pool of OOD detection methodologies. With these enhancements, OpenOOD v1.5 aims to drive advancements and offer a more robust and comprehensive evaluation benchmark for OOD detection research.
Forward citations
Cited by 23 Pith papers
-
FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics
FML-Bench shows a simple greedy hill-climber nearly matches tree search on dense-opportunity tasks while an adaptive agent that broadens search on stagnation outperforms six baselines across 18 tasks.
-
Mitigating Simplicity Bias in OOD Detection through Object Co-occurrence Analysis
OCO mitigates simplicity bias in OOD detection by predicting disentangled representations, dividing patterns into three co-occurrence scenarios from ID data, and applying divide-and-conquer detection.
-
Uncertainty Estimation via Hyperspherical Confidence Mapping
HCM turns neural outputs into magnitude plus unit hypersphere vector and treats uncertainty as the geometric violation of that unit constraint, yielding deterministic estimates for regression and classification that m...
-
Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection
Sparse autoencoders on ViT class tokens reveal stable Class Activation Profiles for in-distribution data, enabling OOD detection via divergence from core energy profiles.
-
One-Step Score-Based Density Ratio Estimation
OS-DRE performs score-based density ratio estimation in one step by approximating the temporal score component with a closed-form RBF frame and providing error bounds from approximation theory.
-
Catalyst: Out-of-Distribution Detection via Elastic Scaling
Catalyst improves OOD detection by multiplicatively scaling baseline scores using channel-wise statistics from pre-pooling feature maps, reducing average FPR by 22-33% on standard benchmarks.
-
Level, Sharpness, and Corpus: Why Zero-Shot OOD Detector Rankings Do Not Transfer
Zero-shot OOD detector rankings do not transfer across domains or models; a complementary-evidence wrapper (CEG) cuts FPR95 without using OOD samples.
-
Self-Poisoning in Adaptive Out-of-Distribution Detection: A Sharp-Threshold Theory and Certified Label-Free Calibration
In adaptive OOD detection, bank impurity follows a mean-field urn law whose kernel slope acts as a reproduction number; a frozen-reserve gate removes the supercritical collapse, and a two-world theorem caps label-free...
-
Evaluating Epistemic Uncertainty: Beyond OOD Detection and Active Learning
Epistemic uncertainty should be judged by how well it ranks reducible error, and a new Pareto-gap diagnostic shows proxy-task rankings can invert.
-
FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics
FML-Bench shows that a simple greedy hill-climber performs nearly as well as complex tree-search agents on ML research tasks, with an adaptive strategy that switches exploration modes outperforming all tested agents.
-
Uncertainty Estimation via Hyperspherical Confidence Mapping
HCM estimates uncertainty in neural network outputs by quantifying violation of a unit hypersphere constraint on the normalized direction vector.
-
A Robust Out-of-Distribution Detection Framework via Synergistic Smoothing
ROSS combines median smoothing with local instability measurement to create a robust OOD detector that outperforms prior methods by up to 40 AUROC points on CIFAR and ImageNet benchmarks while defending symmetrically ...
-
Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs
InterNeg improves OOD detection in VLMs by using inter-modal distance criteria for negative text selection and by inverting high-confidence OOD images into additional negative text embeddings.
-
Cross-Distribution Diffusion Priors-Driven Iterative Reconstruction for Sparse-View CT
CDPIR integrates cross-distribution diffusion priors from a Scalable Interpolant Transformer trained with classifier-free guidance into model-based iterative reconstruction to improve sparse-view CT under out-of-distr...
-
Learning Hyperspherical Time-Frequency Representations for Time-Series Out-of-Distribution Detection
Hyperspherical time-frequency representations learned via von Mises-Fisher likelihood improve OOD detection on UCR and UEA archives using k-NN and Mahalanobis scores over contrastive baselines.
-
Mitigating Simplicity Bias in OOD Detection through Object Co-occurrence Analysis
OCO uses object co-occurrence analysis to divide OOD detection into scenarios based on ID training data patterns for improved near-OOD performance.
-
VOLTA: The Surprising Ineffectiveness of Auxiliary Losses for Calibrated Deep Learning
VOLTA, consisting of a deep encoder with learnable prototypes plus cross-entropy and post-hoc temperature scaling, matches or exceeds ten UQ baselines in accuracy, achieves lower expected calibration error, and perfor...
-
RankOOD -- Class Ranking-based Out-of-Distribution Detection
RankOOD detects out-of-distribution samples by training a model to predict fixed class-specific ranking permutations via the Plackett-Luce loss, achieving a 4.3% FPR95 reduction on near-OOD TinyImageNet.
-
A Systematic Analysis of Out-of-Distribution Detection Under Representation and Training Paradigm Shifts
Benchmark across architectures and shift regimes finds OOD detector rankings shift with representation collapse; proposes NC-based shortlist predictor and PCA filter without extra OOD data.
-
$\Delta \mathrm{Energy}$: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization
ΔEnergy, an energy-change OOD score for CLIP, and its EBM fine-tuning loss simultaneously improve OOD detection and covariate-shift generalization.
-
Out-of-Distribution (OOD) Detectors for Open-Set RF Fingerprinting
Applies OOD detectors to open-set RF fingerprinting via an information-theoretic framework and demonstrates effective tuning without OOD data on the POWDER dataset.
-
Enhancing Few-Shot Out-of-Distribution Detection via the Refinement of Foreground and Background
A new framework adds adaptive background patch suppression via entropy weighting and confusable foreground patch rectification to existing FG-BG methods, significantly improving few-shot OOD detection.
-
Safeguarding AI in Medical Imaging: Post-Hoc Out-of-Distribution Detection with Normalizing Flows
Post-hoc normalizing flows for OOD detection in medical imaging achieve 84.61% AUROC on MedOOD and 93.8% on MedMNIST, outperforming ViM, MDS, and ReAct.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.