REVIEW 17 cited by
Generalized Out-of-Distribution Detection: A Survey
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Out-of-distribution (OOD) detection is critical to ensuring the reliability and safety of machine learning systems. For instance, in autonomous driving, we would like the driving system to issue an alert and hand over the control to humans when it detects unusual scenes or objects that it has never seen during training time and cannot make a safe decision. The term, OOD detection, first emerged in 2017 and since then has received increasing attention from the research community, leading to a plethora of methods developed, ranging from classification-based to density-based to distance-based ones. Meanwhile, several other problems, including anomaly detection (AD), novelty detection (ND), open set recognition (OSR), and outlier detection (OD), are closely related to OOD detection in terms of motivation and methodology. Despite common goals, these topics develop in isolation, and their subtle differences in definition and problem setting often confuse readers and practitioners. In this survey, we first present a unified framework called generalized OOD detection, which encompasses the five aforementioned problems, i.e., AD, ND, OSR, OOD detection, and OD. Under our framework, these five problems can be seen as special cases or sub-tasks, and are easier to distinguish. We then review each of these five areas by summarizing their recent technical developments, with a special focus on OOD detection methodologies. We conclude this survey with open challenges and potential research directions.
Forward citations
Cited by 17 Pith papers
-
Local Hessian Spectral Filtering for Robust Intrinsic Dimension Estimation
LHSD uses spectral filtering on the log-density Hessian to isolate tangent directions from noise and estimate local intrinsic dimension scalably via Stochastic Lanczos Quadrature.
-
TOOD: Task-Aware Out-of-Distribution Score Calibration for Continual Learners
OOD detection in continual learning degrades through task-dependent logit-scale drift and feature-space crowding; a post-hoc per-task energy calibration recovers most of that loss for energy-based detectors.
-
On-Policy Self-Distillation with Sampled Demonstrations Reduces Output Diversity
On-policy self-distillation with sampled demonstrations reduces rollout diversity by amplifying existing probability gaps in the base model, unlike ideal RL which preserves ratios among correct outputs.
-
Variational Inference for Evidential Deep Learning
VI-EDL reformulates evidential deep learning via variational inference to derive an ELBO that limits excessive evidence and a generalization bound that justifies setting Dirichlet parameters to e+1.
-
Debiased Negative Mining Improves Out-of-distribution Detection with Pre-trained Vision-Language Models
Debiased negative mining via Monte-Carlo sampling from ID labels and unlabeled wild data improves OOD detection with VLMs and achieves new state-of-the-art results.
-
Component-Based Out-of-Distribution Detection
CoOD decomposes inputs into components and applies Component Shift Score plus Compositional Consistency Score to improve detection of both standard and compositional out-of-distribution data.
-
Scaling K2 VIII: Short-Period Sub-Neptune Occurrence Rates Peak Around Early-Type M Dwarfs
Short-period sub-Neptunes peak at about 3750 K around early-type M dwarfs, matching pebble accretion predictions, while super-Earths keep rising toward cooler stars.
-
NMINE: Normalized Mutual Information Neural Estimation
A fully neural estimator for normalized mutual information beats a KSG baseline on Gaussian data but fails to deliver its advertised scale-invariance property.
-
Data Selection Through Iterative Self-Filtering for Vision-Language Settings
An iterative bootstrapped self-filtering approach selects balanced clean and diverse subsets from noisy vision-language datasets to train improved CLIP models.
-
Variational Inference for Evidential Deep Learning
VI-EDL trains evidential classifiers with a full Dirichlet KL penalty and a cosine prototype layer, reporting gains in OOD detection but grounding its theory in a flawed ELBO derivation.
-
HEDP: A Hybrid Energy-Distance Prompt-based Framework for Domain Incremental Learning
HEDP uses energy regularization inspired by Helmholtz free energy plus hybrid energy-distance weighting in prompts to improve domain selection and achieve a 2.57% accuracy gain on benchmarks like CORe50 while mitigati...
-
Local Hessian Spectral Filtering for Robust Intrinsic Dimension Estimation
LHSD estimates local intrinsic dimension in high-D spaces by spectral filtering of the log-density Hessian via SLQ to isolate zero-curvature tangent directions.
-
Rare Event Analysis of Large Language Models
Using annealed transition path sampling plus MBAR reweighting, the authors estimate TinyStories-8M completion probabilities for extreme ARI and log-probability values that are unobservable by direct sampling.
-
Energy Landscapes Enable Reliable Abstention in Retrieval-Augmented Large Language Models for Healthcare
An energy-based scoring head trained on dense embeddings improves abstention decisions for medical RAG systems on semantically hard out-of-distribution queries compared to softmax and kNN baselines.
-
Feature Bank Enhancement for Distance-based Out-of-Distribution Detection
Clipping per-dimension outlier features in the training feature bank improves distance-based out-of-distribution detection on ImageNet-1k and CIFAR-10.
-
Edge Case Detection in Automated Driving: Methods, Challenges and Future Directions
The paper delivers a two-level hierarchical classification of edge case detection methods in automated driving, covering AV modules and methodologies, plus evaluation metrics and open challenges.
-
Entropy-Based Non-Invasive Reliability Monitoring of Convolutional Neural Networks
A CNN's activation entropy separates clean from FGSM-attacked image batches in a small VGG-16 test, but fitted binning, tiny samples, and contradictory numbers weaken the claim.
Discussion (0). Sign in to comment.