Pith. sign in

REVIEW 18 cited by

Long-tail learning via logit adjustment

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2007.07314 v2 pith:XPXSDIBO submitted 2020-07-14 cs.LG stat.ML

Long-tail learning via logit adjustment

classification cs.LG stat.ML
keywords labelsadjustmentdominantlabellearninglogittechniquestraining
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Real-world classification problems typically exhibit an imbalanced or long-tailed label distribution, wherein many labels are associated with only a few samples. This poses a challenge for generalisation on such labels, and also makes na\"ive learning biased towards dominant labels. In this paper, we present two simple modifications of standard softmax cross-entropy training to cope with these challenges. Our techniques revisit the classic idea of logit adjustment based on the label frequencies, either applied post-hoc to a trained model, or enforced in the loss during training. Such adjustment encourages a large relative margin between logits of rare versus dominant labels. These techniques unify and generalise several recent proposals in the literature, while possessing firmer statistical grounding and empirical performance.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 18 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. From Perturbation Correction to Geometry-Aware Sampling: Sharpness-Guided Equilibrium Sampling for Balanced Flat Minima in Long-Tailed Learning

    cs.LG 2026-07 conditional novelty 6.0

    Sharpness-Guided Equilibrium Sampling reweights long-tailed training batches using cumulative class counts and SAM perturbation-loss gaps, improving tail accuracy by up to 10.8 points.

  2. Loss Landscape Topology Reveals Why Simple Baselines are Competitive at 3D Point Cloud Segmentation Under Class Imbalance

    cs.CV 2026-07 conditional novelty 6.0

    Uniform cross-entropy is competitive with 11 imbalance-aware losses in point-based 3D segmentation; specialized methods give small, architecture-dependent gains.

  3. Revisiting Scene Graph Generation from the Perspective of Detector-Conditioned Reachability

    cs.CV 2026-07 accept novelty 6.0

    A dual-query scene graph generation method unifies detector-based and query-based reasoning in a single decoder, achieving state-of-the-art results on Visual Genome, Open Images v6, and GQA-200.

  4. Class-frequency Guided Noise Schedule for Diffusion Models

    cs.LG 2026-06 unverdicted novelty 6.0

    Proposes CFRG noise schedule for diffusion models that assigns larger noises to low-frequency classes to improve generation on imbalanced datasets.

  5. Modular Diffusion Models for Structured Visual Recognition

    cs.CV 2026-06 unverdicted novelty 6.0

    Modular Diffusion Models decompose diffusion into task-specific modules to model distributions over structured visual outputs for detection, segmentation, and scene graph generation.

  6. Rethinking Loss Reweighting for Imbalance Learning as an Inverse Problem: A Neural Collapse Point of View

    cs.LG 2026-05 unverdicted novelty 6.0

    Loss reweighting is cast as an inverse problem that dynamically infers class weights to equalize per-class average losses under the Neural Collapse simplex ETF target.

  7. Mitigating Label Shift in Tabular In-Context Learning via Test-Time Posterior Adjustment

    cs.LG 2026-05 unverdicted novelty 6.0

    DistPFN is a test-time posterior adjustment that rescales TabPFN class probabilities to reduce overfitting to the training class distribution under label shift.

  8. Mitigating Label Shift in Tabular In-Context Learning via Test-Time Posterior Adjustment

    cs.LG 2026-05 unverdicted novelty 6.0

    DistPFN is a test-time posterior adjustment technique that mitigates label shift in TabPFN by downweighting the training prior and emphasizing the model's predicted posterior, with a temperature-scaled variant, evalua...

  9. SPECTRA: Spectral Domain-Aware Graph Generation for Imbalanced Molecular Property Regression

    cs.LG 2025-11 unverdicted novelty 6.0

    SPECTRA improves molecular property regression on underrepresented targets via spectral graph generation with rarity-aware budgeting and Laplacian interpolation, paired with edge-aware Chebyshev GNNs, yielding competi...

  10. LoFT: Parameter-Efficient Fine-Tuning for Long-tailed Semi-Supervised Learning in Open-World Scenarios

    cs.LG 2025-09 unverdicted novelty 6.0

    LoFT uses parameter-efficient fine-tuning of foundation models for long-tailed semi-supervised learning, supported by proofs that this reduces hypothesis complexity to minimize balanced posterior error and compresses ...

  11. Mutually Exclusive Multiclass Lesion Segmentation in Neuroimaging: Binary-Guided Weak Supervision with Inter-Class Orthogonality

    eess.IV 2026-07 accept novelty 5.0

    Binary-guided mutual exclusivity with inter-class orthogonality yields accurate multiclass weakly supervised neuroimaging lesion segmentation from image-level labels alone.

  12. Simultaneous Long-tailed Recognition and Multi-modal Fusion for Highly Imbalanced Multi-modal Data

    cs.CV 2026-05 unverdicted novelty 5.0

    A multi-modal extension of multi-expert architectures uses confidence-guided fusion from modality-specific networks to handle long-tailed class imbalance across heterogeneous inputs.

  13. STaR-DRO: Stateful Tsallis Reweighting for Group-Robust Structured Prediction

    cs.LG 2026-04 unverdicted novelty 5.0

    STaR-DRO applies momentum-smoothed Tsallis reweighting to focus learning on hard groups in structured prediction, yielding F1 gains on clinical label extraction.

  14. SciLT: Long-tailed Image Classification under Scientific Image Domains

    cs.CV 2026-04 conditional novelty 5.0

    On scientific long-tailed image tasks, foundation-model fine-tuning gains are limited; SciLT fuses penultimate and final ViT features under dual supervision to improve balanced accuracy.

  15. Dual-Margin Embedding for Fine-Grained Long-Tailed Plant Taxonomy

    cs.CV 2025-12 unverdicted novelty 5.0

    TaxoNet uses a dual-margin objective to reshape decision boundaries in long-tailed fine-grained plant taxonomy, improving rare-class geometry under open-world conditions.

  16. Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond

    cs.AI 2024-08 unverdicted novelty 5.0

    The ADC method automates the creation of large image classification datasets using LLMs and search engines, achieving 79% human agreement and reducing label noise on a 1 million image clothing dataset, while also rele...

  17. Enhancing Oracle Bone Inscription Recognition via Multi-Scale Layer Attention

    cs.CV 2026-06 unverdicted novelty 4.0

    MSLA is a new attention mechanism that models multi-scale and cross-layer interactions to achieve more accurate OBI recognition than prior attention methods.

  18. Dynamic Distillation and Gradient Consistency for Robust Long-Tailed Incremental Learning

    cs.CV 2026-05 unverdicted novelty 4.0

    Gradient consistency regularization and entropy-driven dynamic distillation improve accuracy by up to 5% in long-tailed incremental learning, with strong gains in majority-to-minority task ordering.