Pith. sign in

REVIEW 1 cited by

Order-free Learning Alleviating Exposure Bias in Multi-label Classification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1909.03434 v1 pith:5G7EMA6M submitted 2019-09-08 cs.LG cs.CLcs.SDeess.ASstat.ML

classification cs.LGcs.CLcs.SDeess.ASstat.ML
keywords labelclassificationmulti-labeltrainingapproachbiascombinationsdecoder
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Multi-label classification (MLC) assigns multiple labels to each sample. Prior studies show that MLC can be transformed to a sequence prediction problem with a recurrent neural network (RNN) decoder to model the label dependency. However, training a RNN decoder requires a predefined order of labels, which is not directly available in the MLC specification. Besides, RNN thus trained tends to overfit the label combinations in the training set and have difficulty generating unseen label sequences. In this paper, we propose a new framework for MLC which does not rely on a predefined label order and thus alleviates exposure bias. The experimental results on three multi-label classification benchmark datasets show that our method outperforms competitive baselines by a large margin. We also find the proposed approach has a higher probability of generating label combinations not seen during training than the baseline models. The result shows that the proposed approach has better generalization capability.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Machine learning revolution for exoplanet direct imaging detection: transformer architectures

    astro-ph.EP 2025-08 conditional novelty 6.0 of 10

    A CNN-Transformer sequence model detects injected moving planet signals in synthetic and semi-synthetic JWST high-contrast imaging data, reaching 100% accuracy in one reported run.

Pith tools