REVIEW 3 cited by
Learning Soft Labels via Meta Learning
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
One-hot labels do not represent soft decision boundaries among concepts, and hence, models trained on them are prone to overfitting. Using soft labels as targets provide regularization, but different soft labels might be optimal at different stages of optimization. Also, training with fixed labels in the presence of noisy annotations leads to worse generalization. To address these limitations, we propose a framework, where we treat the labels as learnable parameters, and optimize them along with model parameters. The learned labels continuously adapt themselves to the model's state, thereby providing dynamic regularization. When applied to the task of supervised image-classification, our method leads to consistent gains across different datasets and architectures. For instance, dynamically learned labels improve ResNet18 by 2.1% on CIFAR100. When applied to dataset containing noisy labels, the learned labels correct the annotation mistakes, and improves over state-of-the-art by a significant margin. Finally, we show that learned labels capture semantic relationship between classes, and thereby improve teacher models for the downstream task of distillation.
Forward citations
Cited by 3 Pith papers
-
Combined Image Data Augmentations diminish the benefits of Adaptive Label Smoothing
Soft augmentation improves training for single homogeneous augmentations like random erasing, but gives no net benefit and can reduce corruption robustness when combined with diverse augmentations like TrivialAugment.
-
Learning from Ambiguous Data with Hard Labels
A class-wise positive-unlabeled risk estimator trains classifiers from ambiguous data with hard labels and beats label-noise baselines on synthetic mixed-image benchmarks.
-
GovRelBench:A Benchmark for Government Domain Relevance
A new Chinese government-domain benchmark uses a ModernBERT model trained on subjectively assigned, Beta-diffused relevance labels to score LLM responses.
Discussion (0). Continue with ORCID to comment.