REVIEW 3 cited by
MixText: Linguistically-Informed Interpolation of Hidden Space for Semi-Supervised Text Classification
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
This paper presents MixText, a semi-supervised learning method for text classification, which uses our newly designed data augmentation method called TMix. TMix creates a large amount of augmented training samples by interpolating text in hidden space. Moreover, we leverage recent advances in data augmentation to guess low-entropy labels for unlabeled data, hence making them as easy to use as labeled data.By mixing labeled, unlabeled and augmented data, MixText significantly outperformed current pre-trained and fined-tuned models and other state-of-the-art semi-supervised learning methods on several text classification benchmarks. The improvement is especially prominent when supervision is extremely limited. We have publicly released our code at https://github.com/GT-SALT/MixText.
Forward citations
Cited by 3 Pith papers
-
Backtranslation and paraphrasing in the LLM era? Comparing data augmentation methods for emotion classification
Backtranslation and paraphrasing produce competitive or better classification gains than zero-shot and few-shot generation when augmenting a low-resource emotion dataset.
-
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
AKD combines adversarially sampled synthetic exercises with Direct Preference Optimization to fine-tune small code models, but its reported gains over standard fine-tuning are not supported by its own tables.
-
The Efficiency of Pre-training with Objective Masking in Pseudo Labeling for Semi-Supervised Text Classification
CformerM extends Cformer with LDA-based objective masking during pre-training and reports consistent, modest accuracy gains over Cformer and baselines across four text datasets.
Discussion (0). Continue with ORCID to comment.