REVIEW 3 cited by
Strong Baselines for Neural Semi-supervised Learning under Domain Shift
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Novel neural models have been proposed in recent years for learning under domain shift. Most models, however, only evaluate on a single task, on proprietary datasets, or compare to weak baselines, which makes comparison of models difficult. In this paper, we re-evaluate classic general-purpose bootstrapping approaches in the context of neural networks under domain shifts vs. recent neural approaches and propose a novel multi-task tri-training method that reduces the time and space complexity of classic tri-training. Extensive experiments on two benchmarks are negative: while our novel method establishes a new state-of-the-art for sentiment analysis, it does not fare consistently the best. More importantly, we arrive at the somewhat surprising conclusion that classic tri-training, with some additions, outperforms the state of the art. We conclude that classic approaches constitute an important and strong baseline.
Forward citations
Cited by 3 Pith papers
-
AKD : Adversarial Knowledge Distillation For Large Language Models Alignment on Coding tasks
AKD combines adversarially sampled synthetic exercises with Direct Preference Optimization to fine-tune small code models, but its reported gains over standard fine-tuning are not supported by its own tables.
-
Geodesic Flow Kernels for Semi-Supervised Learning on Mixed-Variable Tabular Dataset
GFTab, a semi-supervised tabular method with variable-specific corruptions and geodesic flow kernel similarity, reports the best F1 on about half of 21 mixed-variable benchmarks with sparse labels.
-
Semi-supervised Thai Sentence Segmentation Using Local and Distant Word Representations
A Bi-LSTM-CRF with n-gram embeddings, self-attention, and modified Cross-View Training improves Thai sentence segmentation F1 to 92.5% on Orchid and 88.9% on UGWC, and punctuation restoration overall F1 to 65.2% on IWSLT.
Discussion (0). Continue with ORCID to comment.