REVIEW 3 cited by
A Survey on Transfer Learning in Natural Language Processing
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Deep learning models usually require a huge amount of data. However, these large datasets are not always attainable. This is common in many challenging NLP tasks. Consider Neural Machine Translation, for instance, where curating such large datasets may not be possible specially for low resource languages. Another limitation of deep learning models is the demand for huge computing resources. These obstacles motivate research to question the possibility of knowledge transfer using large trained models. The demand for transfer learning is increasing as many large models are emerging. In this survey, we feature the recent transfer learning advances in the field of NLP. We also provide a taxonomy for categorizing different transfer learning approaches from the literature.
Forward citations
Cited by 3 Pith papers
-
CSMF: Cascaded Selective Mask Fine-Tuning for Multi-Objective Embedding-Based Retrieval
CSMF sequentially fine-tunes a two-tower EBR model with selective parameter masks, then serves a weighted linear combination of exposure, click, and conversion scores from one 64-dimensional index, improving offline a...
-
The Muddy Waters of Modeling Empathy in Language: The Practical Impacts of Theoretical Constructs
Empathy tasks with fine-grained definitions and labels directly tied to construct components transfer better to other empathy tasks than tasks with abstract or adjacent labels.
-
Speaker Diarization for Low-Resource Languages Through Wav2vec Fine-Tuning
Fine-tuning Wav2Vec 2.0 on a custom Kurdish corpus is reported to cut speaker diarization error by 7.2 percentage points and raise cluster purity by about 13 percentage points, though the paper contains conflicting numbers.
Discussion (0). Continue with ORCID to comment.