Pith. sign in

REVIEW 1 cited by

Learning to Perturb Word Embeddings for Out-of-distribution QA

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2105.02692 v3 pith:POESHTOG submitted 2021-05-06 cs.CL

classification cs.CL
keywords datamodelmodelstrainedwordeffectiveembeddingmethod
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

QA models based on pretrained language mod-els have achieved remarkable performance on various benchmark datasets.However, QA models do not generalize well to unseen data that falls outside the training distribution, due to distributional shifts.Data augmentation (DA) techniques which drop/replace words have shown to be effective in regularizing the model from overfitting to the training data.Yet, they may adversely affect the QA tasks since they incur semantic changes that may lead to wrong answers for the QA task. To tackle this problem, we propose a simple yet effective DA method based on a stochastic noise generator, which learns to perturb the word embedding of the input questions and context without changing their semantics. We validate the performance of the QA models trained with our word embedding perturbation on a single source dataset, on five different target domains.The results show that our method significantly outperforms the baselineDA methods. Notably, the model trained with ours outperforms the model trained with more than 240K artificially generated QA pairs.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. FPAN: Mitigating Replication in Diffusion Models through the Fine-Grained Probabilistic Addition of Noise to Token Embeddings

    cs.CV 2025-05 conditional novelty 6.0 of 10

    Probabilistically adding high-intensity noise to individual token embeddings during fine-tuning reduces replication in Stable Diffusion by up to 28.78% in the paper's experiments, with unchanged or improved FID.

Pith tools