REVIEW 5 cited by
Searching for Optimal Subword Tokenization in Cross-domain NER
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Input distribution shift is one of the vital problems in unsupervised domain adaptation (UDA). The most popular UDA approaches focus on domain-invariant representation learning, trying to align the features from different domains into similar feature distributions. However, these approaches ignore the direct alignment of input word distributions between domains, which is a vital factor in word-level classification tasks such as cross-domain NER. In this work, we shed new light on cross-domain NER by introducing a subword-level solution, X-Piece, for input word-level distribution shift in NER. Specifically, we re-tokenize the input words of the source domain to approach the target subword distribution, which is formulated and solved as an optimal transport problem. As this approach focuses on the input level, it can also be combined with previous DIRL methods for further improvement. Experimental results show the effectiveness of the proposed method based on BERT-tagger on four benchmark NER datasets. Also, the proposed method is proved to benefit DIRL methods such as DANN.
Forward citations
Cited by 5 Pith papers
-
StructCoh: Structured Contrastive Learning for Context-Aware Text Semantic Matching
StructCoh, a graph-enhanced contrastive learning framework for text semantic matching, reportedly outperforms prior methods on legal and plagiarism benchmarks, but the reported results are not reproducible from the pa...
-
Using External knowledge to Enhanced PLM for Semantic Matching
A BERT-based model that injects WordNet lexical-relation signals into attention and adaptively fuses them claims consistent accuracy improvements on 10 semantic matching datasets.
-
Comateformer: Combined Attention Transformer for Semantic Sentence Matching
Comateformer replaces softmax attention with a product of tanh similarity and sigmoid dissimilarity scores, and reports consistent gains on ten semantic matching datasets.
-
Multi-Granularity Reasoning for Natural Language Inference
Stacking element-wise multi-layer BERT interactions and DenseNet yields modest NLI gains over BERT/RoBERTa baselines on standard benchmarks.
-
Boosting Neural Language Inference via Cascaded Interactive Reasoning
A new feature-extraction module for NLI combines all BERT layers with element-wise sentence-pair interactions and a DenseNet, reporting average gains of roughly one point on ten benchmarks, though evaluation inconsist...
Discussion (0). Continue with ORCID to comment.