Pith. sign in

REVIEW 2 cited by

Superbizarre Is Not Superb: Derivational Morphology Improves BERT's Interpretation of Complex Words

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2101.00403 v3 pith:AEPM27M5 submitted 2021-01-02 cs.CL

classification cs.CL
keywords bertinputwordscomplexplmssegmentationderivationalgeneralization
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

How does the input segmentation of pretrained language models (PLMs) affect their interpretations of complex words? We present the first study investigating this question, taking BERT as the example PLM and focusing on its semantic representations of English derivatives. We show that PLMs can be interpreted as serial dual-route models, i.e., the meanings of complex words are either stored or else need to be computed from the subwords, which implies that maximally meaningful input tokens should allow for the best generalization on new words. This hypothesis is confirmed by a series of semantic probing tasks on which DelBERT (Derivation leveraging BERT), a model with derivational input segmentation, substantially outperforms BERT with WordPiece segmentation. Our results suggest that the generalization capabilities of PLMs could be further improved if a morphologically-informed vocabulary of input tokens were used.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Model Decides How to Tokenize: Adaptive DNA Sequence Tokenization with MxDNA

    q-bio.GN 2024-12 conditional novelty 7.0 of 10

    A learnable tokenization module with mixture of convolution experts and deformable convolution improves DNA foundation model performance on Genomic and Nucleotide Transformer Benchmarks.

  2. KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation

    cs.CL 2025-07 conditional novelty 5.0 of 10

    KinyaColBERT, a morphology-aware two-tier ColBERT retriever, reports large MRR gains over multilingual baselines and commercial APIs on a new Kinyarwanda agricultural retrieval benchmark.

Pith tools