Pith. sign in

REVIEW 1 cited by

Second language Korean Universal Dependency treebank v1.2: Focus on data augmentation and annotation scheme refinement

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2503.14718 v1 pith:ONOYDOE5 submitted 2025-03-18 cs.CL

classification cs.CL
keywords languagekoreantreebankannotationdatadatasetsfine-tuningmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We expand the second language (L2) Korean Universal Dependencies (UD) treebank with 5,454 manually annotated sentences. The annotation guidelines are also revised to better align with the UD framework. Using this enhanced treebank, we fine-tune three Korean language models and evaluate their performance on in-domain and out-of-domain L2-Korean datasets. The results show that fine-tuning significantly improves their performance across various metrics, thus highlighting the importance of using well-tailored L2 datasets for fine-tuning first-language-based, general-purpose language models for the morphosyntactic analysis of L2 data.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. UD-KSL Treebank v1.3: A semi-automated framework for aligning XPOS-extracted units with UPOS tags

    cs.CL 2025-06 reject novelty 5.0 of 10

    A semi-automated XPOS-to-UPOS alignment pipeline for the UD-KSL treebank is reported to improve tagging and some parsing accuracy, but the comparison is confounded because the test gold labels also change between conditions.

Pith tools