Continued pre-training on Taiwanese legal text plus instruction tuning did not consistently improve legal reasoning over base models or LoRA routes, and DPO and ORPO alignment degraded accuracy.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Continual Pre-Training is (not) What You Need in Domain Adaption
Continued pre-training on Taiwanese legal text plus instruction tuning did not consistently improve legal reasoning over base models or LoRA routes, and DPO and ORPO alignment degraded accuracy.