A full training pipeline for SOAE-specific LLMs combines continual pre-training, two-stage curriculum SFT, and distilled speculative decoding; the authors report retained general ability, domain gains, and 1.39-1.52x speedup.
In: ICASSP 2025-2025 IEEE International Conference on Acous- tics, Speech and Signal Processing (ICASSP)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
SOAEsV2-7B/72B: Full-Pipeline Optimization for State-Owned Enterprise LLMs via Continual Pre-Training, Domain-Progressive SFT and Distillation-Enhanced Speculative Decoding
A full training pipeline for SOAE-specific LLMs combines continual pre-training, two-stage curriculum SFT, and distilled speculative decoding; the authors report retained general ability, domain gains, and 1.39-1.52x speedup.