LANG combines language-adaptive hint guidance, progressive decay, and difficulty-tailored learning horizons in RL to boost non-English reasoning performance while preserving language consistency.
Understanding the Repeat Curse in Large Language Models from a Feature Perspective
2 Pith papers cite this work, alongside 3 external citations. Polarity classification is still indexing.
2
Pith papers citing it
3
external citations · OpenAlex
fields
cs.CL 2years
2026 2representative citing papers
citing papers explorer
-
LANG: Reinforcement Learning for Multilingual Reasoning with Language-Adaptive Hint Guidance
LANG combines language-adaptive hint guidance, progressive decay, and difficulty-tailored learning horizons in RL to boost non-English reasoning performance while preserving language consistency.
- Breaking the Likelihood Trap: Variance-Calibrated Modulation for Large Language Model Decoding