WizardMath applies RLEIF to produce open-source LLMs that reach new state-of-the-art math reasoning scores on GSM8k and MATH, with the 70B variant surpassing GPT-3.5-Turbo, Claude 2, Gemini Pro, and early GPT-4.
It surpasses all existing open-source state-of-the-art models, showcasing the effectiveness and robustness of the RLEIF approach proposed in our study
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2023 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct
WizardMath applies RLEIF to produce open-source LLMs that reach new state-of-the-art math reasoning scores on GSM8k and MATH, with the 70B variant surpassing GPT-3.5-Turbo, Claude 2, Gemini Pro, and early GPT-4.