LLMs improve reasoning in math, science, and coding by generating and self-filtering their own training data through cycle-consistency, factuality, and correctness checks on unlabeled prompts.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2026 1verdicts
UNVERDICTED 1representative citing papers
citing papers explorer
-
Self-Verified Distillation: Your Language Model Is Secretly Its Own Synthetic Data Pipeline
LLMs improve reasoning in math, science, and coding by generating and self-filtering their own training data through cycle-consistency, factuality, and correctness checks on unlabeled prompts.