Pith. sign in

Verifiers: Environments for llm reinforcement learning

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

method 1

citation-polarity summary

fields

cs.CL 1

years

2026 1

verdicts

CONDITIONAL 1

roles

method 1

polarities

use method 1

representative citing papers

Ask-E: An Environment for Calibrated Question Generation

cs.CL · 2026-08-07 · conditional · novelty 6.0

A language model trained only to write questions that split two weaker solvers improves at solving math problems, while even frontier models calibrate less than half the time.

citing papers explorer

Showing 1 of 1 citing paper.

  • Ask-E: An Environment for Calibrated Question Generation cs.CL · 2026-08-07 · conditional · none · ref 7

    A language model trained only to write questions that split two weaker solvers improves at solving math problems, while even frontier models calibrate less than half the time.