Pith. sign in

T ree RL : LLM reinforcement learning with on-policy tree search

4 Pith papers cite this work, alongside 1 external citations. Polarity classification is still indexing.

4 Pith papers citing it
1 external citations · OpenAlex

fields

cs.CL 2 cs.LG 2

years

2026 4

verdicts

UNVERDICTED 4

representative citing papers

Self-Improving Language Models with Bidirectional Evolutionary Search

cs.CL · 2026-05-27 · unverdicted · novelty 6.0

Bidirectional Evolutionary Search augments autoregressive expansion with evolutionary recombination operators and dense backward subgoal feedback to produce better candidates than standard best-of-N or tree search for language model self-improvement.

citing papers explorer

Showing 4 of 4 citing papers.