Pith. sign in

G1: Teaching llms to reason on graphs with reinforcement learning

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

citation-role summary

background 1

citation-polarity summary

years

2026 5

verdicts

UNVERDICTED 5

roles

background 1

polarities

background 1

representative citing papers

TheoremGraph: Bridging Formal and Informal Mathematics

cs.IR · 2026-06-24 · unverdicted · novelty 7.0

TheoremGraph builds a unified statement-level dependency graph across informal arXiv math and formal Lean code via parsing, embeddings, and LLM validation, releasing the data and APIs for search and retrieval.

Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key

cs.AI · 2026-05-07 · unverdicted · novelty 6.0 · 3 refs

RL training compute for logical reasoning follows a power law with horizon depth whose exponent rises with logical expressiveness, yielding better downstream transfer when models train on richer logics.

citing papers explorer

Showing 5 of 5 citing papers.