Pith. sign in

Deep reinforcement learning from human preferences

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.CL 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

Who Reasons in the Large Language Models?

cs.CL · 2025-05-27 · conditional · novelty 6.0

The authors hypothesize, with evidence from merging and freezing experiments, that the attention output projection is the primary locus of mathematical reasoning in large language models.

citing papers explorer

Showing 1 of 1 citing paper.

  • Who Reasons in the Large Language Models? cs.CL · 2025-05-27 · conditional · none · ref 8

    The authors hypothesize, with evidence from merging and freezing experiments, that the attention output projection is the primary locus of mathematical reasoning in large language models.