Pith. sign in

Paper Citation Record · LEDGER

Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2411.04282.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.04282 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:51:06.663262Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 28fc5344-66f0-4006-81e0-516a0875b960 · inbound

BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning cites this paper.

BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T22:20:13.669361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:20:13.669361Z digest=sha256:c36d75708df22b806b6cf51cf66e4696916e5fba632e7010f2e70c6f7a4bb7d2

Observation 35a6b2f9-8d39-4f06-adff-8953b346837d · inbound

MetaSC: Test-Time Safety Specification Optimization for Language Models cites this paper.

MetaSC: Test-Time Safety Specification Optimization for Language Models Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T11:15:44.960566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:15:44.960566Z digest=sha256:5ed28ebd0f9164d143c304219946a8ef1ab0cd644cf9a8a1f1774ff9c86015d5

Observation 6204eb4b-af09-4a0a-95fb-96107ab3a866 · inbound

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models cites this paper.

Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T08:40:41.796532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T08:40:40.910461Z digest=sha256:1b9cd4dec58a12edb645f255b8d9863bc1ce35db681739db3c3c847fc59da78d

Observation 2d8fbe99-9288-484c-8f15-03077a6ed928 · inbound

Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents cites this paper.

Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:51:06.663262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:51:06.663262Z digest=sha256:0f727da3efedd6cd6b7e96c1c4c558834b6671c466b1b1ac78108ad1e4c318cb

Observation 60267126-4892-44c3-83d9-e841b54957bd · inbound

LARES: Latent Reasoning for Sequential Recommendation cites this paper.

LARES: Latent Reasoning for Sequential Recommendation Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T14:59:15.456174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:59:15.456174Z digest=sha256:942bcd7111b564d14de91e7f4a6e8e195908ae2b388ba8b503b80413ec2d6ac5

Observation 4f4b8e15-5e15-4a2f-ad02-412bf176573e · inbound

Iterative Self-Incentivization Empowers Large Language Models as Agentic Searchers cites this paper.

Iterative Self-Incentivization Empowers Large Language Models as Agentic Searchers Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:20.874472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:20.874472Z digest=sha256:2c78161d70e82f33c7f3b34e71342f29118c8fabaedd89346213c87773b7c5c0

Observation 52ebb6e6-cb33-47bf-9fd6-91affe7684c7 · inbound

Reinforcing General Reasoning without Verifiers cites this paper.

Reinforcing General Reasoning without Verifiers Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:50.020849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:35:50.020849Z digest=sha256:0bc59c66f375bfa9b3d0f014059767a7a9a7ce22b9232634a589caf0c0947b86

Observation 4c99523a-ece3-43b6-88ec-2b64a9fcaa15 · inbound

Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty cites this paper.

Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:38:09.456278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:38:09.456278Z digest=sha256:f67b0d56983c62d1bf4415f4f0ff5ec168a7feebad7ce87b0235ac8eaf684131

Observation 179665dc-1fcf-454b-9c3c-6bf43aa56ab8 · inbound

Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks cites this paper.

Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:52:14.073815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-19T09:48:56.990745Z digest=sha256:c806410b77c62b875e3d5364946ae38bd78c3bb456f500b9a16610990eb62676

Observation 6baeb213-9b97-482c-b2d0-cea8eb952187 · inbound

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards cites this paper.

CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:26.853811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:26.853811Z digest=sha256:13ba1bfae4168903eb25adb20fa5d43c733ec46384b3a72a9a125a6259b74085

Observation 5abe9b40-224b-4dc7-95d2-83f3a27f7637 · inbound

Implicit Reasoning in Large Language Models: A Comprehensive Survey cites this paper.

Implicit Reasoning in Large Language Models: A Comprehensive Survey Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:36.459749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:36.459749Z digest=sha256:0097ab86acceee3781844f86ab43bb92153ce542d824dbea37e6d924ff1a32f6

Observation d5bd85dc-7ea3-422f-b2ec-095d848a1e5d · inbound

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification cites this paper.

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:50.125223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:50.125223Z digest=sha256:ca0495387c8f000ddab9450cfa5fe003296641ca559ca69b2fe03c0657b4b4e8

Observation e026c71f-ed98-4660-954a-bb8ca40016ef · inbound

Coupled Variational Reinforcement Learning for Language Model General Reasoning cites this paper.

Coupled Variational Reinforcement Learning for Language Model General Reasoning Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T16:43:54.497117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:43:54.497117Z digest=sha256:3a4f7bd1e1163352749882f21501f9612bf9b3dd44c6cad84f3c73afc490b7ae

Observation fc69de2b-418d-467d-b49d-aa322210ffed · inbound

LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning cites this paper.

LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T04:03:51.020124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:03:51.020124Z digest=sha256:301678b274fda1b57f8fc662cc0547734fb898c8c7d98875343c117eb4265318

Observation edbde4cb-91e8-49c1-8443-c29144631805 · inbound

Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning cites this paper.

Self-Supervised Bootstrapping of Action-Predictive Embodied Reasoning Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 68

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T14:04:11.993239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T14:03:48.795572Z digest=sha256:0756f5fa5235ec846ad34084f329d24d03d494fc2c5eecb85c8bb28f81f3dfde

Observation 307f65aa-17c1-42ea-b2ea-26740158adaf · inbound

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction cites this paper.

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:26:41.674341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T07:26:35.767179Z digest=sha256:bd85bf0049eab604e13239c8c3917dca727f21d03076b63f881c9d635035e809

Observation 14b95965-dd5e-484d-b895-68bc47d1329c · inbound

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook cites this paper.

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-13T14:03:01.974171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:03:01.974171Z digest=sha256:89db50da7a2c9f040d68753ea843a4b452ac6bf901338eb9e36653591b363973

Observation a628f43c-6c0d-408f-a078-4e7751847f9d · inbound

LASAR: Latent Adaptive Semantic Aligned Reasoning for Generative Recommendation cites this paper.

LASAR: Latent Adaptive Semantic Aligned Reasoning for Generative Recommendation Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:46:28.958348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T04:58:07.060409Z digest=sha256:de61399e267a728b6fcf4502ec08a527ae53762d38406242c57efe72cacf68ec

Observation 7e883577-6c8a-4afc-b647-d44119e552e6 · inbound

LaME: Learning to Think in Latent Space for Multimodal Embedding via Information Bottleneck cites this paper.

LaME: Learning to Think in Latent Space for Multimodal Embedding via Information Bottleneck Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T13:48:21.001531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-27T07:31:39.234875Z digest=sha256:133430a7eca505d2e20b3ed6daba087f80de684feee21f90e999dfc034817a9e

Observation 332b0302-aef3-4d0c-9719-984935300d68 · inbound

Unsupervised Causal Abstractions Discovery cites this paper.

Unsupervised Causal Abstractions Discovery Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-26T20:49:57.368019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T20:45:15.226889Z digest=sha256:0ff33be3836dc5a7a83b6b43060ee49d3a297b1a85c090c139558e2fa2eec94c

Observation f9dc57ef-f1d8-41dd-a289-b8f0f9abc9b9 · inbound

When are likely answers right? On Sequence Probability and Correctness in LLMs cites this paper.

When are likely answers right? On Sequence Probability and Correctness in LLMs Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T15:09:54.101012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T01:55:09.626278Z digest=sha256:1c853eddc32cae6b03c8b57c5a6b5781e6f7be9c9aebec9992e5a7bd10a1492f

Observation ba45b263-4c6f-4fe2-9d36-c4d3b742c98f · inbound

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops cites this paper.

Recursive Self-Improvement in AI: From Bounded Self-Refinement to Autonomous Research Loops Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding

Reference 101

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T03:45:55.533680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-09T03:36:57.168246Z digest=sha256:90fb87a0399aa344ec4a9665bba0d96b4817a351d63cb638ad92d6faf7e0473b