Pith. sign in

Paper Citation Record · LEDGER

Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2504.05419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.05419 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 39 of 39 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:13:12.982125Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 666e2612-48b2-4d5a-83da-d9edd88670b6 · inbound

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective cites this paper.

Deciphering Trajectory-Aided LLM Reasoning: An Optimization Perspective Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 130

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:12.982125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:12.982125Z digest=sha256:9f99cd4c8a92a8e523d1d8e27b738a73f808dcbab72362a8592b4c99a3c31b83

Observation 2532eceb-f695-4c11-b079-920187603302 · inbound

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency cites this paper.

Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:19:20.228761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:19:20.228761Z digest=sha256:865dfcd6f09c9dc889439595cee7c5f6751f69fdf8fd793cefb48e2bcf1874c3

Observation 8a2025f4-41bc-4e8f-b31d-7ca5b723881d · inbound

The Geometries of Truth Are Orthogonal Across Tasks cites this paper.

The Geometries of Truth Are Orthogonal Across Tasks Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:12:50.596167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:12:50.596167Z digest=sha256:8a25eeebf379871681ccebf19414208bedd66a2476f422d91f77374036f5e2dc

Observation 2a21f87a-3f81-4075-8720-f0df015b0e62 · inbound

Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement cites this paper.

Exploring and Exploiting the Inherent Efficiency within Large Reasoning Models for Self-Guided Efficiency Enhancement Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:49.924159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:49.924159Z digest=sha256:2d2a3f6c62f3ac881be13b246d5d5abc620475d4546ce5b66e4dc95f97cfc509

Observation 3936ad77-716d-4947-a689-efe6d943f63d · inbound

Real-Time Progress Prediction in Reasoning Language Models cites this paper.

Real-Time Progress Prediction in Reasoning Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:09.871354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:55:09.871354Z digest=sha256:95d8c8694d80b868767c26513454f882de341f07453b5a11fda0df86f763c4a5

Observation 23c1314f-b0f2-4e6e-96f3-05dd4da8282f · inbound

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs cites this paper.

CoRE: Enhancing Metacognition with Label-free Self-evaluation in LRMs Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T19:16:31.411480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:16:31.411480Z digest=sha256:9a0ddcb834af3c5a2ecafd494de61ad5297d5826ab3e5d21a9af2ea46a1a73ea

Observation 1e65bdff-a7a5-4a94-bb64-b63464b4fe62 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 237

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:18.051568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:54:18.051568Z digest=sha256:ff4b2d4ed950564b8178c64556cf622f61280cbf1c92da6e90dc250765940415

Observation 0c49d9c9-e8d3-4485-9753-cb07cd765495 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:03.846399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:03.846399Z digest=sha256:cb4c654aecfb3c9501730811888068cdee8f46ad5f69a4f479c90a1e139aab00

Observation 8d6fa640-ea61-411e-a4f8-da6b9361be6a · inbound

Rethinking LLM Parametric Knowledge as Post-retrieval Confidence for Dynamic Retrieval and Reranking cites this paper.

Rethinking LLM Parametric Knowledge as Post-retrieval Confidence for Dynamic Retrieval and Reranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T23:34:30.832661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:34:30.832661Z digest=sha256:21367143231208d52210f9b459649f7fe1a9b3cc283a2992753e035ecab0918a

Observation 768c89b5-0e74-42c4-865e-5124e83f0e87 · inbound

Entropy After </Think> for reasoning model early exiting cites this paper.

Entropy After </Think> for reasoning model early exiting Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:52:35.578407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T11:51:58.579048Z digest=sha256:ccd9c502a3bfe678d826420922b967bafb0dcc3d0cc052b72fad325204c89135

Observation c9db14a0-1101-485b-ab7c-eda414fd952a · inbound

Learning More from Less: Unlocking Internal Representations for Benchmark Compression cites this paper.

Learning More from Less: Unlocking Internal Representations for Benchmark Compression Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T06:01:59.124009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:01:59.124009Z digest=sha256:e32c170281896120d61923bd5717c93863710281951060ebd4975afe466f39fb

Observation 47325669-6598-4559-a7be-eaa5878c0911 · inbound

Conformal Thinking: Risk Control for Reasoning on a Compute Budget cites this paper.

Conformal Thinking: Risk Control for Reasoning on a Compute Budget Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T07:47:32.666284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T07:45:53.184670Z digest=sha256:313df3b69971220fd5b77e8293c513a63666054c54cc4fbf9ab4b69acb46b219

Observation ded4de4c-6ae2-4e38-aba3-a867ef938e09 · inbound

Emergent Manifold Separability during Reasoning in Large Language Models cites this paper.

Emergent Manifold Separability during Reasoning in Large Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:35.060139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T20:12:05.809065Z digest=sha256:e0eb4fc2c424a2f6990c59f767492c58e2a8fc06773a9381d63579d4b08e4f1d

Observation 0cc44ec1-660b-4072-b918-96deb8804ab0 · inbound

How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality cites this paper.

How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:35:51.917676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:26:04.301213Z digest=sha256:d910d360b105e8094eb4d607105da222e9a6dc449db479538800818b5d730c84

Observation 2f2e76be-f53a-4869-be90-2dc30e3354a6 · inbound

LLM Reasoning Is Latent, Not the Chain of Thought cites this paper.

LLM Reasoning Is Latent, Not the Chain of Thought Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:53:04.614744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:49:05.178087Z digest=sha256:3d05d854e996c5c579106e143567f520db5302901261ab35da1d1d0209976dc3

Observation 5f202c30-32b6-46dd-9e5f-9b5578072dcc · inbound

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping cites this paper.

TRACES: Tagging Reasoning Steps for Adaptive Cost-Efficient Early-Stopping Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:29:47.262377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:26:45.372232Z digest=sha256:5ed6f665cb3fa8d09b35825aac0ddba9887f100fcf4a6b7e58eb578082e50d2c

Observation 12528287-ffee-497e-90e6-4cd667937eb6 · inbound

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation cites this paper.

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:10.371417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T06:38:32.413172Z digest=sha256:d6688101c9023c7239d9e1caad9b83cdc33192f2a674698d0b9590ac934e6b11

Observation 72c7a5b4-7de5-4919-a122-727a24f18f64 · inbound

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation cites this paper.

Do Prompt-Elicited Trajectories Reflect Training-Time Reward Hacking? A Systematic Study on Monitoring Training-Time Reward Hacking in Code Generation Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:05:40.207092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T09:56:48.019404Z digest=sha256:9f75d8b176a7c2936d132c456df4de4ddad2561e3a82687b5f75cc77a134f91b

Observation 30b3732b-6007-4ad6-8a90-e05a311398f9 · inbound

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models cites this paper.

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:21:08.888017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T17:20:19.586214Z digest=sha256:08bf088c2aaeefa324ec28d9f2209893c157646fdbd033352e21432c7d9eafa8

Observation 743a19ff-25c2-4fe5-baf3-d38cf596103e · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:00:56.545726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T00:54:25.549158Z digest=sha256:e99d31ce1c82bb81018944b65a2c27d99b4b2303a440253e76f42d365a721e03

Observation ba999c67-75b8-4b6f-8675-af4b50895d01 · inbound

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight cites this paper.

Behavior Cue Reasoning: Monitorable Reasoning Improves Efficiency and Safety through Oversight Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:29:52.735984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T08:29:09.122055Z digest=sha256:278229fdf015be5c92dee7e6a1eee4109592a4d9704233d0193d8bea328434e2

Observation b83c8184-7213-4655-b137-7250c9e2aee1 · inbound

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling cites this paper.

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:10:58.214344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:54:34.000216Z digest=sha256:53bf502dea123096a777265a53f8856f12d5743129aa064a9f20f65e071a50a0

Observation 806489af-17ad-4a9b-9c82-2aac09eb005b · inbound

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling cites this paper.

LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:12:28.383220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T07:09:02.672233Z digest=sha256:f0605f95147878043f27264995d54924bdbf67da37cf183a16f00af0977ae5fd

Observation f2733e87-406e-44a2-918b-26f98d3f87b2 · inbound

Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal cites this paper.

Hidden Error Awareness in Chain-of-Thought Reasoning: The Signal Is Diagnostic, Not Causal Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:31:25.390664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T05:11:51.012343Z digest=sha256:26ac8b8c3e925eebde3a78e388bcef56ca78a3fe9dbb2758df473bcf38db1a41

Observation b626de8d-581c-4f7c-9b28-4eaf8ae8f688 · inbound

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking cites this paper.

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:46.315000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:58:46.417344Z digest=sha256:2e56fd795c86eff99837d6353ea9cc0ca96a34543248efca7371a31c098d4749

Observation 859c5d38-85c4-446f-92dc-24b86657a5ae · inbound

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking cites this paper.

Grounded or Guessing? LVLM Confidence Estimation via Blind-Image Contrastive Ranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-19T17:12:41.168062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T17:10:13.032423Z digest=sha256:bd60b5f1d70b1ed076834e949d09ec556de8780496f9787bafa48179ce336400

Observation 7e90ac0e-dfd1-4d35-b8dd-28d61f644204 · inbound

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards cites this paper.

When Self-Belief Misleads: Active Label Acquisition for Reinforcement Learning with Verifiable Rewards Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:04:01.363599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T22:58:46.313028Z digest=sha256:408e76ce1cc5af9914e87ea6013f3a96aa518f75823abe3eba3aeac6070d1853

Observation 6db0d441-f816-455a-b782-27e796065c49 · inbound

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking cites this paper.

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:40.955133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T16:59:25.127619Z digest=sha256:4f54217a3d3c886d9c660f8243026920a9b860f82317db1087bcc5737d33f078

Observation aaf993bc-7338-4ca7-a8f8-b87c3cca3282 · inbound

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling cites this paper.

Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 93

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.894680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T10:25:10.559953Z digest=sha256:4265c7fb3eeeaadbe530dbaa0f307cee6fe78631776823937325da906574a4cc

Observation ae8514c4-2188-4734-91ae-0a0933d89460 · inbound

When is Your LLM Steerable? cites this paper.

When is Your LLM Steerable? Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.591198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:01:21.682285Z digest=sha256:2a3dba1dbf51789cd9a158f5c2b2afab05d3a7289d66bd99de52f6e0c3257de6

Observation 9c75a31b-a1b5-4f76-93b5-bce026bddc92 · inbound

Dynamic Rollout Editing for Reducing Overthinking in RL-Trained Reasoning Models cites this paper.

Dynamic Rollout Editing for Reducing Overthinking in RL-Trained Reasoning Models Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-27T01:20:20.414019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T01:19:55.164835Z digest=sha256:ae93dd20ad35601eca92f9d8a71f588d0e1ddf63846cf23bdcb6035a2b7d7b51

Observation 1e05eb4e-4a2a-413b-874c-35d26c5d29d5 · inbound

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning cites this paper.

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 111

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:38:55.839168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T01:13:11.483599Z digest=sha256:2311e83af5297f014f02f58d68d9bfffb24b28e761c2f6e9e5a4dc843c52ff2a

Observation 3c41cc19-79b3-4f5b-baec-6e243190baac · inbound

Does the Same Token Mean the Same State? MoE Routing as Signal for Reasoning Control cites this paper.

Does the Same Token Mean the Same State? MoE Routing as Signal for Reasoning Control Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 27

Resolution
malformed identifier
arxiv_id, observed 2026-07-04T10:19:47.420429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T08:59:28.038685Z digest=sha256:0cd049919b402714ce24b70f56ebd5ee8a00127d340b0912dffc578d07be8213

Observation 4dfb2210-428c-4b74-aa3c-6a15235cbd3a · inbound

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents cites this paper.

When Agents Commit Too Soon: Diagnosing Premature Commitment in LLM Agents Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:49:45.696278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T08:32:24.949332Z digest=sha256:e590e21df0abcc40a4e8448ed0c543ea05b9d70437a4ad3ce97b8b375314acb1

Observation b90db3b7-ca9b-40f9-8560-2806f98cedc5 · inbound

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents cites this paper.

Plans Don't Persist: Why Context Management Is Load Bearing for LLM Agents Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:49:46.404977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T08:27:28.236073Z digest=sha256:792fe5de562b2cf5ec0b296ea4460fa08b161a0defb6a69e1f2c5e33753f5c17

Observation 6305eb4e-aec8-4061-8d2d-f359e683fe93 · inbound

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning cites this paper.

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T14:00:54.863716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T14:00:54.863716Z digest=sha256:3d7283806760be262e91b2dab45a22eb1bb6269defbf9fd68df0f356cb471c3c

Observation 499c5b30-9477-40a3-a998-f70e2869604d · inbound

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning cites this paper.

LLMs as a Jury: Cross-Model Consensus Can Outperform Process Reward Models for LLM Reasoning Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T07:28:23.769012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:28:23.769012Z digest=sha256:303e98136935cb419639f1f63627523cc00cd93ed5e1195c5d90fbbb6c98694d

Observation 8ca5437a-6d42-489d-be36-7413dc947a89 · inbound

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making cites this paper.

Multi-Head Latent Control: A Unified Interface for LLM Agent Decision Making Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T02:42:07.720174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:42:07.720174Z digest=sha256:2c2b872c5f178f83797618388c023fd5705e4f66adcac4c427cbe0ea1f1033c4

Observation acefe7b9-9aad-474f-8d33-c81a64081175 · inbound

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes cites this paper.

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes Reasoning Models Know When They're Right: Probing Hidden States for Self-Verification

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T07:49:36.297426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:49:36.297426Z digest=sha256:4dc2375007677cca56b948e0182427591de8b1ce09984e3f1124f399c3c6c2a4