Pith. sign in

Paper Citation Record · LEDGER

A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2505.02665.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.02665 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:06.133968Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 06417ac6-4910-4b6a-9e4d-268a7dda7ebb · inbound

Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training cites this paper.

Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:06.133968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:36:06.133968Z digest=sha256:45fac2fe82c268b8e8c5270bed45133f5ab12b5b22a9994e48b1685897b7d12b

Observation 25187283-2789-43d2-bb10-3d4dbaa2f15a · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:40.984595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:40.984595Z digest=sha256:b36fecc4fd9354433c27a4b45f2ea8334d3c67cc4689fac9ac1e3758e928ed32

Observation 84b62c6d-43ad-400b-a5bc-797323656475 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 146

Resolution
unresolved
no resolver link, observed 2026-08-06T17:54:17.558495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:54:17.558495Z digest=sha256:c0fd056eb4535b4e906ebf8efcd00550602dba046ae0811e311f901989094a65

Observation 8a3c3de1-12c5-48fe-a1f7-78a51acc03f4 · inbound

Controllable LLM Reasoning via Sparse Autoencoder-Based Steering cites this paper.

Controllable LLM Reasoning via Sparse Autoencoder-Based Steering A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T12:20:59.992436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T12:20:59.992436Z digest=sha256:8bc005743cac81c3dc526657a526e95bef3a0af756880fdc573324e8288482f4

Observation 97d1e28f-2a9e-4715-8d16-eaea9d63518f · inbound

On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective cites this paper.

On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:13:58.707550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:09:37.588841Z digest=sha256:c3c5a0ea2c5bfcefbe4aab7e0e0a76e35d60eb09084f7d8fc83f2d76d56ef119

Observation 6a01c40b-8bf6-4f19-83c8-8c80039401c0 · inbound

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation cites this paper.

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:46:46.314216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T06:40:12.178170Z digest=sha256:16cbffc7830c5a6a062067f4db0065c43e4418c30ce5ca6bb45f487d56db993b

Observation 41f03b39-a909-424e-ac8b-d561e8d3ac7c · inbound

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation cites this paper.

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-28T06:41:42.778580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T06:40:12.178170Z digest=sha256:6eb96462edc3e4b5f32c2933a987fca0b051d7d0b86e9f742ff7057836435689

Observation 13dcd7ac-5c18-4264-8acd-293dd104c7f3 · inbound

A First-Principles Theory of Slow Thinking and Active Perception cites this paper.

A First-Principles Theory of Slow Thinking and Active Perception A Survey of Slow Thinking-based Reasoning LLMs using Reinforced Learning and Inference-time Scaling Law

Reference 122

Resolution
verified exact
local_arxiv, observed 2026-07-10T11:37:03.240425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-10T11:32:24.374377Z digest=sha256:0aa0c86504a4baad3228e07bc1c1dedfdfa1e578def21a24778ed6cbe0841968