Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2412.14135.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T23:33:52.937067Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 647261e7-faf9-46f4-8233-cd650ecf5e9e · inbound
HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ff67168c-269a-401b-ae22-67cffee03c28 · inbound
Search-o1: Agentic Search-Enhanced Large Reasoning Models Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 454dedb2-e3f9-4e64-9a26-2b809b0efec5 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e9b1f88b-5d58-4037-921a-64db6211d813 · inbound
Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3372b021-128e-463a-9bf3-b4508077d489 · inbound
TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1b73dc4-bc02-4a79-bf8c-24778fe13642 · inbound
Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 266
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9ffcf2c6-bd94-470f-b32f-550aa083dcb3 · inbound
Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 251
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 58f25752-65fa-4ac8-bbaa-43dd3fcbd1fa · inbound
CAT: Confidence-Adaptive Thinking for Efficient Reasoning of Large Reasoning Models Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.