Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:07:02.121912Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2506.06009.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T06:07:02.121912Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 246ca00c-68a2-4c84-bd15-805de75a1b75 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 939f3c85-c716-43fe-821c-736597ed7817 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a23e769-4d86-4c64-9a8e-9603b5d2a836 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47b801cc-3afb-4edc-95fa-edd60c593ab1 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Critique-out-Loud Reward Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c55dfe77-2b6c-49ab-b566-62fab778e8b4 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c04a003-9f19-46bf-b99d-54b7299b459f · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6839484a-01ee-4961-8192-6cd9c52c977e · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement UltraFeedback: Boosting Language Models with Scaled AI Feedback
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7813797b-1ee3-4912-9539-35b925349318 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 380dbae0-c436-4b05-80cd-67d470103764 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70eb2ecf-2649-4acf-a152-1efe251e6f61 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement KTO: Model Alignment as Prospect Theoretic Optimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffc07461-ffdc-44a3-8845-da654ca7584f · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1deacb34-0eb5-4a6d-8403-2b305905c92a · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99d2a96c-a0a0-4b41-b672-bf9620c8ca44 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6509816e-6a6d-47dd-a93c-8d7edb8b66f2 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 204fcbc4-8020-47c9-b96e-6a8676ecf97c · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 705a9ac5-f8aa-43a1-acfa-18a3e9b4fbe4 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Training Language Models to Self-Correct via Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7932206-d0a4-4c75-bb3e-9b990646251b · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d099228-8a1c-4a64-9c96-40bce31657d7 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Hashimoto
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ff1999d-45be-4193-9751-ae2c244ffcd8 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9d85ba7-afeb-4fd9-a028-9647961ba4e8 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Let's Verify Step by Step
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71669695-99c8-4e6d-8853-be1d52f87454 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d449e7df-6801-4529-a0c1-68995e4c7361 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda1544f-8451-41de-9077-a474bf655e98 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48c893e3-c9f5-446b-931b-6ef14232d5c1 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement s1: Simple test-time scaling
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd036561-2dc0-4c25-ae10-2a2b251de0ae · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7faebdb5-bce8-4568-9819-96e8c32d9373 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Disentangling Length from Quality in Direct Preference Optimization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d7ca1f7-6c45-49c7-8011-f26477e5e3b3 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement O1 Replication Journey: A Strategic Progress Report -- Part 1
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 240fdee3-e6fb-4e72-ab6c-d236343dbeff · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Recursive Introspection: Teaching Language Model Agents How to Self-Improve
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0289a2f-f48d-4367-9d62-3fb1d7e8a19a · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3c7f2da-8a66-47c0-b9b9-cb3b863d71f2 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d62d28c-59f8-4d8d-a541-ff514de9200d · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f14bae8-f08e-4f62-ba2d-9cb358d8eb81 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 429d58db-9a7f-433a-92e6-8cd9a2964ec3 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b414345b-aa49-4c5e-a267-65b1183c2d77 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ca97041-5f81-428c-8287-3dc77c8239d5 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Proximal Policy Optimization Algorithms
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a8cd4f1-5e19-4997-bd32-5b4e9ae2bd3e · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ecba1e32-ca90-4164-998b-5c71dbe261b2 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e4046cb-f43d-485f-a977-3299fb9f0052 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5792110-ad88-4671-964a-a5b8859a5d18 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfb59995-fcf9-4ed4-a7c5-2584175072a5 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 617291b0-867b-4816-97f9-e09c6666a867 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement DRT: Deep Reasoning Translation via Long Chain-of-Thought
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 491b5ac9-a129-4783-ba7c-6607f705e9b3 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95bf17a7-838b-45b7-9e5f-4411dfae52a6 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76841248-9d69-4b5e-b366-47d44e28e2b8 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 861754ea-92b0-4b77-ab47-7d256420d134 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 922c3bed-baff-4497-b9b8-297001d79be1 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ec39e4-7325-4bf2-93cd-bd83a1ca71c3 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Self-Rewarding Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33364e88-0301-43cb-82b7-a273cb02dc70 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Understanding the Dark Side of LLMs' Intrinsic Self-Correction
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19cb7ee1-9951-4eca-8ed0-a9302ad030b9 · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement o1-Coder: an o1 Replication for Coding
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1d0106d-abe3-4a86-b6f0-07b718b24a5d · outbound
Unlocking Recursive Thinking of LLMs: Alignment via Refinement Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.