Pith. sign in

Paper Citation Record · LEDGER

Unlocking Recursive Thinking of LLMs: Alignment via Refinement

As of 7 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2506.06009.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06009 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T06:07:02.121912Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved48
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 246ca00c-68a2-4c84-bd15-805de75a1b75 · outbound

This paper cites online" 'onlinestring :=.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.972708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.972708Z digest=sha256:b2fd2adcde9d90feac34bae89cfa2ca0c1e28b04c604e4b2e88260ce147c9a63

Observation 939f3c85-c716-43fe-821c-736597ed7817 · outbound

This paper cites write newline.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.976315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.976315Z digest=sha256:6d4154ebb5cb6e27feaeb57ffdb9dcba128a4e2a103c1ca36e2bf11931358fc3

Observation 6a23e769-4d86-4c64-9a8e-9603b5d2a836 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.979925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.979925Z digest=sha256:6ded455fabbcd9a05c63c13156f2cf6ac0fd7eaff94282979233a95b817c9547

Observation 47b801cc-3afb-4edc-95fa-edd60c593ab1 · outbound

This paper cites Critique-out-Loud Reward Models.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Critique-out-Loud Reward Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.983009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.983009Z digest=sha256:de4f06004957ad3e42286afde13caef5a921b147760258aac768db0542ae550a

Observation c55dfe77-2b6c-49ab-b566-62fab778e8b4 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.551702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:01.986290Z digest=sha256:c973bb83e5bd1fff31248191d259a3bc66d56dad800e06d2b23ee9aa09f6fb0b

Observation 6c04a003-9f19-46bf-b99d-54b7299b459f · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.989412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.989412Z digest=sha256:d8d038b3bd9f5a57da865010192b778732a7feb6366af84ac73fa9814b55b0ce

Observation 6839484a-01ee-4961-8192-6cd9c52c977e · outbound

This paper cites UltraFeedback: Boosting Language Models with Scaled AI Feedback.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement UltraFeedback: Boosting Language Models with Scaled AI Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.992586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.992586Z digest=sha256:6539df375b416e9578fcd411354ce107dd9f40a550483296808331dc94fbc7eb

Observation 7813797b-1ee3-4912-9539-35b925349318 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:01.995800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:01.995800Z digest=sha256:7528f32f4a75e70a9e42737bd73e11bacaf0c5739c08c6a70a5c1a4cd6fa9b12

Observation 380dbae0-c436-4b05-80cd-67d470103764 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.542765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:01.999001Z digest=sha256:f5fcfa9dd1c390f2c6b341ef775b3ab18e8e2b1b92bdcbac7c7c8f4dce2f2e7a

Observation 70eb2ecf-2649-4acf-a152-1efe251e6f61 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement KTO: Model Alignment as Prospect Theoretic Optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.002272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.002272Z digest=sha256:609850e23c329f235bcdd49db0b0088931dce380446d51c60f8f0b0465aa9af6

Observation ffc07461-ffdc-44a3-8845-da654ca7584f · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.005898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.005898Z digest=sha256:b01df24fdca840049241e5e0d5484d18f8f6618018fb6355e0318082b8e069ff

Observation 1deacb34-0eb5-4a6d-8403-2b305905c92a · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.533322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.009364Z digest=sha256:c2a4a094b30562518e901e73053287e35677abc0e76318a12a50e9bdcd03eae9

Observation 99d2a96c-a0a0-4b41-b672-bf9620c8ca44 · outbound

This paper cites SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.012071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.012071Z digest=sha256:88030485576beb028488d7775ea01d62c106b53a4c203380c32a4dc04059038c

Observation 6509816e-6a6d-47dd-a93c-8d7edb8b66f2 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.015165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.015165Z digest=sha256:681ab87989110cb9e06e38ccc9eacf4765d9197889d4f4cf9e18773dafd04cb7

Observation 204fcbc4-8020-47c9-b96e-6a8676ecf97c · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.517796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.018445Z digest=sha256:ceee8bd7ce2295cd6a00d0d81f8ae02ddcc1c70ef2639bbfdc8edb409161bdbb

Observation 705a9ac5-f8aa-43a1-acfa-18a3e9b4fbe4 · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Training Language Models to Self-Correct via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.021435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.021435Z digest=sha256:8dfb1de7a562c0f2d5b2afdd1d19db57fcfaf6fb04f840a1cbd1979f4f04d48e

Observation a7932206-d0a4-4c75-bb3e-9b990646251b · outbound

This paper cites From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement From Crowdsourced Data to High-Quality Benchmarks: Arena-Hard and BenchBuilder Pipeline

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.024636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.024636Z digest=sha256:a1e5d69a88c9868e2df6f838737a49f941181cdaf6b13f45c78700d9b8dd6b8a

Observation 9d099228-8a1c-4a64-9c96-40bce31657d7 · outbound

This paper cites Hashimoto.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Hashimoto

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.027559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.027559Z digest=sha256:68f8be09c1f59b62c207c9b4738ff7065d727ded6f2ff024d5406e8a116c183a

Observation 9ff1999d-45be-4193-9751-ae2c244ffcd8 · outbound

This paper cites Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Fennec: Fine-grained Language Model Evaluation and Correction Extended through Branching and Bridging

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T06:07:02.326602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.030197Z digest=sha256:7dd09180ff9dace685fe449651f363daeffdeb45b3a084467afd9e33dad123b5

Observation e9d85ba7-afeb-4fd9-a028-9647961ba4e8 · outbound

This paper cites Let's Verify Step by Step.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Let's Verify Step by Step

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.033102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.033102Z digest=sha256:485f3b1d3d34b0abae58e649c130bcafa059cd01600d8e51adf61d64fd6f8daf

Observation 71669695-99c8-4e6d-8853-be1d52f87454 · outbound

This paper cites Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.036124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.036124Z digest=sha256:6cf2578a60c64f1463a0eaad83554b15c62f3e906f9f9f0c04b0f701ce420747

Observation d449e7df-6801-4529-a0c1-68995e4c7361 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.039027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.039027Z digest=sha256:a6f7ecb773a699d92cf8aa1d896258b693e56ae04eff1313c11aeb64bff3e555

Observation cda1544f-8451-41de-9077-a474bf655e98 · outbound

This paper cites Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.041738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.041738Z digest=sha256:19ab2ad123f47d26de2707d850fcba64709633d75ca2c1c2e03c5c7b343775bb

Observation 48c893e3-c9f5-446b-931b-6ef14232d5c1 · outbound

This paper cites s1: Simple test-time scaling.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement s1: Simple test-time scaling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.044561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.044561Z digest=sha256:cd943178e38f76c60271e06f5b8bbb208c07f89bec4452aa9a6bab77279fb8c8

Observation dd036561-2dc0-4c25-ae10-2a2b251de0ae · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.497164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.047780Z digest=sha256:d1330939bdb3f31d519a18b0027504a3c61370c57ec6da953ded91b3cd679eb4

Observation 7faebdb5-bce8-4568-9819-96e8c32d9373 · outbound

This paper cites Disentangling Length from Quality in Direct Preference Optimization.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Disentangling Length from Quality in Direct Preference Optimization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.051648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.051648Z digest=sha256:1b469241b814007f7274b3358ff6b06dfdd04b77955cb5b4a23343594f579753

Observation 2d7ca1f7-6c45-49c7-8011-f26477e5e3b3 · outbound

This paper cites O1 Replication Journey: A Strategic Progress Report -- Part 1.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.055033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.055033Z digest=sha256:f436332a51a3a7e9b06d9aa2eefc9d5a607c6d5de41e29c92b11e978a05dbb67

Observation 240fdee3-e6fb-4e72-ab6c-d236343dbeff · outbound

This paper cites Recursive Introspection: Teaching Language Model Agents How to Self-Improve.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Recursive Introspection: Teaching Language Model Agents How to Self-Improve

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.058112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.058112Z digest=sha256:1de2677ea0c3084813f992670b9c342f5515a860d4fa751a46e451dd1d3a3844

Observation a0289a2f-f48d-4367-9d62-3fb1d7e8a19a · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.488026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.061385Z digest=sha256:3204cf875d0e06a2a0712ff3bfc4e6f90c64102b1bc3c5f80730a6af9833fd6e

Observation e3c7f2da-8a66-47c0-b9b9-cb3b863d71f2 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.064199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.064199Z digest=sha256:1b3013bffb6ed9dcbe641d40b2a4dd38fa73d7ddbb27015ab9e961944fc1bf0b

Observation 8d62d28c-59f8-4d8d-a541-ff514de9200d · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.066957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.066957Z digest=sha256:e547fd279e4e4b8a00c3a837c892934e0c8ff3b72b339451da31fc3ee0baa640

Observation 0f14bae8-f08e-4f62-ba2d-9cb358d8eb81 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.069605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.069605Z digest=sha256:11ee5b29393d6641c22da700c3ca443dfd722ee751892d180b09f5d4a8678e8c

Observation 429d58db-9a7f-433a-92e6-8cd9a2964ec3 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.072269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.072269Z digest=sha256:4e14bfe418a8bae1e7ec9f74b5d291f239a7b2c9b71a115d78ff412611e1d03b

Observation b414345b-aa49-4c5e-a267-65b1183c2d77 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.453762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.075154Z digest=sha256:8481f1ab2123674d4b17e9607bfc9aa2cd612e5a090e710a90d01f017bd5544d

Observation 7ca97041-5f81-428c-8287-3dc77c8239d5 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Proximal Policy Optimization Algorithms

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.078218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.078218Z digest=sha256:fc44556f6157bec68ff6a03d807407e36326e7c0601537071f31921ed16a60ff

Observation 3a8cd4f1-5e19-4997-bd32-5b4e9ae2bd3e · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.444742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.081251Z digest=sha256:ed89bad87d613b603f370685c74004244b568dc6289d00470237bec94afdf4bc

Observation ecba1e32-ca90-4164-998b-5c71dbe261b2 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.084050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.084050Z digest=sha256:abf4f0002f9fb44cf6482a4a0971de4e3c99d7ce0bfdcb7bbb304717a3feda7a

Observation 5e4046cb-f43d-485f-a977-3299fb9f0052 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T06:07:02.435618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.087091Z digest=sha256:3b24b1ab105d78b708d334a2175a78550eac319cfd9d7bba854210fa11059d57

Observation e5792110-ad88-4671-964a-a5b8859a5d18 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.089755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.089755Z digest=sha256:bffccaa964421e91558643bd2699c173a8ff08d1eb5e92eeffa7eac2510f7c7d

Observation bfb59995-fcf9-4ed4-a7c5-2584175072a5 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.092644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.092644Z digest=sha256:738162a7feff6374974378ae420ac6f2ccac4bc411db1ec8c0b9ceeff0847ce6

Observation 617291b0-867b-4816-97f9-e09c6666a867 · outbound

This paper cites DRT: Deep Reasoning Translation via Long Chain-of-Thought.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement DRT: Deep Reasoning Translation via Long Chain-of-Thought

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.095155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.095155Z digest=sha256:40049f2e4442a316bd8b30979d7d3a4108ca6aa1a62e57563d4dcbab5849f216

Observation 491b5ac9-a129-4783-ba7c-6607f705e9b3 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.098175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.098175Z digest=sha256:9ea2f26f66c9a67d3622bedf2a8fd9e8983edcaa89466effbe60d0c236184c31

Observation 95bf17a7-838b-45b7-9e5f-4411dfae52a6 · outbound

This paper cites PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement PandaLM: An Automatic Evaluation Benchmark for LLM Instruction Tuning Optimization

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.101272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.101272Z digest=sha256:6712e8bbeabb0178a815ac6df27598474ab704ba16617126d640c449849b19c6

Observation 76841248-9d69-4b5e-b366-47d44e28e2b8 · outbound

This paper cites Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.104060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.104060Z digest=sha256:482753863ec1fe7b81670c75962c65ad89c8f3076bee44619bff1171386bde4d

Observation 861754ea-92b0-4b77-ab47-7d256420d134 · outbound

This paper cites an unresolved cited work.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.107069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.107069Z digest=sha256:eae4c19431480ca26e430e4b1a3d6b4eddd28065f018d896dd3971c911e96503

Observation 922c3bed-baff-4497-b9b8-297001d79be1 · outbound

This paper cites Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.109745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.109745Z digest=sha256:874a4c1361918e6af8e4db474adb5ff990af3224174cb2250ffc9f7228f84573

Observation 25ec39e4-7325-4bf2-93cd-bd83a1ca71c3 · outbound

This paper cites Self-Rewarding Language Models.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Self-Rewarding Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.112916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.112916Z digest=sha256:e881bdd8ed9aa212e9f51bb8cca19fa9b8722da0d9a134688cf7bf38dd2ff871

Observation 33364e88-0301-43cb-82b7-a273cb02dc70 · outbound

This paper cites Understanding the Dark Side of LLMs' Intrinsic Self-Correction.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Understanding the Dark Side of LLMs' Intrinsic Self-Correction

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T06:07:02.175915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T06:07:02.115821Z digest=sha256:58f004267d17fcd1a06b7c118d82081b204704522572b78254a8e3e38ed1715a

Observation 19cb7ee1-9951-4eca-8ed0-a9302ad030b9 · outbound

This paper cites o1-Coder: an o1 Replication for Coding.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement o1-Coder: an o1 Replication for Coding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.118750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.118750Z digest=sha256:b4425eb5cfc03f05fdd6894569f3ede84e32b97ca2fa1710a499f35a98d649f1

Observation b1d0106d-abe3-4a86-b6f0-07b718b24a5d · outbound

This paper cites Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions.

Unlocking Recursive Thinking of LLMs: Alignment via Refinement Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T06:07:02.121912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:07:02.121912Z digest=sha256:72830e7e961c0b5f88926f6565dbaac482255031235ec7c1382a0ff031e1cbb9

Pith citing papers

No inbound Pith citation observations are available.