Pith. sign in

Paper Citation Record · LEDGER

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

As of 8 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 5 inbound Pith citation observations for arXiv:2502.06773.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06773 v1

Coverage vector

measured 87 of 87 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:25:53.635679Z

measured 92 of 92 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:41.579717Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:57:41.592060Z

Reference resolution

87 of 87 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved66
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ecc36f18-1e9b-40fe-93ff-1e108ea1af3a · outbound

This paper cites Phi-4 Technical Report.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Phi-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.261354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.261354Z digest=sha256:a24966165fb095805aa14c586a20579c9b427ba8dfad17207df984b508a80e69

Observation b582b0e9-a95e-4d3e-9f94-0fb09582e460 · outbound

This paper cites Amc 2023.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Amc 2023

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.267178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.267178Z digest=sha256:054772e1908d43b772d83445cb9a0b518ecafa4541e8736bc1835f40447468ab

Observation eece90db-4388-4ddf-9c31-7642bde6225a · outbound

This paper cites Aime 2024.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Aime 2024

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.271469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.271469Z digest=sha256:d199e061aca0b3ba59d50b039967a9b033f4b1e22ed1f618eae19b487f1cbcde

Observation b5f48994-4ec5-4212-9c7f-2cefbe971e23 · outbound

This paper cites Numinamath-cot.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Numinamath-cot

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.275821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.275821Z digest=sha256:bf42594653f39bb591dbad0dc7d4ba92ce42e09e2ba7f99f313efac6e0eb8a99

Observation 6a10645d-8bfd-4e75-b884-8f82f92de4b7 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.280245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.280245Z digest=sha256:851e09acaa7fd8fbd362ba04919498698086aae0f3d679a18745c3188be84bd3

Observation c281d1d5-4ccc-459e-985e-eb611e13f7f9 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Constitutional AI: Harmlessness from AI Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.284895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.284895Z digest=sha256:73bef441d66d34cee3df95f5178335bbd8326f87d3fd9f7af6ec297949321bcd

Observation 1082ab03-ef0f-4ca1-97a8-9f0e7f0cbfcb · outbound

This paper cites Scaling test-time compute with open models, 2024.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Scaling test-time compute with open models, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.290041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.290041Z digest=sha256:00d155528e6a751539c482bf9af5d359ed60874bb364c907f4f57127cda7cda6

Observation 0b217de4-d4a0-4bb4-ad06-04e7fdc930c6 · outbound

This paper cites Open-r1: a fully open reproduction of deepseek-r1.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Open-r1: a fully open reproduction of deepseek-r1

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.294520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.294520Z digest=sha256:491a63ca7e45f45e7f346ed365268bd11c1c4712933e5d4b58f99b3f28ee74b2

Observation 82bdc4bf-793f-4cd1-ad2b-fd0c52762730 · outbound

This paper cites LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.299016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.299016Z digest=sha256:3f44181525632d289425c5f074f5e4d3183104aea824c3cd8dbed820c7727f2e

Observation 3357ea57-02a9-44c2-937e-6a19461ad863 · outbound

This paper cites Self-Improving Robust Preference Optimization.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Self-Improving Robust Preference Optimization

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.303921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.303921Z digest=sha256:b4dc0513eccac553c14f5e4ab63e98921873dc2ab6b01b88424dfa0545074a76

Observation 2402dfe1-17ca-45d9-8aa6-e55610864943 · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition AlphaMath Almost Zero: Process Supervision without Process

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.308231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.308231Z digest=sha256:d4158bcce979b2745ecf221deb894ddcf4761f987cb61d64e6b443105e855d41

Observation ce1aee1b-0de3-4e4b-9af5-2daeb8862248 · outbound

This paper cites Codeforces dataset.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Codeforces dataset

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.312356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.312356Z digest=sha256:b4ebb6e9128e44867b8753d5fc76ab006763f7a2fbeba0b3ac832c15b067f279

Observation 1e9f82f3-63a8-4862-bab1-3583f27d8217 · outbound

This paper cites Process reinforcement through implicit rewards.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Process reinforcement through implicit rewards

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.316251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.316251Z digest=sha256:e749af57db8215e8342af273154bd1aba36241fcd9ce6535a77c707b7b37e707

Observation b2b9e052-b993-479d-b1e8-f235fbb6df7b · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.891480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.320253Z digest=sha256:974554271604bccfd399895c89048b92940c4d305e003db18b6d3a9c3caeefa9

Observation 51617f6d-02d7-4ade-b1ae-f7ee7d57f972 · outbound

This paper cites Flash A ttention-2: Faster attention with better parallelism and work partitioning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Flash A ttention-2: Faster attention with better parallelism and work partitioning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.875355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.324290Z digest=sha256:dea7b034c788db80ce9b99a2b4977e4c6371085cb4bb8d00ff325f4a1733dc2b

Observation c4041b13-e1a1-492a-9fae-53fd468c599b · outbound

This paper cites Ai achieves silver-medal standard solving international mathematical olympiad problems, 2024.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Ai achieves silver-medal standard solving international mathematical olympiad problems, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.860648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.328130Z digest=sha256:bcee7edea972ecd523a01b9241c8c6ba25a0aaa807041462064d1f6fc585e824

Observation bdd185cc-189c-4585-a351-a8aeff4a8531 · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.332278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.332278Z digest=sha256:fec3692da575786cb3e486ea49fc68f44c7c37100e94fef7d0d71f22920a5bf1

Observation 56005d59-b58d-4a7d-a27f-8341ea7a820f · outbound

This paper cites Introducing gemini 2.0: our new ai model for the agentic era.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Introducing gemini 2.0: our new ai model for the agentic era

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.846667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.336397Z digest=sha256:0b45f1201913daa4c1605a4d0d91952a66b8e79bdd13991e08cd44b024fa735b

Observation c7e1636c-8b03-4263-91b1-ae5239e0bc6a · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.340082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.340082Z digest=sha256:ea1d21025ef3e55b4bd776ccff051caf5dc52386ede7eb4fcdd09d74b2b682bf

Observation 6d640c67-b92c-4e0d-9e83-c349fecca3a1 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Measuring Massive Multitask Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.344239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.344239Z digest=sha256:b1f5f953c6a23c3113f9e0ee77b1361a5f333084eab130e2b7f9cfffd0237063

Observation 388a53cd-7a7e-43c0-90f1-bcc3a536fd76 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Measuring Mathematical Problem Solving With the MATH Dataset

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.348541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.348541Z digest=sha256:c980fda3e3bfcd0ea0cd20dec1d2fd74f133e814a31f2e59fcb052e2f24a478c

Observation c2702d0c-f576-4c21-8bb6-35cca6a3d8e0 · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Large Language Models Cannot Self-Correct Reasoning Yet

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.352790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.352790Z digest=sha256:5fad971ddc971a9c41249d84a1895111e95c7b2100d54c781aec4f257fdbbfc3

Observation 943e4545-6612-4952-a7c6-53c23682a60a · outbound

This paper cites Teaching Large Language Models to Reason with Reinforcement Learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Teaching Large Language Models to Reason with Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.357179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.357179Z digest=sha256:1c28efb2457799941118efe1bc52bf5822b537a2972d6ea18401719116bc2c60

Observation a28b3f64-5a69-4c93-831e-8e7869e8ff38 · outbound

This paper cites O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey -- Part 3: Inference-time Scaling for Medical Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.361801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.361801Z digest=sha256:a768fd002678433ff14517a014cadcbde4217e28dd55783ae497a79d9208a37c

Observation b1e892ab-f379-458d-9888-707cb8c2c758 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Reasoning with Language Model is Planning with World Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.366490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.366490Z digest=sha256:3c488cb08e854040b107753b66f277b96ce78049ae5b69a54df313aa1bacce13

Observation 5c92a674-1136-4135-92ae-a8c9385107c5 · outbound

This paper cites GPT-4o System Card.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition GPT-4o System Card

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.370738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.370738Z digest=sha256:baa464612efae261feb274cadc6e4d14a2f9d855e1768e36b31aad384eb9bf65

Observation 89738edf-e17b-4fb8-aa32-0e92980a4dfa · outbound

This paper cites T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.374934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.374934Z digest=sha256:e6b864ac7bf6f3ff0f9a3dcdc67ba02a5f6c595820d1f241e543061ec598c5e8

Observation 71c91751-5438-45e2-bc14-697b74db75eb · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.379420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.379420Z digest=sha256:861e0b26db4381d61c446df7b079c68beb3ef885477051b2df2461aae9dc94d8

Observation 6b593b83-c156-43af-984f-0629a2a6616e · outbound

This paper cites O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.383766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.383766Z digest=sha256:7ab9372079b793a74eaec6c3b29c63a13ea9df2f75f14cfef555f550a6a1734d

Observation 9a181384-00a6-44af-a6d7-0d6ebdc467e3 · outbound

This paper cites Enhancing LLM Reasoning with Reward-guided Tree Search.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.387950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.387950Z digest=sha256:e052c6c77acdc285d9de4359f07a73717e336ce646e204f43aec8b1a37011486

Observation 03b4dae8-94e7-44d4-935e-32a759b793f1 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.392275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.392275Z digest=sha256:23ae14c3b2e424d4f07dfff47557a05e9d6b58c74c486cbed527f2b1b960a725

Observation a9ce09a3-0a28-4bbe-baa8-a3c690682860 · outbound

This paper cites OpenAI o1 System Card.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition OpenAI o1 System Card

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.396401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.396401Z digest=sha256:7ee76122eccc805d4792085e62b35c33b55734ccb88de083cc43c9420172b584

Observation 9fd80777-38eb-4813-b345-735b42bb596c · outbound

This paper cites Reinforcement Learning with Unsupervised Auxiliary Tasks.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Reinforcement Learning with Unsupervised Auxiliary Tasks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.401082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.401082Z digest=sha256:2816c41b08f090db8b14457aaae9c736c59db8f205eb006f8e5cf4d4a7957d85

Observation 9101808f-b24d-40c7-8fe0-e762b331cb81 · outbound

This paper cites Kimi k1.5: Scaling reinforcement learning with llms.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Kimi k1.5: Scaling reinforcement learning with llms

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.832565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.405463Z digest=sha256:50c1cd56950f7a90f961777a9bdcec059593728f6385caad7a90d48cd2eefb1f

Observation 79363580-8c3b-4201-904c-95637ea29156 · outbound

This paper cites MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition MindStar: Enhancing Math Reasoning in Pre-trained LLMs at Inference Time

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.409568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.409568Z digest=sha256:94b3d2a900be9ade91b113447b042efb4c37360da8b5df9ea2192185ce2610c8

Observation ab2f45c7-cd1e-4198-985a-51258f2f78d0 · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Training Language Models to Self-Correct via Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.413796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.413796Z digest=sha256:23fcb13da44ae4764b0a3de0cc30f13f5a2c61b1fca5df821f6f6a3d1ab6bde5

Observation 30c427f1-2c02-460d-8113-97c8523d761f · outbound

This paper cites Numinamath: The largest public dataset in ai4maths with 860k pairs of competition math problems and solutions.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Numinamath: The largest public dataset in ai4maths with 860k pairs of competition math problems and solutions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.818246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.418178Z digest=sha256:ff2d6beb34aa6b256161d52ae3885995f98bab0e65b98a07c54f0876962c5c44

Observation dd0171c7-9298-42b6-be98-39bf2838d347 · outbound

This paper cites Let's Verify Step by Step.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Let's Verify Step by Step

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.422225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.422225Z digest=sha256:9265dbacbf0d261aead000fdf648f630628d57c71dcd78fff95106922ffb443c

Observation a58e83aa-0365-45a6-87bb-477e0bcc5451 · outbound

This paper cites Chain of Thought Empowers Transformers to Solve Inherently Serial Problems.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Chain of Thought Empowers Transformers to Solve Inherently Serial Problems

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.426639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.426639Z digest=sha256:04bfa411012b3b4a9aa5bd72423dff4d352b63d5ca6d71674e30bd0561ea0dc1

Observation fe092817-e977-4101-87d0-9c9ab5f6a8fd · outbound

This paper cites WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition WizardMath: Empowering Mathematical Reasoning for Large Language Models via Reinforced Evol-Instruct

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.430999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.430999Z digest=sha256:15072eff1a440f2923473edf63a1f9edc66262436450acf510f1c0476156a145

Observation 85937927-b0c4-406b-bc68-afea3a066a6a · outbound

This paper cites Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.435670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.435670Z digest=sha256:315e7bfe292937a808a7c38f6d7584c5a44f40873e9365e406ac2e430a883de5

Observation 6ad11a03-cdce-412c-842b-a4a40d04aaa1 · outbound

This paper cites Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.439518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.439518Z digest=sha256:055ffb05d76d3be971db5a4d7bed15720581372b3dec9e90b35c6e73ef358804

Observation 15b8f081-e7f6-412a-8cb6-8426317c24a6 · outbound

This paper cites Llama-3.1-8b.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Llama-3.1-8b

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.793295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.443750Z digest=sha256:0e063c59303f9f11c83661c04c7525b85193b990111626bd72bdad86be1d177c

Observation b208dd9e-f17c-4bc1-a706-1b09b8c00692 · outbound

This paper cites Ray: A distributed framework for emerging \ AI \ applications.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Ray: A distributed framework for emerging \ AI \ applications

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.779383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.447529Z digest=sha256:74b52be4a7118d3d6227816ecc5217fdfdb2ed1545f9c11bd1eda68d3aafa9a8

Observation ac323491-468e-4ade-8b7c-7ae2a9fea4f6 · outbound

This paper cites The Expressive Power of Transformers with Chain of Thought.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition The Expressive Power of Transformers with Chain of Thought

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.451451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.451451Z digest=sha256:88e05e93231307c634076e34e0bb53a156f347eb8160682a8d044460d226c63f

Observation b2b5d512-62b9-456e-893a-26c3ce31dee5 · outbound

This paper cites s1: Simple test-time scaling.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition s1: Simple test-time scaling

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.455635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.455635Z digest=sha256:1920510406a8cd4fc9abb935c1bf195a8bc2e8c939dd7ef25814b83421f9e8f2

Observation e3e9fde9-2d69-418b-9f3b-12c5c3db937a · outbound

This paper cites Sky-t1: Train your own o1 preview model within \ 450.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Sky-t1: Train your own o1 preview model within \ 450

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.764289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.459617Z digest=sha256:2b43313a7a17c87b52da16cd81449996c71f6a234c52785280ea5b4366f73772

Observation b7089b6b-670b-4e80-a58d-dc265ff58890 · outbound

This paper cites an unresolved cited work.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:25:54.748210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.463706Z digest=sha256:e46f80749a96b56d19e8f0bc0ea8d49f123803883d55028deee8e1310071a719

Observation 72bc56d0-1ef5-4f09-ac0f-67b5db078475 · outbound

This paper cites Learning to reason with llms.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Learning to reason with llms

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.733808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.468071Z digest=sha256:6c02c57c61ec7722fa473d382beb080853835ee1c8c1a8943475c9eccb3aaa6b

Observation a0a00aeb-efcb-44cc-bc5b-b1de23dd640a · outbound

This paper cites Math-500.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Math-500

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.719735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.472120Z digest=sha256:384b3cf159d918be2d6a1245915fa1e41c334316a2be33fc0e9fcb1e09c13c54

Observation dce3a3b1-830d-44f5-a017-d2b828112f30 · outbound

This paper cites Openai humaneval.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Openai humaneval

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.705136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.478200Z digest=sha256:46a1c2c3e857183cebab985c6e4c18895b56c9911767a76360cda7f838535abf

Observation ecc12a85-1bea-422a-a208-8604b0edfa99 · outbound

This paper cites Openai o1-mini advancing cost-efficient reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Openai o1-mini advancing cost-efficient reasoning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.691110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.482502Z digest=sha256:07f48a36d2ee80f366bbe449d49652c0383fd7a572b46856f8c513a93b6f5d74

Observation eb58fd79-d437-4c38-b627-6b5ec21ecaad · outbound

This paper cites Training language models to follow instructions with human feedback.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Training language models to follow instructions with human feedback

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.486672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.486672Z digest=sha256:6a8b0b51e22acb5adf94eb7bfd89ebe0ab2daaaf54151cfdbb8f687653c4cbac

Observation aa5ce6f6-2465-4fbc-9d8a-af09777caf75 · outbound

This paper cites O1 Replication Journey: A Strategic Progress Report -- Part 1.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.490904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.490904Z digest=sha256:13ae0a4dcd8930b70853a44b46543951bb0bae3d5d7ddbf48c159d58a258c447

Observation 8b4943d0-f8c2-4144-99f0-d7676e972f86 · outbound

This paper cites Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solvers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.496098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.496098Z digest=sha256:74d7dacd3f300c2496c5e879017c70a61c8a70a643cfcab2e2e895431d8c77dd

Observation c1a0fb65-ef3a-426b-8805-2aef8b147292 · outbound

This paper cites Qwen-2.5-32b.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwen-2.5-32b

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.665820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.500841Z digest=sha256:ac1c2abc6b711b3d2382795867f3bd598bae0ae677d8d95f46117c520db0eed0

Observation 1283d6a6-60a1-46c0-ad8d-ccc8592df38b · outbound

This paper cites Qwq-32b-preview.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwq-32b-preview

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.651382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.506146Z digest=sha256:079c554620b01317309d47f770dd34a76ed72ea321fed6391c4f3538c7ad7abb

Observation 8ade35d2-357e-4fcc-906c-02673a1a5888 · outbound

This paper cites Qwq-longcot-130k-cleaned.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwq-longcot-130k-cleaned

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.635311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.510751Z digest=sha256:f184989951c211dab1bf4e623294857e94e38af7599637cf1f746684be426ff3

Observation b67dcc04-5003-4918-a1a2-cff852db7ea0 · outbound

This paper cites Qwq: Reflect deeply on the boundaries of the unknown.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwq: Reflect deeply on the boundaries of the unknown

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.620823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.515039Z digest=sha256:338cb1e23642ffc6f65b6ec5a7c2effb6a552a8296b17820d743c1c458b5648b

Observation 41f5cea6-2626-484a-90ac-1707ef95cbc1 · outbound

This paper cites Recursive Introspection: Teaching Language Model Agents How to Self-Improve.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Recursive Introspection: Teaching Language Model Agents How to Self-Improve

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.519496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.519496Z digest=sha256:c9290cef4851eb6f4de93d64782e0d01ddbd2a4ed7487945cf3305244766c9bc

Observation 73ac72b7-cd88-409a-8e9e-983c7e3826ee · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.524212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.524212Z digest=sha256:0a680d7681ae68f33f3cce95d814209c678074da6eab90ffcbc43593b6d46c2c

Observation 118b5f30-a0b3-47a3-8ce1-de26512b8c3e · outbound

This paper cites A program for the machine translation of natural languages.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition A program for the machine translation of natural languages

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.606496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.528679Z digest=sha256:f817a8361581f9cff755bb6fed6d6118e9a6dc56d01e9e777db5036088ece5fe

Observation 1a275c71-3f7d-4a64-b783-e466d9d6137f · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.532654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.532654Z digest=sha256:b151b68798fdaf759d45920604063936942ce1c164f2667895d7d39e588402f4

Observation 46209a82-f91d-419f-94a7-731aeaee85d3 · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.536838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.536838Z digest=sha256:e453902213a827364f8f719e9c59fdefd2629da8d93bfcebeac1631b3659596b

Observation b8806b55-4f4f-45fc-819c-638cf7f0127a · outbound

This paper cites Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Dualformer: Controllable Fast and Slow Thinking by Learning with Randomized Reasoning Traces

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.541676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.541676Z digest=sha256:1ae35f3e67f2990bd70d0d9aefa234c60e4dc95b61182e34c8e381d16b3c3973

Observation 568c3934-2dd7-4746-81f8-88f9b644773d · outbound

This paper cites Proximal Policy Optimization Algorithms.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Proximal Policy Optimization Algorithms

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.545928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.545928Z digest=sha256:cf7710b1fec2ad8a6f318cd16d343165ef9db24477f7e76cc9fafad1ba59714e

Observation 400e6bbb-f577-4f4a-91de-f675b930a8bf · outbound

This paper cites Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.550735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.550735Z digest=sha256:acdaab94f5db2feacdd745d3e0ba0d193ac5eba6899e8c8495d38dad6e5fcda9

Observation 26cbd230-74a1-45e0-90b6-77d1584b1b29 · outbound

This paper cites Solving olympiad geometry without human demonstrations.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Solving olympiad geometry without human demonstrations

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.591672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.554997Z digest=sha256:3bac650f7a6fb6b640c19f42466a56928c6b3e01b345e1acc2d9db6a37eee501

Observation c967d194-9948-4029-b6f7-2b1915b48a62 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Solving math word problems with process- and outcome-based feedback

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.558980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.558980Z digest=sha256:bc738b258863ae8d46367cdd6befd93c4b93c7740d765d30e20b167419fb499b

Observation 078bb60e-14d7-4bdd-8e97-d4c805a86e85 · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.563368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.563368Z digest=sha256:f36ddbec35667d57f8cb782c5b36ad013a80ac2e81cd1aa3f60dc184adfb0f4b

Observation 02e6b3c6-4fce-4229-bd2a-91e10e91dc0b · outbound

This paper cites Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.567513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.567513Z digest=sha256:a42a8beaf8b8b69cd02e8f33f9c3761b699849ac26c2319c3a752ca5b8f11143

Observation 8889cdf4-a380-4569-9c9b-92db64076048 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.571619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.571619Z digest=sha256:a2fd5538d190c78d3cb360d1f839a8a815eca1686c77aa546d127a832fb9376c

Observation 8c7e0202-04a5-4727-a56e-85f8bf1c08e8 · outbound

This paper cites An empirical analysis of compute-optimal inference for problem-solving with language models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition An empirical analysis of compute-optimal inference for problem-solving with language models

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.576399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.576225Z digest=sha256:00ed0e83a8cccc4c0bb0da613c16d16bb1f364aedbc1c03464b0d706b0ed1224

Observation 490f703c-c4cf-40e5-a257-a31153a9bfa5 · outbound

This paper cites Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.580233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.580233Z digest=sha256:3d97f29bbe6983a818c5db0dbd40899407e98deab68b9eeeb8cdfd450e247baf

Observation dc1766c5-126c-43cc-9c11-535b88589061 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.584409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.584409Z digest=sha256:889528bed9a7e22b82f45ab6af56c2f1f79622e8d854b248f66e95bcad7ec3cf

Observation 1fae56f8-5d36-4880-b537-d2c5179ab3c5 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.588274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.588274Z digest=sha256:3dd3ccc5d47f163b647e30a6c913c572480e586d55232ad479b1f1d42f323ff7

Observation 04a629f7-ff2c-4e73-a168-ca0f29f98e55 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.592132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.592132Z digest=sha256:7709d7a53c5cea0deb9d35dc5f19a7d6f5d234c71f3bd17a973aae0a7d6a0de2

Observation 62e15922-f60b-4ea6-811f-e6b923915a3d · outbound

This paper cites MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.596276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.596276Z digest=sha256:3db2d0385a8645297b5f7706fd9960173fe0e7b7f1ceb4ea5bf2736fe209b46c

Observation 9754a7d2-d866-4e3e-9803-e7d1ff172adf · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.600448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.600448Z digest=sha256:c0def8c15ae33edd910487565fd0ca1772d83f40fd19f5ce075c0d9a9eeeab55

Observation 6a28b955-a2f4-4089-9f84-33f32dc4962a · outbound

This paper cites Scaling Relationship on Learning Mathematical Reasoning with Large Language Models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Scaling Relationship on Learning Mathematical Reasoning with Large Language Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.604965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.604965Z digest=sha256:0f141fac92bec177c50bcf8ddf85363042a03c765f1968d1abb35880684ec79b

Observation 47af6924-d35a-47f8-b9c0-07fe3f934934 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Tree of thoughts: Deliberate problem solving with large language models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.609543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.609543Z digest=sha256:23b307b08c67b503a8dd38972ca0a7f34a456082b21035948c67f3ffe407cf5e

Observation 1049b08b-fa90-4bb7-b99c-ebca1098589a · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.613796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.613796Z digest=sha256:6735b0bf44135b1281473e03ed04cbbed80cb948c91ac66d90bb3a8ab69e6c9b

Observation a6936156-37c3-4b86-a63d-0d780164f52d · outbound

This paper cites 7b model and 8k examples: Emerging reasoning with reinforcement learning is both effective and efficient.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition 7b model and 8k examples: Emerging reasoning with reinforcement learning is both effective and efficient

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:25:54.549628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T14:25:53.618399Z digest=sha256:ee0fd0ef1e96ef9924a5fbefb30aba6b4216ddc03ce8bd6240411761428abc1b

Observation 0f9bad28-0108-4b8f-bf7c-123fe2271a93 · outbound

This paper cites Small Language Models Need Strong Verifiers to Self-Correct Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Small Language Models Need Strong Verifiers to Self-Correct Reasoning

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.622557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.622557Z digest=sha256:7c4cf87d770c5c35d2ca32d53bc51b97aae62ed00a346921ab058b526a2a4b07

Observation 388520e0-f5bf-489f-a245-0ef8e55cbf75 · outbound

This paper cites Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.627005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.627005Z digest=sha256:c945874be7202a998d4de19268415858923d5038d8c94e41f81ec565733f3907

Observation 407f26e5-679b-458a-860b-e946dfad7d99 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.631264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.631264Z digest=sha256:6bcab896dcfb22a60b6205b74319a025c346dd698c55d1a30d8994b6ab224286

Observation e4199f7f-f4b0-424c-8252-8677147d40fd · outbound

This paper cites ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.635679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.635679Z digest=sha256:fec0dced225f08db1a880c586c1273d5e41c5051dbb91e49dfa00c6c5b43c6df

Pith citing papers

Observation 9c08c4c8-1eff-47cd-8605-016f003dce04 · inbound

Phi-4-reasoning Technical Report cites this paper.

Phi-4-reasoning Technical Report On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:40:25.854460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T03:40:25.706499Z digest=sha256:63ee6966e8deb1538d1c8b4bcf4a6cd43c2b57d8337febf81fb13bc46c16180f

Observation a41acd6a-a1c2-4b54-8f11-fec289847cce · inbound

LLM-First Search: Self-Guided Exploration of the Solution Space cites this paper.

LLM-First Search: Self-Guided Exploration of the Solution Space On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:41.579717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:41.579717Z digest=sha256:e291f978704b7869bd265760f2bc899aa1abe45688dadf57807c0271a5ead2d0

Observation 0ed703d5-4526-4335-9935-6d8a042fb6da · inbound

Reasoning-Finetuning Repurposes Latent Representations in Base Models cites this paper.

Reasoning-Finetuning Repurposes Latent Representations in Base Models On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:50:00.873671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:50:00.873671Z digest=sha256:3c524ad8c9ab4b65d45ea2de617979e39610495eff5e2ea07a82a2a252ee32cb

Observation 83d99c6c-72e6-4eff-8a06-18fbb0536bef · inbound

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete cites this paper.

Rethinking the Role of Positional Encoding: Sliding-Window Transformers without PE Remain Turing Complete On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T22:06:16.178705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-28T15:48:48.046003Z digest=sha256:9e272a3e22c882354d745ff36e775bfe3d99edef0a681c996cefa3059249f884

Observation 7b5fd31f-e1af-469f-91b7-bb688a19f061 · inbound

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes cites this paper.

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes On the Emergence of Thinking in LLMs I: Searching for the Right Intuition

Reference 290

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:57:41.593548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T12:59:51.091008Z digest=sha256:e7584e931aea3d1ed2991b12b459d861326f425b4144df8a49aa601349d25136