Pith. sign in

Paper Citation Record · LEDGER

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

As of 22 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 36 inbound Pith citation observations for arXiv:2411.16489.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.16489 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:07:48.934782Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:23:45.383186Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T18:07:42.247566Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3be880df-c1eb-4573-8879-50788f2ff16c · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.250761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.841797Z digest=sha256:f57732c51b790db0445a9c3be6a25157484c1675b6c6912669797d47e84360d6

Observation 56311ecc-126c-432d-92c8-07b55654ce1c · outbound

This paper cites SAFETY-J: Evaluating Safety with Critique.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? SAFETY-J: Evaluating Safety with Critique

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.846726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.846726Z digest=sha256:a6d21765994ca00f96adda40d6df4344f5f87793971dd983f1545b5971ed7018

Observation 4cd5ffe2-d175-49d0-b690-a5b3f25fb82b · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.236647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.851359Z digest=sha256:b1429007dd08ca1f23fba49e7885aca82b78adbc5bf36556e2de442475548f89

Observation 172945e6-9bbb-437e-8818-197fcfcef603 · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.222231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.855706Z digest=sha256:a398d2b488adb25e9cafa30f96492ed5b439aabb9dea38e310505adaed2f332f

Observation f744c515-47f7-455d-a333-d908b74df1e8 · outbound

This paper cites O1 Replication Journey: A Strategic Progress Report -- Part 1.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? O1 Replication Journey: A Strategic Progress Report -- Part 1

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.860045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.860045Z digest=sha256:d11f8f7534665a036696f8118b4d568ab334277449c510399f40243c97121788

Observation d107d821-572a-402f-a8d0-fca5d8542f9b · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.864430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.864430Z digest=sha256:869f6ca2295ba82c01840027a231d6fd94113f868c5085077b7b42a3557ce43b

Observation 12d9312a-c90e-4715-88c8-2ae643b8e94f · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.207412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.869443Z digest=sha256:edf5019a34b9f7ecc7fad2533b326d9bd5bf19282468e6aa2cb92b085b201ac0

Observation 4c57e2a7-cf79-4414-a756-016578c6533e · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.193421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.873612Z digest=sha256:aa7e74402c57d93b1129837dde3110f53e22aedde7aa356e477e070dd3bc160c

Observation 90c4a1d5-1ebf-4326-95bb-e8e1cf3ab03a · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.179697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.877679Z digest=sha256:f479d954a85a91765f2472b2f61dc441b68c77694a94e486a2659df6c188eef5

Observation 4ed22f88-f48c-4067-9227-96267eb39f45 · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.165821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.881833Z digest=sha256:9405ad1afaa0d6fd7c6f1509c945fe1c9e9a8d500162cf8b2d84f1500cb6c13b

Observation a486b82a-463a-4453-a425-1ac3cbf7b7b6 · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.151284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.885876Z digest=sha256:1957419c6bb9c7861fd459b8888d895ab8e0eac22f15e3ddcaac5cca737f0662

Observation 0c82c8ff-fc68-47f6-bc4a-ad208539e709 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.890377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.890377Z digest=sha256:73b1cf5c387bc2702e784957f331a99060dd45b1d5216534fe3e72bd97c26204

Observation e34ef943-f0b5-4d86-bc6f-10cae8658b79 · outbound

This paper cites Measuring short-form factuality in large language models.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Measuring short-form factuality in large language models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.894841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.894841Z digest=sha256:a44e1352f4b8f92b61eb1430fbd9dcac94128d93a2dae2b2fa88903dbd5ac419

Observation e21f5184-e3c5-4b69-a975-5801fbf822a8 · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.899312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.899312Z digest=sha256:f2223b52909a1586bb76c01960cd17ce19e0a8ddb9f595b5dd3f053e627afcda

Observation 2926c0c2-d735-4f6b-9007-2d01487d3dd7 · outbound

This paper cites WizardLM: Empowering large pre-trained language models to follow complex instructions.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? WizardLM: Empowering large pre-trained language models to follow complex instructions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.907837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.907837Z digest=sha256:0d6bfbabca1b86e583a145080609d3cf570b671e61d4b2a9c0aa00c668bee149

Observation f9cf0810-6b54-470e-a1f6-a2de28de31b7 · outbound

This paper cites Qwen2 Technical Report.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Qwen2 Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.912230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.912230Z digest=sha256:d6a809953908f42bc0209253ab75201bf5bb1ed325fa0befb8a646ead5a4a0a2

Observation 98debf8b-c6f1-4da7-899e-6ac0199f1268 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.916845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.916845Z digest=sha256:01556004515503a0111705d327b01d9ca9dbc06e1be6451701a8b3b75df4ad4d

Observation afb7f50d-6e80-4458-8f0f-fa3910b44744 · outbound

This paper cites MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.921933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.921933Z digest=sha256:821227c5d339546630bbbf878b06a42aec3c28ce1e14dd898bdb1de79e857be9

Observation 1a9eebcf-acdb-4b9d-ba2d-326b615ac4e4 · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.119184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.926492Z digest=sha256:4515cf0e952d671bc3eddf84aad5eca700f521e3ae945d47c21455bf484abc5a

Observation c7b0a243-9e7a-415d-99ba-e9a27fc7f44d · outbound

This paper cites an unresolved cited work.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-12T13:07:49.105007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T13:07:48.930618Z digest=sha256:355ec04ba34a36eea107cd25c732c93a9bfc3f79678d5ae809068e9434e6972c

Observation 861e9aa1-cad4-4c87-8eef-94e5eb205783 · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Fine-Tuning Language Models from Human Preferences

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.934782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.934782Z digest=sha256:7d093a965bfcaaa85b9b274f8462e2a2cfcce5cba451e10e459b2a3b2cc1e2a5

Observation 296b1e12-1e7d-4a6f-ab73-fd69eb06e4cd · outbound

This paper cites Advances in neural information processing systems, 35:24824–24837.

O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson? Advances in neural information processing systems, 35:24824–24837

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-12T13:07:48.903437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:07:48.903437Z digest=sha256:4e4849bad3244da9166bd3fff53df9e863e8c3eef26c8ac2be9842216e013494

Pith citing papers

Observation 66aa41ce-2d57-435e-9b1f-b2da403bfdad · inbound

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems cites this paper.

Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:35:31.460514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T00:35:31.375020Z digest=sha256:ee020334e014b15242eaef70fd608c9738efbc1ffbc2df500f4153daae052218

Observation e3ad2e09-c05e-4176-bd6b-08c3c5986e11 · inbound

Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective cites this paper.

Scaling of Search and Learning: A Roadmap to Reproduce o1 from Reinforcement Learning Perspective O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 8

Resolution
malformed identifier
no resolver link, observed 2026-08-11T12:30:38.353753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:30:38.353753Z digest=sha256:ea9867e9a2c3b23e5b95825ad02062335c220b17a8ce05c89262104e7cf6766c

Observation 3cc02538-4436-4905-b06a-e67da6f96eed · inbound

DRT: Deep Reasoning Translation via Long Chain-of-Thought cites this paper.

DRT: Deep Reasoning Translation via Long Chain-of-Thought O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T05:32:00.307512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:32:00.307512Z digest=sha256:aa27b508b86bab1f703946168678ac463764f0bcf51cb16ad9980e109a1a8e8a

Observation f0602ee0-5a42-4edf-8b40-80ff454a2088 · inbound

Search-o1: Agentic Search-Enhanced Large Reasoning Models cites this paper.

Search-o1: Agentic Search-Enhanced Large Reasoning Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:36:27.697612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T17:36:27.515468Z digest=sha256:01d36bd512b90b8e1e1e843fed735bf2fed9ea1303067b50924e3726f7a93879

Observation 22b49c8f-5b87-473f-b689-220da3b35a17 · inbound

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search cites this paper.

Satori: Reinforcement Learning with Chain-of-Action-Thought Enhances LLM Reasoning via Autoregressive Search O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T11:57:48.758762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:57:48.758762Z digest=sha256:25a1722ef5c6cb51b3e21057b7e7dfcb4d1b1c041bd972b94bf697eb90b7a4af

Observation 45e8bf9c-1239-4860-b16f-1e329ffb8e8a · inbound

LIMO: Less is More for Reasoning cites this paper.

LIMO: Less is More for Reasoning O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 135

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T02:11:37.579657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-17T02:11:36.932541Z digest=sha256:079df3499950e4e40daeb2a0f2f956cd9a9c0c3827710a012bec9f297dc16f75

Observation e5b2e5d4-df87-4057-8644-e02a0a752806 · inbound

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools cites this paper.

Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T22:03:46.368270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:03:46.368270Z digest=sha256:598f8ab8958152fb04bae4948c4dfc4145cc7b290f08e0a35b97963cd27012f0

Observation 061c1b28-cf6b-4c55-893e-19b4692aa84a · inbound

Iterative Deepening Sampling as Efficient Test-Time Scaling cites this paper.

Iterative Deepening Sampling as Efficient Test-Time Scaling O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T19:23:26.463205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:23:26.463205Z digest=sha256:2c5c0b4a55a1eb8d2a82f446e19544df5817c10823a2769e3d87fc10f8e18391

Observation 624dac24-83cc-4a27-be2e-2c9f3a305274 · inbound

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling cites this paper.

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:40:35.955396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:40:35.955396Z digest=sha256:b04b1864baf33f7c1699247c1d1a4fceecafc2e10fa42175d2663439c5535adb

Observation 6b593b83-c156-43af-984f-0629a2a6616e · inbound

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition cites this paper.

On the Emergence of Thinking in LLMs I: Searching for the Right Intuition O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:53.383766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:25:53.383766Z digest=sha256:470ff666d65bdc620f9346a903150b42aaa3dadc8e7c94aa68224a24ef84dbdc

Observation 199ead34-9050-4d22-a51a-9b03c6f5914e · inbound

Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning cites this paper.

Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:27:50.987846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:27:50.987846Z digest=sha256:a4825bc9a9149943a7fd8aad065657bafe45cda73d8615b8d897b6d5f0e74cbf

Observation 116bdcc9-7804-4ae6-8eda-34eae7c4aea3 · inbound

Typhoon T1: An Open Thai Reasoning Model cites this paper.

Typhoon T1: An Open Thai Reasoning Model O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T22:52:54.335881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:52:54.335881Z digest=sha256:c35a7c411725f8da8b154fea191a3c7b88037e32a7be7e1cc5fd334823e2e744

Observation c9e87381-d63c-49e1-ba34-32b38573dfa0 · inbound

Dynamic Chain-of-Thought: Towards Adaptive Deep Reasoning cites this paper.

Dynamic Chain-of-Thought: Towards Adaptive Deep Reasoning O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T20:51:16.684149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T20:51:16.684149Z digest=sha256:27b4c6f669c5652832ed6ee6daef2430a3ac63c31e6389c447be235c90979c2f

Observation a7a32e34-f39e-419b-ac8f-77a746c94ba2 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.613149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:255b4f699ea2b53e2147775d757ff7934b13cdeb4a06dd351b1d6d6b61e7929c

Observation 2b513696-1e65-4f9c-b23c-63ea13f83fdf · inbound

ToolRL: Reward is All Tool Learning Needs cites this paper.

ToolRL: Reward is All Tool Learning Needs O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T00:26:48.328841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T00:26:48.291431Z digest=sha256:9a5c16e875dc4044062c555a6202659e04e90ff8404fa12413b4b77086cecd55

Observation d1d64196-4f45-4af1-9615-00bb001e7f69 · inbound

Tina: Tiny Reasoning Models via LoRA cites this paper.

Tina: Tiny Reasoning Models via LoRA O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:23:45.383186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:23:45.383186Z digest=sha256:8bdeaae9f0073132facf62c322e55840759d2f1faf0c363dbfd7d4c630d8545c

Observation b18b6175-4216-42d1-913a-6ff614733e84 · inbound

Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think cites this paper.

Beyond the Last Answer: Your Reasoning Trace Uncovers More than You Think O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T05:26:00.368244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:26:00.368244Z digest=sha256:49674be3b233b86512f979633ab830aec6666b06d505c5d01bef01b4a94bdf3f

Observation 8dc6e9ce-ab5e-4226-9370-90e052f1ec01 · inbound

Quantitative Analysis of Performance Drop in DeepSeek Model Quantization cites this paper.

Quantitative Analysis of Performance Drop in DeepSeek Model Quantization O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T00:58:05.729259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:58:05.729259Z digest=sha256:cb027ee1f609c03339c63d27a6554b0b57290eeb4d771252f2f0a349ef6ee973

Observation 2f7a70cb-b9c7-4c3a-8e27-6db77aebffb1 · inbound

Long-Short Chain-of-Thought Mixture Supervised Fine-Tuning Eliciting Efficient Reasoning in Large Language Models cites this paper.

Long-Short Chain-of-Thought Mixture Supervised Fine-Tuning Eliciting Efficient Reasoning in Large Language Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T23:56:46.231895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:56:46.231895Z digest=sha256:154c3025a7c823968caef21bebabf76db08d07ad7945118d91baddc2a99e9ebc

Observation fa15ad93-73f0-469b-887e-3b1abbcd8c2b · inbound

Crosslingual Reasoning through Test-Time Scaling cites this paper.

Crosslingual Reasoning through Test-Time Scaling O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:00.481057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:00.481057Z digest=sha256:bc2051eddee3bb04fd81a361606bfa74a448ed8b4e93e9d3eb827d54e2b8903f

Observation 25972b8f-464b-49e8-94e8-f8f5994bcaff · inbound

The Hallucination Tax of Reinforcement Finetuning cites this paper.

The Hallucination Tax of Reinforcement Finetuning O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:44:32.005515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:44:32.005515Z digest=sha256:f776dc747c5a60f46f4d5cb097738503c1a543433246450eb21d29d2638295e2

Observation 13095c2b-6ad1-4105-8fc5-ab73c61da6ea · inbound

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models cites this paper.

DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:21.824586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:21.824586Z digest=sha256:eea7665e7fe1d2e0b77d5ad37150054ee4feb647bee2e2ad5359aa1ddce4fc9c

Observation 93068a75-31cb-47d2-9e61-f089585532d5 · inbound

Which Data Attributes Stimulate Math and Code Reasoning? An Investigation via Influence Functions cites this paper.

Which Data Attributes Stimulate Math and Code Reasoning? An Investigation via Influence Functions O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:07:47.214181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:07:47.214181Z digest=sha256:f6018c9ea53dce391fd8b032fba87b98df383f067a1d6198f137ff2f0169ee5f

Observation c194fdc6-7f42-4f4f-aa43-261835bef716 · inbound

Pangu Embedded: An Efficient Dual-system LLM Reasoner with Metacognition cites this paper.

Pangu Embedded: An Efficient Dual-system LLM Reasoner with Metacognition O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:16:13.699004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:16:13.699004Z digest=sha256:7bcec64fa13baf9f352ff15cdd5e9d0021f58b455df45532635e138f5cbb60f0

Observation 4730ea76-6a7a-4f55-a2d8-0e6386a5c80f · inbound

Discriminative Policy Optimization for Token-Level Reward Models cites this paper.

Discriminative Policy Optimization for Token-Level Reward Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:53:01.360490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:53:01.360490Z digest=sha256:e2b23d345a70df5e94d1605df99fa51c1fdc82ace891ce7652ce4f24e15d2679

Observation 8ce8cbb9-e0b3-4779-90bd-79f1d93cbdd2 · inbound

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time cites this paper.

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:41.259399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:41.259399Z digest=sha256:600385a06323ebf7d01137fdd9eaec075fcf0554b62a3946cfbab3aa015858eb

Observation 138d07e0-6681-424c-993c-7834b5af1842 · inbound

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL cites this paper.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.181809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.181809Z digest=sha256:764e4bd68013e24ae0c4661b1241e84f91acc92d830398201024240185247901

Observation d7c69f25-f84b-4e19-b3ab-0bfb19014bd0 · inbound

A Survey on Model Extraction Attacks and Defenses for Large Language Models cites this paper.

A Survey on Model Extraction Attacks and Defenses for Large Language Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:08.994979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:08.994979Z digest=sha256:9275fd08f5ba1299f75674a5c032a37afa86c7d4086050deff23cf105d5c3b21

Observation 07735104-4146-4919-a208-5bb64e9c917f · inbound

Decoupling Knowledge and Reasoning in LLMs: An Exploration Using Cognitive Dual-System Theory cites this paper.

Decoupling Knowledge and Reasoning in LLMs: An Exploration Using Cognitive Dual-System Theory O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:22:13.260427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:22:13.260427Z digest=sha256:21b501fbdb9c02bc97acd886f6ab8695a447311ac6c740411657c73a3b2878d6

Observation 24ce318b-8b3f-498d-af59-d535234a24c3 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T10:38:59.955343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:38:59.955343Z digest=sha256:8a123a17cdcbebdaf32102c97f9e781c8d65a68f07e470f90e0bee4ffeff959c

Observation 7f5dc798-99ad-4a9a-a694-de37ab8f8aae · inbound

Memory in the Age of AI Agents cites this paper.

Memory in the Age of AI Agents O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 192

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T18:18:20.614997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-11T18:18:19.911342Z digest=sha256:7a0abe73a8bcec3d4a59592a3980856fab8698e1270b60519e7ab5c4edbc75f0

Observation aaf9560f-921a-406b-ab95-de86c7eaa519 · inbound

CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning cites this paper.

CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:20:57.502337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T13:20:49.919833Z digest=sha256:1db238d406c188f6b6123241bc112a3d834e81ef05ec8a92285d84662e03cb3d

Observation 4d92913c-2ecf-41cb-8814-f3b7dc39b92d · inbound

Logic-Regularized Verifier Elicits Reasoning from LLMs cites this paper.

Logic-Regularized Verifier Elicits Reasoning from LLMs O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:51:10.845910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-08T10:54:01.229934Z digest=sha256:3d0c6abf8347652fe6807e10e3210033afaed06e21c3320bcfc406533f7c3c41

Observation d3173cee-78ea-4144-98aa-f3dc22ec6d2c · inbound

AIPO: Learning to Reason from Active Interaction cites this paper.

AIPO: Learning to Reason from Active Interaction O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:06:31.465733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T01:17:28.124867Z digest=sha256:dfcd49a48d989fe718dc647d2e103e6fed299c2b66909ead793119757c6c20b1

Observation 6a6304ac-de7a-463d-8d0a-d4809d91903a · inbound

AIPO: Learning to Reason from Active Interaction cites this paper.

AIPO: Learning to Reason from Active Interaction O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T18:07:42.250014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T18:07:27.492419Z digest=sha256:4f7d8ce2c57192a5f01be15bc99f2bd8eae21949c747ba5d9e2906f71469d283

Observation 13462202-4359-43b5-b29a-21b1cd625d5f · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:45.482027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:45.482027Z digest=sha256:4c78395858ec770ea94c9a197055f9e9ed88379b3df268e1a5662b4b133bec02