Pith. sign in

Paper Citation Record · LEDGER

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL

As of 18 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2506.02338.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02338 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:30:40.435211Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3a620c53-2971-4c03-be1c-26aed213bb56 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.095403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.095403Z digest=sha256:e53022b0a2e2df22443c48972648227c4b1c82ea706958df2a946d6b6557608b

Observation fafcd2b1-2580-4a82-a3e3-1cafd5624b9c · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.109082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.109082Z digest=sha256:bbe9868116c6ce157bddd730b313a6f91ed5ee1b9a779b7fe026f38434cb75d8

Observation 470914d3-e96b-4022-8e51-24794ef3b387 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.120327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.120327Z digest=sha256:15eabf6ca9cdfe3148b59722113e4d5dc9398a305e937187d7220174ff434de3

Observation 8336a38d-68dd-4876-8394-68288d115880 · outbound

This paper cites The Llama 3 Herd of Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.129217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.129217Z digest=sha256:9630b25cfb067fa0cc8fa5cad5232aa48920f8b9f042dcd0da210ac2b3bfc068

Observation abc0b61a-31c9-4586-8bf7-70f503b7e8fe · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.139932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.139932Z digest=sha256:ae8564faf0d6815a3146ff3f22e75f5088e7ec29def409fdc713b228b6f372cd

Observation 8144c61d-b6dc-4247-8e83-bd410c59b013 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Measuring Massive Multitask Language Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.149670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.149670Z digest=sha256:f5e857b32de2bf6393442101fb094345fddf4fd22094f991a60c737d1c678f18

Observation a407ff1e-8f7f-4c45-9e09-ba3f39fc24a4 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Measuring Mathematical Problem Solving With the MATH Dataset

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.159435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.159435Z digest=sha256:5dee8277b454c2639698c83a2246df0e96c4f63d445711607f11b040e2c601d0

Observation 0d5b3497-a783-48d8-bd38-4be515430b98 · outbound

This paper cites Unsolved Problems in ML Safety.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unsolved Problems in ML Safety

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.166528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.166528Z digest=sha256:35e3e999da05fc53fcdbf1298b16e9cd6e76f8d59a67fc0e1b0010f4a7e52ea4

Observation c36feee3-b382-499b-9777-4e8baec108c3 · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.175491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.175491Z digest=sha256:87f6166376ba8d54ddd52f04befbcdb1b0fa27d5466694a681e20e0402badae1

Observation 138d07e0-6681-424c-993c-7834b5af1842 · outbound

This paper cites O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.181809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.181809Z digest=sha256:f9898a87617223b1550cb327031a38c30f404583208ccf9f4d8a2a7c3eda498f

Observation 6a63a3a0-0fdf-47cb-a435-65664274dc02 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.170900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.190888Z digest=sha256:93392357ee2a286e909ce8d7e3b9fd6e35168f6f9f2e7a8d683e7e5326878104

Observation 75a3efd2-fc16-4e71-93ea-e76bcfddf22d · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.198182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.198182Z digest=sha256:b392a4f2f1b3406b6b3b81436cffd03cd92ae32e9ccd0c37a6d8ed908d9c983f

Observation af14dcbf-dada-4fe0-a10c-0f134e11cae6 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.137299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.207318Z digest=sha256:d6fa3b08723bddb6dc05f98bbd85ee21583c487d856e76d657de643012fb436a

Observation 59e8e1e9-4b76-4c12-8e4d-39a0f264ed32 · outbound

This paper cites Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.214833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.214833Z digest=sha256:72564c7b9cd4ab3f30f5960ccbb74b7992d0eecde0aa6d293ae0625d5a7fcfda

Observation fbd92be8-c84b-4151-9a72-4c0cbb2b1a24 · outbound

This paper cites Let's Verify Step by Step.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Let's Verify Step by Step

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.222867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.222867Z digest=sha256:9330a3d97eeba6e4f0fff4a8fba19acf6bfd7d53a12f6c16a779e1ee6be33e45

Observation 196780c7-d2c6-4134-9ae4-6c4362406864 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.232268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.232268Z digest=sha256:30bb6a67d0cbca606a4c0109a2c3a4dc7a3de718694c0ba3146e667bb419eb10

Observation 38c2129e-14fb-4c30-b090-1a8270f5421c · outbound

This paper cites s1: Simple test-time scaling.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL s1: Simple test-time scaling

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.243867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.243867Z digest=sha256:899c1d438f41d9d0a146237f4bb7b391b747ebf72627cecabe7b888c199cf50f

Observation 5791fb73-f390-4b0c-a883-11e76bfef89e · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.101056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.249490Z digest=sha256:5b6b82b154864e032f55e71f8026d436004873120f8f83f7565e5a1fe9817d09

Observation 133c582f-5501-454a-ae21-ee900206634e · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.061042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.256938Z digest=sha256:683e45996aaf17de2ba4cc1a8af6dd1136fc5a5cd2228f818a72d6ba5f5a9848

Observation 0bfe602b-2997-480c-bc68-fc6e76f3d866 · outbound

This paper cites Training language models to follow instructions with human feedback.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Training language models to follow instructions with human feedback

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.263653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.263653Z digest=sha256:cebcd57b9ee3f8900f7c9d72ef86980adeb7b672c575af2c15b7825c9bf99397

Observation 7386175c-0d20-478c-8879-b89f2fd30cf5 · outbound

This paper cites BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.270967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.270967Z digest=sha256:37d023877c2d223c171719d438ec25005723f325575c88b51bcba633392412a4

Observation 51ad3838-3c8f-4fc4-8ea4-36de2e5e33f6 · outbound

This paper cites Qwen2.5 Technical Report.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Qwen2.5 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.279569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.279569Z digest=sha256:2c7c4dd594ff3b5cfdb3eb6a38ce3c299a7e8fcfc4d4a496da16acbefcd61676

Observation fd945479-6501-437d-ac78-15869f69d8b4 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.292039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.292039Z digest=sha256:0f0300374e275079b7c2c6ea4530decbe9589c471c6a7ade97b7afc97d2aea81

Observation 9c46240b-2e4d-4188-8cd7-ad583572395d · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.299149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.299149Z digest=sha256:4a01a8eaaf23bb12aad2ea2aeeecfa8093f66c0d6c6cee8bbc41d82d21a63c19

Observation 505fc5cd-2c32-4803-932d-35a185277239 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.304456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.304456Z digest=sha256:aa4c708f6308c26924a6affaa8d42ae2d5a3e2d5381e41f092c39a8a2b85c37a

Observation c45ae7de-2d83-443e-908b-eef6cfad4427 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.020936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.310291Z digest=sha256:f1f62448ec59d7724e6ec1ea0617b15e3398f90d80d8585bf071e4fc635ecb64

Observation 0981fa35-cb9a-4e76-baab-9ea9177ab3f3 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:41.977143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.319370Z digest=sha256:3e54e4085291e2f4a3b18bb1b897e785232914f7866fbf104c78fd91a01bec2c

Observation 70e64e28-d9e0-4cd0-89fb-8b2cf5bcde62 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.327281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.327281Z digest=sha256:c687dd237c753504e119ca23892764f9a6c8cdfe74b75f4d9c68b6babf9fc760

Observation c47bf320-e49c-45d4-a33f-144415becc0a · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LLaMA: Open and Efficient Foundation Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.332643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.332643Z digest=sha256:aefe152032e79adb7f20e333c1b1c658295074bf0adfdc2d351dcc043d99ace3

Observation 37be43d1-c47d-4fa3-bea9-6bf92ae3e265 · outbound

This paper cites Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.339030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.339030Z digest=sha256:2d59d0bd58a6672e67654541977ae79fbd9c5cb73f62f60fe6e7021738061866

Observation 53536604-b829-430a-a54e-2a8efd3ec156 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.351279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.351279Z digest=sha256:38cc806424106e47338fc7241e5958ea91b5b61c9479ef3e7998defb33e1e6ac

Observation 4f37b2ca-202f-4c65-a786-0e39b2586b8f · outbound

This paper cites Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.358899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.358899Z digest=sha256:8bbd474bf44eeccb7ef21dc39e7da98e077dec32e4f40800833a56b937fc78b2

Observation db56cecf-9f5a-4d29-9bcf-863d9e379bfa · outbound

This paper cites RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.365512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.365512Z digest=sha256:884f80fe30e4e613335e53deefc6b7f7de47953eb33aa39339b57de5918feed5

Observation a35a4bf7-4597-45b2-9057-8ef4b19a54bf · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.373089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.373089Z digest=sha256:523222c7e14d0c08cb1a40f694c3318bace5be828934b03ec2a664962307e063

Observation 0bb73033-dc53-47b9-83ae-7fc1a637e5ab · outbound

This paper cites LIMO: Less is More for Reasoning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LIMO: Less is More for Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.380185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.380185Z digest=sha256:5726dcfce9f2bcebeca8c0a24be54e52e8e772fcaba818557b4d11e1f4e40b9c

Observation c969852d-8e08-4927-be52-38b9736a4cd4 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.388167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.388167Z digest=sha256:32add457da69f3823334018483aba9cf611e70f97678b4d849aec8daed8ab7f9

Observation 0fa1eb0f-0c63-4a2e-97db-e2b700904a24 · outbound

This paper cites FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.394611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.394611Z digest=sha256:d4aae2884e15824e0c67cac35fb4a46a3be5066dab5977ccc2e4f006da8c3d01

Observation ab44c3fb-adf0-49fd-91b6-4a4f2c95a478 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.401314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.401314Z digest=sha256:508c24162deb969bde0d37b34fce480cee6b9ab0a6da84ef1ed772aaac41504e

Observation e5ff97dd-070b-4d5c-be49-0b54664f2870 · outbound

This paper cites o1-Coder: an o1 Replication for Coding.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL o1-Coder: an o1 Replication for Coding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.410198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.410198Z digest=sha256:2e643421e47493a687ba35d09a12208a8b66432a47229533da7bc7c57676090b

Observation a2a8c277-e738-4e85-b9bd-732e42193099 · outbound

This paper cites LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.416440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.416440Z digest=sha256:1e6ca11b65b51658ea6646d10daef9c8c834d3cab52bc0bdf116c055f2d6eeae

Observation bd8639f7-bc2e-41b5-8c96-f5c1e861a5b0 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.422849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.422849Z digest=sha256:4447288a0957616267f3c6e5c62ff6ba886cbc76c32e1f1f30cc0baf17407349

Observation 43c14ce9-5f8a-4e73-bf90-58096c7406c1 · outbound

This paper cites online" 'onlinestring :=.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL online" 'onlinestring :=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.428196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.428196Z digest=sha256:9e2ab33676752401bcd343d6c6d7cac7bf5704e63ade6ee25d49c4d0e2ac2938

Observation 5f10d5a3-57b7-431d-ba84-a96b80a47512 · outbound

This paper cites write newline.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL write newline

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.435211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.435211Z digest=sha256:c3fb3587c2ba7c879541bab76c0088f5f80c584a0ea373a13ad25eaad5dcc913

Pith citing papers

No inbound Pith citation observations are available.