Pith. sign in

Paper Citation Record · LEDGER

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 0 inbound Pith citation observations for arXiv:2506.02338.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02338 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:30:40.435211Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3a620c53-2971-4c03-be1c-26aed213bb56 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.095403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.095403Z digest=sha256:3dd1a67bb1438af13742dfa794364bc41d34fb22f1b82a0c997797f113b5fe6c

Observation fafcd2b1-2580-4a82-a3e3-1cafd5624b9c · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.109082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.109082Z digest=sha256:6be17a7a4a1dba59850101066ff7f8c85b96ed54993ea32834aa66081d52aabb

Observation 470914d3-e96b-4022-8e51-24794ef3b387 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.120327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.120327Z digest=sha256:636a5da04929bc488a99aea20d68a2f786dfa4524002ef5675e44f2465748075

Observation 8336a38d-68dd-4876-8394-68288d115880 · outbound

This paper cites The Llama 3 Herd of Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.129217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.129217Z digest=sha256:934ae4ea1b11e293ade6ab562ba20e4e2f25e9c60477b332e75962daad3a93d6

Observation abc0b61a-31c9-4586-8bf7-70f503b7e8fe · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.139932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.139932Z digest=sha256:72b998ef6af8249d686c770db90edb20ffeea63d4d4b27286236afacdd5f0efe

Observation 8144c61d-b6dc-4247-8e83-bd410c59b013 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Measuring Massive Multitask Language Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.149670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.149670Z digest=sha256:fd8cea474fb50c5a8ecf8ced88408496f5cb39bf5c376435977aee5f46582d29

Observation a407ff1e-8f7f-4c45-9e09-ba3f39fc24a4 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Measuring Mathematical Problem Solving With the MATH Dataset

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.159435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.159435Z digest=sha256:0498415985724208ddde40d3929f6cb21e6d839655520c500f7c0468e754caaf

Observation 0d5b3497-a783-48d8-bd38-4be515430b98 · outbound

This paper cites Unsolved Problems in ML Safety.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unsolved Problems in ML Safety

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.166528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.166528Z digest=sha256:ecc3ea747153fc2170c0142d7f3385777dc71e65039790f5a29df5653784a208

Observation c36feee3-b382-499b-9777-4e8baec108c3 · outbound

This paper cites OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.175491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.175491Z digest=sha256:4fb81cc21ab7d99e9f8f2ff5a9fa1ef1840f8b5d981bcf334823a2b4471d7e62

Observation 138d07e0-6681-424c-993c-7834b5af1842 · outbound

This paper cites O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL O1 Replication Journey -- Part 2: Surpassing O1-preview through Simple Distillation, Big Progress or Bitter Lesson?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.181809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.181809Z digest=sha256:faf02760f106ed331f9a9c9316889222ebc8c0f09d96b118478056ab64199ff6

Observation 6a63a3a0-0fdf-47cb-a435-65664274dc02 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.170900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.190888Z digest=sha256:4cc7577bcceb9f67c78100c1a3361e5e31f1a03336e74404386a0e83451ed41e

Observation 75a3efd2-fc16-4e71-93ea-e76bcfddf22d · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.198182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.198182Z digest=sha256:fc382205832b37c35d6b967c9b65d0dcce07bf802b2d311dc8d36ea3456be863

Observation af14dcbf-dada-4fe0-a10c-0f134e11cae6 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.137299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.207318Z digest=sha256:cee2cbaaae529c80d0ca3eaf9fe5c842244a00179c04b231c15e658eb1633968

Observation 59e8e1e9-4b76-4c12-8e4d-39a0f264ed32 · outbound

This paper cites Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.214833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.214833Z digest=sha256:c670fbfe3bfd2515705d44c535ee86e0d3b638229467a469e2157a70359118eb

Observation fbd92be8-c84b-4151-9a72-4c0cbb2b1a24 · outbound

This paper cites Let's Verify Step by Step.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Let's Verify Step by Step

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.222867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.222867Z digest=sha256:d25ae421fb1f7cd34a63008820f005bcc022c656595c8715ede4c66ebf7e9408

Observation 196780c7-d2c6-4134-9ae4-6c4362406864 · outbound

This paper cites Ring Attention with Blockwise Transformers for Near-Infinite Context.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Ring Attention with Blockwise Transformers for Near-Infinite Context

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.232268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.232268Z digest=sha256:c2d9d494d55497452e1ff2a61ac22b54981aafb6541409008c8c318d457bf249

Observation 38c2129e-14fb-4c30-b090-1a8270f5421c · outbound

This paper cites s1: Simple test-time scaling.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL s1: Simple test-time scaling

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.243867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.243867Z digest=sha256:81f7cb17725ad289b34b5b373c0be0cd174cbfd3c2c1d0c36fb46562a10c6d47

Observation 5791fb73-f390-4b0c-a883-11e76bfef89e · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.101056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.249490Z digest=sha256:ae288d2038039a899e5ceb9e47c38664b951e2e85e9d0cdf85347fb4cb4d71ff

Observation 133c582f-5501-454a-ae21-ee900206634e · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.061042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.256938Z digest=sha256:8d9b49cb3bd46744568a34862b983d19d4fd9fa942533f6bb5be92cb5e5b43d5

Observation 0bfe602b-2997-480c-bc68-fc6e76f3d866 · outbound

This paper cites Training language models to follow instructions with human feedback.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Training language models to follow instructions with human feedback

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.263653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.263653Z digest=sha256:55f8004da9767db679bf9d598ddd25fed4cf43ea8deadf6a394f6b64f8903f74

Observation 7386175c-0d20-478c-8879-b89f2fd30cf5 · outbound

This paper cites BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL BOLT: Bootstrap Long Chain-of-Thought in Language Models without Distillation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.270967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.270967Z digest=sha256:e660dd7957ab97469c9fe811ff7f9a1561b3f3106437ece137ee618fef37bfaf

Observation 51ad3838-3c8f-4fc4-8ea4-36de2e5e33f6 · outbound

This paper cites Qwen2.5 Technical Report.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Qwen2.5 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.279569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.279569Z digest=sha256:2339d585e40b9cbc876d698d195d0b1afa8122e7b32cd742d88689ef289bc431

Observation fd945479-6501-437d-ac78-15869f69d8b4 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.292039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.292039Z digest=sha256:346e0d72c1f9e24d745ef722788999efe05aa9d5d43f18a8567036f6fb1b2a6b

Observation 9c46240b-2e4d-4188-8cd7-ad583572395d · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.299149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.299149Z digest=sha256:0003681dccef109f2cadd1888129f62da50c4a63b5ebc643c8fe99a3441d7782

Observation 505fc5cd-2c32-4803-932d-35a185277239 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.304456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.304456Z digest=sha256:f86eacd5b5c85306b666eb9ac8f3b24f0132436b08e8f9bb8fa6157b786072f7

Observation c45ae7de-2d83-443e-908b-eef6cfad4427 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:42.020936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.310291Z digest=sha256:edd0a475557b6ead8322106a66170c63c376f4d22cdd7fb99d7939e2bd4dd4ef

Observation 0981fa35-cb9a-4e76-baab-9ea9177ab3f3 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:30:41.977143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T11:30:40.319370Z digest=sha256:ee64b4d84b235dfb44653c202aa4c86e773d1791a3395493afb0014e326238f4

Observation 70e64e28-d9e0-4cd0-89fb-8b2cf5bcde62 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.327281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.327281Z digest=sha256:852ebc186e5a31559752a3ad22ffa4e9cba1c021a803ded8a0e483fcc75e80aa

Observation c47bf320-e49c-45d4-a33f-144415becc0a · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LLaMA: Open and Efficient Foundation Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.332643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.332643Z digest=sha256:2d6cf10ad7b6b1e97cd46aa15daecda84809aaf3f535033b9070bb74a192811a

Observation 37be43d1-c47d-4fa3-bea9-6bf92ae3e265 · outbound

This paper cites Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Q*: Improving Multi-step Reasoning for LLMs with Deliberative Planning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.339030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.339030Z digest=sha256:d9c7b01fcad33c4b7ce5c282b5ec291524dec250f940ba88a95e40d4f01f7a47

Observation 53536604-b829-430a-a54e-2a8efd3ec156 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.351279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.351279Z digest=sha256:a75025e144f0f85a31a9e774ce0b5f2fc38c431ed9cd27f94b63d9153cf2bb7d

Observation 4f37b2ca-202f-4c65-a786-0e39b2586b8f · outbound

This paper cites Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.358899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.358899Z digest=sha256:2cc77d8c1a928e43ab7b06b5f40a9537c123fe7cda1f009f38389df4e0039353

Observation db56cecf-9f5a-4d29-9bcf-863d9e379bfa · outbound

This paper cites RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.365512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.365512Z digest=sha256:b72078b1e5a1abbb63101f237da0c9f39868da1fc5e5f8425bc1e9b0d6350e23

Observation a35a4bf7-4597-45b2-9057-8ef4b19a54bf · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.373089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.373089Z digest=sha256:2fd10d70fd3b453c5b0bb67da71c271843a266e3e3a9f23b11a4ab9d48601a71

Observation 0bb73033-dc53-47b9-83ae-7fc1a637e5ab · outbound

This paper cites LIMO: Less is More for Reasoning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LIMO: Less is More for Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.380185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.380185Z digest=sha256:c94af00a50ce9b7d3cbbc1637474075bc822781c4284992aa7ad882082eab6d4

Observation c969852d-8e08-4927-be52-38b9736a4cd4 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.388167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.388167Z digest=sha256:f05c3692ed2d47718ea9ceeb8618962f5bb1b61f820b2f655462650cb5cf18ac

Observation 0fa1eb0f-0c63-4a2e-97db-e2b700904a24 · outbound

This paper cites FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL FineMedLM-o1: Enhancing Medical Knowledge Reasoning Ability of LLM from Supervised Fine-Tuning to Test-Time Training

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.394611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.394611Z digest=sha256:6c0cf479a7f69f85f24c0c2370dea6a18c8cc9ffcb31a13f68731e005410c4eb

Observation ab44c3fb-adf0-49fd-91b6-4a4f2c95a478 · outbound

This paper cites LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LLaMA-Berry: Pairwise Optimization for O1-like Olympiad-Level Mathematical Reasoning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.401314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.401314Z digest=sha256:e11058a9d8bdb9baf70feba561a2cf524fdb4ea5cca8a4238885fdedf3cb2859

Observation e5ff97dd-070b-4d5c-be49-0b54664f2870 · outbound

This paper cites o1-Coder: an o1 Replication for Coding.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL o1-Coder: an o1 Replication for Coding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.410198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.410198Z digest=sha256:f572a5936499152949c904d23fc74d30622a0fea0025f601c03b6668cd75a5ea

Observation a2a8c277-e738-4e85-b9bd-732e42193099 · outbound

This paper cites LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.416440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.416440Z digest=sha256:dfb87e2450ac2dab0af4169d1389727822757bb92a9b954e9460aa5d16a62cca

Observation bd8639f7-bc2e-41b5-8c96-f5c1e861a5b0 · outbound

This paper cites an unresolved cited work.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.422849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.422849Z digest=sha256:33622fd08c619ec9888bfa5fbc38177f2adfd97d64ea0406c2ede2a0debbeea5

Observation 43c14ce9-5f8a-4e73-bf90-58096c7406c1 · outbound

This paper cites online" 'onlinestring :=.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL online" 'onlinestring :=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.428196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.428196Z digest=sha256:b6d383c922120aa467b13c12e8960d65cd3e122ec032429d1dfb89f84d1cf1a0

Observation 5f10d5a3-57b7-431d-ba84-a96b80a47512 · outbound

This paper cites write newline.

One Missing Piece for Open-Source Reasoning Models: A Dataset to Mitigate Cold-Starting Short CoT LLMs in RL write newline

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:30:40.435211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:30:40.435211Z digest=sha256:5d0fd19ad4dbd19b6b25a13c80a93696e025941dd540d47a31d0b8a3d7071853

Pith citing papers

No inbound Pith citation observations are available.