Pith. sign in

Paper Citation Record · LEDGER

Understanding Reasoning from Pretraining to Post-Training

As of 6 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2607.16097.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.16097 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T21:26:36.017309Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved51
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 78c1ecb4-d656-4c8d-ba3f-de8d4f0e6350 · outbound

This paper cites Human-aligned Chess with a Bit of Search.

Understanding Reasoning from Pretraining to Post-Training Human-aligned Chess with a Bit of Search

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.196479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.196479Z digest=sha256:2926a5714dde8da00f2b7123e8377fd5c948f318c6ac17dc30e0915b09a52dd7

Observation 24b31eb4-1cc8-4f77-b48a-95641d23a195 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Understanding Reasoning from Pretraining to Post-Training Advances in Neural Information Processing Systems , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.245778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.245778Z digest=sha256:4609652adfa8b18c69ec2eaae54f387ebee206e47d74dd5be064555eef1d5728

Observation b0bc2b55-2ce2-4f5c-936b-f1642c36a28b · outbound

This paper cites Qwen3 Technical Report.

Understanding Reasoning from Pretraining to Post-Training Qwen3 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.325889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.325889Z digest=sha256:d5f4fa98c362ee1bc560cd8ace521483335414dacb3a20fd91d533d907ef1dbd

Observation 4ca9d9e2-8735-488a-ab05-c3bda2e0d27b · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Understanding Reasoning from Pretraining to Post-Training HybridFlow: A Flexible and Efficient RLHF Framework

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.388270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.388270Z digest=sha256:89abfd6222a9a09c67ffa842ab585cd7ab575b050b3f09be25d2e8ba042c31b8

Observation cebc0c53-dddb-4c49-8821-ec843289af4d · outbound

This paper cites Scaling Laws for Neural Language Models.

Understanding Reasoning from Pretraining to Post-Training Scaling Laws for Neural Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.456353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.456353Z digest=sha256:aaca981ee55dbdb876e65e8658714faf356b53ead39cee765022bd2ecfcc552b

Observation c354859d-0e8c-448e-8406-5892327cbe32 · outbound

This paper cites Training Compute-Optimal Large Language Models.

Understanding Reasoning from Pretraining to Post-Training Training Compute-Optimal Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.517791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.517791Z digest=sha256:784044665215bf581297084afbcda027dcf65c7f1ba61d9e6eaaa4d4c0b792f0

Observation 45fc3c09-1649-4265-b169-8500af46fc1a · outbound

This paper cites arXiv preprint arXiv:2509.21016 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2509.21016 , year=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.556078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.556078Z digest=sha256:78c1158378bbd5d7d661c90700a82ffbdfa0f7504a522f0f6c7471cb59a52352

Observation 626e3a08-4e60-4c35-907e-4d49baadecca · outbound

This paper cites Physics of Language Models: Part 2.1, Grade-School Math and the Hidden Reasoning Process.

Understanding Reasoning from Pretraining to Post-Training Physics of Language Models: Part 2.1, Grade-School Math and the Hidden Reasoning Process

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.620292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.620292Z digest=sha256:f0e0c0534f4221b5a19e31b824c7bb1a2961deb8ea1758e7f3ba55b0c56f5390

Observation b6ae222a-ec23-46e0-b510-42e7fb47b6a1 · outbound

This paper cites arXiv preprint arXiv:2509.25123 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2509.25123 , year=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.665546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.665546Z digest=sha256:696be63f7c000eaee93fce4a2e97d92912e762415682b07b84f9c244b47ddcc9

Observation b84ad023-1d0f-4acb-8a1b-f1c422cb1437 · outbound

This paper cites The Art of Scaling Reinforcement Learning Compute for LLMs.

Understanding Reasoning from Pretraining to Post-Training The Art of Scaling Reinforcement Learning Compute for LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.726200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.726200Z digest=sha256:cb2169e1045826e7f45adb803efb7c707c53d8d5f20bd38a417d58f237515cdb

Observation c70e37f8-63be-4dba-b5d7-78eeefc9767f · outbound

This paper cites arXiv preprint arXiv:2512.07783 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2512.07783 , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.746663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.746663Z digest=sha256:a851d756312d01010fc5f464f2507fe58591bdf0cfea13519c5d6d532271de6e

Observation 92239618-bc87-45f7-950c-e08f769ec455 · outbound

This paper cites arXiv preprint arXiv:2506.16029 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2506.16029 , year=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.826713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.826713Z digest=sha256:531d89de1e3b56f6825dafd2a2a4930719907e79d2fe19dcc4667e703459e81a

Observation 1a3a5358-4373-4775-9e9e-8879069c263a · outbound

This paper cites Forty-first International Conference on Machine Learning , year=.

Understanding Reasoning from Pretraining to Post-Training Forty-first International Conference on Machine Learning , year=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:32.962357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:32.962357Z digest=sha256:e44514c6d259b93290e061aff4b46bf19f8b39c7320732e4bf2b3f3bc9e4e6bf

Observation 5b0e64e7-b806-40ff-a5b4-d7e81e974fcf · outbound

This paper cites When Can LLMs Learn to Reason with Weak Supervision?.

Understanding Reasoning from Pretraining to Post-Training When Can LLMs Learn to Reason with Weak Supervision?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.106406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.106406Z digest=sha256:79f61a8b7d348665ec79439dfb608363698616d0815a30d0fcc380cf2ba159dc

Observation a45fb5e0-7c75-4a33-82a4-4f8f70bba4ed · outbound

This paper cites Nature , volume=.

Understanding Reasoning from Pretraining to Post-Training Nature , volume=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.207178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.207178Z digest=sha256:4fa40121cf2113d0d39052833713ab88a98de6dd5176ee082d8141ae4c15e932

Observation 9b6e379a-522a-4a4b-8ec7-34e3ae3ab709 · outbound

This paper cites Large Language Model Guided Tree-of-Thought.

Understanding Reasoning from Pretraining to Post-Training Large Language Model Guided Tree-of-Thought

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.265609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.265609Z digest=sha256:ed95b559b33856a7c766dad8a343875422781858d83e25f74c0c549676e6328c

Observation 08f6bcbd-825e-4aa6-a44a-bc9b8da73d75 · outbound

This paper cites Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm.

Understanding Reasoning from Pretraining to Post-Training Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.359440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.359440Z digest=sha256:9708dd2cba5a8b6a7c21b0777a930f5828382ae03a096b6e269b00575f51ddef

Observation 4a5557aa-a76f-4593-86ce-5065c85fd2cc · outbound

This paper cites Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?.

Understanding Reasoning from Pretraining to Post-Training Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.421318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.421318Z digest=sha256:a55f059bbc4cadb774ef6be29161e3cacf546753fcee6d01d3c161b325c7ecad

Observation 4799bc1c-e8c3-4fe1-95fe-fecb6d52b39d · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Understanding Reasoning from Pretraining to Post-Training Advances in Neural Information Processing Systems , volume=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.498829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.498829Z digest=sha256:3e5fa98782d41037f472715569356d6782ae67985ad0e3771affa60550586ca4

Observation 409cf8b7-2a5f-430f-8756-d6a2e22989b6 · outbound

This paper cites arXiv preprint arXiv:2510.15020 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2510.15020 , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.572158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.572158Z digest=sha256:7ae94c697b9c524d54af031801d7ee6ff4578b9a6b63d455934fa20e35400df4

Observation 8806f68c-2a92-4957-9c86-840f6788134a · outbound

This paper cites Reasoning with Sampling: Your Base Model is Smarter Than You Think.

Understanding Reasoning from Pretraining to Post-Training Reasoning with Sampling: Your Base Model is Smarter Than You Think

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.626064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.626064Z digest=sha256:1ab04c5a76c5ee6a38d94709076cc00ddae5583c2493a7e06a2706748fd69ce2

Observation 8a1576f0-0758-4d8b-93da-12f169afbe46 · outbound

This paper cites arXiv preprint arXiv:2603.24844 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2603.24844 , year=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.688075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.688075Z digest=sha256:e0d06ee811e1d02f049fa2720fc3075717017636a4bdf7302428886ac215e0a6

Observation 3108f45c-5ee9-4bf9-99c1-a35a01ca66e3 · outbound

This paper cites International Conference on Learning Representations , volume=.

Understanding Reasoning from Pretraining to Post-Training International Conference on Learning Representations , volume=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.743899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.743899Z digest=sha256:7861677c9d088315399da6f6952b9f767ef2676280392d78084a28e03200b007

Observation b0ff0808-fe3f-49b6-8fff-7ecd19d83944 · outbound

This paper cites arXiv preprint arXiv:2510.03264 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2510.03264 , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.794436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.794436Z digest=sha256:f224ce9ce19a5455bb8fdeddb8967c926a9172d4fca77a8015a582d24bb66b4b

Observation 9a5d67da-066c-4e9a-9def-35c2a8f57ba4 · outbound

This paper cites RL Excursions during Pre-Training: Re-examining Policy Optimization for LLM training.

Understanding Reasoning from Pretraining to Post-Training RL Excursions during Pre-Training: Re-examining Policy Optimization for LLM training

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.858637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.858637Z digest=sha256:04c1c348929273a0184d9ed1466eaefb910b1e8480128c6a672000d441e06c98

Observation ea635115-ac45-4d79-8c01-04e2eec82f04 · outbound

This paper cites Overtrained Language Models Are Harder to Fine-Tune.

Understanding Reasoning from Pretraining to Post-Training Overtrained Language Models Are Harder to Fine-Tune

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.911069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.911069Z digest=sha256:4c795034c92f1325eaae9f91fedce83a9d77d593c7ebe5a1430f20c806c4b4a5

Observation 23efc518-9d68-4987-9e7c-d30238763bf6 · outbound

This paper cites DeepSeek-V3 Technical Report.

Understanding Reasoning from Pretraining to Post-Training DeepSeek-V3 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:33.973602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:33.973602Z digest=sha256:ffb0e6ce5602664683985626f357be573f191e4de8ab03e7561172cfd2b4bd14

Observation 784ade2e-5bac-4cbe-abaf-ac0d8f2f6811 · outbound

This paper cites 2003 , publisher=.

Understanding Reasoning from Pretraining to Post-Training 2003 , publisher=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.025800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.025800Z digest=sha256:d04d1bf981a18718164211220c0889a5c2a6189f345b02588c93a10e3b1823ab

Observation fdab7b28-675e-48c1-853b-0c0ea6516391 · outbound

This paper cites an unresolved cited work.

Understanding Reasoning from Pretraining to Post-Training Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.113763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.113763Z digest=sha256:7a994a0973584738c81dc6779472d65a3a464bb92ed56871bf18865392338d9c

Observation 6d29db7a-b219-4134-ae32-310a7918d01d · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Understanding Reasoning from Pretraining to Post-Training Training Verifiers to Solve Math Word Problems

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.227557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.227557Z digest=sha256:63df00b4fbdb0804ac01c3b3b9f8d966fb8182d711aa6b47b1426b3ffbf83de6

Observation 77660dd6-614d-41b8-b6d4-c78b36dbc429 · outbound

This paper cites Is Best-of-N the Best of Them? Coverage, Scaling, and Optimality in Inference-Time Alignment.

Understanding Reasoning from Pretraining to Post-Training Is Best-of-N the Best of Them? Coverage, Scaling, and Optimality in Inference-Time Alignment

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.271964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.271964Z digest=sha256:4f5bd9cdfcadfe23b6c63be9ab156a2a50d3b81c1ec596bcfd7afe9501fdc771

Observation 8bd31739-576a-4b1c-85d7-c12eb9916ab5 · outbound

This paper cites Olmo 3.

Understanding Reasoning from Pretraining to Post-Training Olmo 3

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.361691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.361691Z digest=sha256:ba9d3e1128c96f34a7ebf22593466478f96b72e1431cabdeb32ec5ba87526b79

Observation e725fcb1-eca8-4e91-8e3f-0c858c3bdd0b · outbound

This paper cites Notion Blog , volume=.

Understanding Reasoning from Pretraining to Post-Training Notion Blog , volume=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.431145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.431145Z digest=sha256:789d15ce301a0d16711556de35f64096589e337551d1b9541fa8c923bd038f92

Observation cd460ae0-bd66-41f7-a510-52a33a7802e0 · outbound

This paper cites 2 OLMo 2 Furious.

Understanding Reasoning from Pretraining to Post-Training 2 OLMo 2 Furious

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.509727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.509727Z digest=sha256:65c29a29fa2e962d9e883161f3e70333f5b055396af0d02c25c23d7d0b237619

Observation 937a5172-5451-4180-b85c-fd5f55e941b1 · outbound

This paper cites Hugging Face repository , howpublished =.

Understanding Reasoning from Pretraining to Post-Training Hugging Face repository , howpublished =

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.573030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.573030Z digest=sha256:0d42a0e2eff99b600c8dd285ceec5a26872dea30b0b9d4696d22c7d48ee058e2

Observation b6844668-ff3a-4f18-bbd6-daf0aabebbad · outbound

This paper cites arXiv preprint arXiv:2512.15489 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2512.15489 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.633423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.633423Z digest=sha256:c6ccdcaba56caa906c367885c957cc4a4bada05dd5563afc6e7991d6d7bac1ed

Observation 28e949f9-a828-4357-bdb5-b64bf1caffa6 · outbound

This paper cites International Conference on Learning Representations , volume=.

Understanding Reasoning from Pretraining to Post-Training International Conference on Learning Representations , volume=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.713790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.713790Z digest=sha256:6ab6f2b2bd43fc5fb2c775d1b8d105c1d72864aa432cd0b6dedf5fed968f39e0

Observation 53615201-a48e-4812-9c48-2a8b6d14d0b2 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , articleno =.

Understanding Reasoning from Pretraining to Post-Training Proceedings of the 41st International Conference on Machine Learning , articleno =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.786841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.786841Z digest=sha256:db46ae7599fd72c576875d561ea5df769156afa804002d535b00f8e277d34704

Observation 709bf90f-e0d2-48f8-b25b-5349a2feb3bb · outbound

This paper cites arXiv preprint arXiv:2604.01411 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2604.01411 , year=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.855268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.855268Z digest=sha256:4f1e68addfcf9f656fea1df29634c2a3ddd0ed87869623ff317ec0920a1de8d4

Observation 9ab0fc64-d5c4-4525-b062-90ff395db5a3 · outbound

This paper cites arXiv preprint arXiv:2503.19551 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2503.19551 , year=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:34.944396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:34.944396Z digest=sha256:aefa637454aa163c26f4e4ccc9cb15eff46f05e064de82892c9d49b44ee75906

Observation e1aa124e-f5a7-40f8-8bbb-47722f4c3f60 · outbound

This paper cites arXiv preprint arXiv:2603.12151 , year=.

Understanding Reasoning from Pretraining to Post-Training arXiv preprint arXiv:2603.12151 , year=

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.036633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.036633Z digest=sha256:fef18961000b8bc5c9f4847de4111327c9da7e6106f10e32d72323d03c73a988

Observation 4785bc8e-996f-4b65-a777-71c970f5b7cc · outbound

This paper cites an unresolved cited work.

Understanding Reasoning from Pretraining to Post-Training Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.100440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.100440Z digest=sha256:1b9f735acdd642f1289c2488290f9d7e1352f8e036369586c2a62997884a6694

Observation 5b9ac02a-dbd0-464c-a7fa-f04b92ac1b3f · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Understanding Reasoning from Pretraining to Post-Training Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.191672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.191672Z digest=sha256:c550af1287b4db717b4021520fee4ab1740d680314149865d3d7aadfb8bc1f7d

Observation 9d0ee7bb-462d-4f8d-b77c-7d7be8780496 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Understanding Reasoning from Pretraining to Post-Training DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.256751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.256751Z digest=sha256:1ae275b426fe9023f234a9a839a90654a67f765e83dab22bfbdd07924606dcf4

Observation 3e830194-4084-4a75-8762-4b6451a1e9db · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

Understanding Reasoning from Pretraining to Post-Training SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.371952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.371952Z digest=sha256:d3bdd22e31bad81a8bdbda6b72dfa5a6d545aa5a62198cdeaf10d709d8e5764c

Observation 0efc0bd5-37b3-490e-86aa-af80a7317620 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Understanding Reasoning from Pretraining to Post-Training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.468782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.468782Z digest=sha256:ede29d580bc946bf936107e072280c87fe31d112b6833d6e48c24388444382ee

Observation 473d8c2d-476f-4065-836f-32d1657426be · outbound

This paper cites Google AI , volume=.

Understanding Reasoning from Pretraining to Post-Training Google AI , volume=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.533985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.533985Z digest=sha256:6ab7cf24cf57f41bf5469346735ac2a5d197349b90d6fbdf6994fb56652f30d3

Observation 84f78f46-bfb5-49b4-852a-90d2f2a76c47 · outbound

This paper cites nature , volume=.

Understanding Reasoning from Pretraining to Post-Training nature , volume=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.635555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.635555Z digest=sha256:72507132edad8a43a4c30f11e17c121fa352f80c18570711bdff4f1e5c795225

Observation 6e532969-9a8e-4aa0-953d-7b2c4afde200 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Understanding Reasoning from Pretraining to Post-Training Advances in Neural Information Processing Systems , volume=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.793794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.793794Z digest=sha256:1c1fdb61b541cd6545a6516303526fe0723e371924ac60ac0677fc6e70e25a87

Observation 27fc059a-610b-4a7b-897f-6a6bfd8f15ce · outbound

This paper cites Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision.

Understanding Reasoning from Pretraining to Post-Training Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:35.909813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:35.909813Z digest=sha256:768c55b2726214cb0985e06d8d1474b05769c1e4a6a8e2c0c97969416da0d5cf

Observation 7fa24bc0-fc18-4e03-b808-19433a53a2e6 · outbound

This paper cites Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset.

Understanding Reasoning from Pretraining to Post-Training Nemotron-CC-Math: A 133 Billion-Token-Scale High Quality Math Pretraining Dataset

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T21:26:36.017309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:26:36.017309Z digest=sha256:e8b939ac0c8236297602ebe5ed9e4b4e992ecc296c621312971eee5e38375919

Pith citing papers

No inbound Pith citation observations are available.