Pith. sign in

Paper Citation Record · LEDGER

Can A Gamer Train A Mathematical Reasoning Model?

As of 8 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2506.08935.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08935 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:03:22.166335Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3cbc36e8-0509-45b7-a539-35431a402049 · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

Can A Gamer Train A Mathematical Reasoning Model? Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.393043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.393043Z digest=sha256:35cb275eb13cec4182440902eaf164d68995a1323b385c7ae8e932371d773e27

Observation 544a754f-e876-4d6b-994c-6ceef92d32e3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Can A Gamer Train A Mathematical Reasoning Model? LoRA: Low-Rank Adaptation of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.566380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.566380Z digest=sha256:6ec08d65ada81c50352406e64aa7cad3e6c0714d33b9c75924e634da5c25a553

Observation 04986e11-49e8-4dd6-8ace-19c6b744973f · outbound

This paper cites s1: Simple test-time scaling.

Can A Gamer Train A Mathematical Reasoning Model? s1: Simple test-time scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.780247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.780247Z digest=sha256:75e51e279d0136fff739ad446eec2ce0ce9bfeaba74bef3fd1fe240732001daa

Observation d2866b02-8b67-4a79-a955-e77869c69d5c · outbound

This paper cites https://github.com/Jiayi-Pan/TinyZero.

Can A Gamer Train A Mathematical Reasoning Model? https://github.com/Jiayi-Pan/TinyZero

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:03:22.834765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:03:21.813208Z digest=sha256:6d3dc67f313222867ed1fead5afb92d8e906ef92f48e06f31cfe84d6588111c1

Observation 53c2a64f-45d1-4f5a-922b-6db29596c5a8 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Can A Gamer Train A Mathematical Reasoning Model? DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.885652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.885652Z digest=sha256:65db5220f4da19132e1e86730c7f42dbf8d84b27ceffd7e750a1580f0b602ff8

Observation 75a29c77-a704-4f57-a59c-4acba6a27bd1 · outbound

This paper cites https://novasky- ai.github.io/posts/sky-t1.

Can A Gamer Train A Mathematical Reasoning Model? https://novasky- ai.github.io/posts/sky-t1

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:03:22.629238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:03:21.978742Z digest=sha256:a376d859744d0a31e9efab0522e294694b6c26891130687ef0f15eb744fb5620

Observation d9a622f5-a6be-4aaf-8554-b7ede7922f75 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Can A Gamer Train A Mathematical Reasoning Model? Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:22.072238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:22.072238Z digest=sha256:c4e870a22a64f515f73ef2c3b175d7ecfbaffaaff308bd555eed7a96f0236bec

Observation 3f308fce-c9de-428b-9bb8-328833fb15f1 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Can A Gamer Train A Mathematical Reasoning Model? Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:22.166335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:22.166335Z digest=sha256:48da0e521e095a2576dfb490104b461fa0e273318c37533932b3727419fb68e8

Observation a06b9743-567f-473e-a005-628e5a6b21be · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Can A Gamer Train A Mathematical Reasoning Model? Measuring Massive Multitask Language Understanding

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.491475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.491475Z digest=sha256:cd3b70ae9461f6ad0a4a81a141d43f633df79a3d0686f8f93ff6b85b82e8f4ff

Observation d8486f86-9c96-4722-b765-98df0273a797 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Can A Gamer Train A Mathematical Reasoning Model? Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.225062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.225062Z digest=sha256:686aa0ceee27ce98f406653fed8c3cc43098836cf75e3c463661da945b5b8201

Observation 50b2e901-02a1-4f37-904a-fa46d9a881be · outbound

This paper cites Solving Quantitative Reasoning Problems with Language Models.

Can A Gamer Train A Mathematical Reasoning Model? Solving Quantitative Reasoning Problems with Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.738356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.738356Z digest=sha256:af67d57290cb37b4582870509d5956a8735e74bd48f4fe22cb9a676029fcc619

Observation 79ab7925-0d95-4377-8d10-0c77e92cd292 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Can A Gamer Train A Mathematical Reasoning Model? FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.279182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.279182Z digest=sha256:59b0aa4ba3bc02eda1878a541a7aa19d6ea69331733a9665fc8be353d9032f35

Observation 1fa0c303-c60d-4cd1-9227-57027454b319 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Can A Gamer Train A Mathematical Reasoning Model? Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.654953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.654953Z digest=sha256:041c1ffca46eaeca1deab4835d97cd9247fd144f44acd56e7f3c6fda8dd53fb6

Observation d67c9a9d-fc43-45a7-a8bc-3fa3bd0ceb49 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Can A Gamer Train A Mathematical Reasoning Model? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.349189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.349189Z digest=sha256:fc4672cfeebe04cb1dd21d37fb6da1a075835ddc9ff0ee36d56b7d37ef673ea5

Pith citing papers

No inbound Pith citation observations are available.