Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:36:35.797792Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 4 inbound Pith citation observations for arXiv:2506.04821.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:36:35.797792Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T05:46:26.938277Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
21 of 21 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 60bb10ba-e173-40fe-a364-2bbf2a34bc19 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning H.; Ol s \'a k, M.; Yang, X.; Nguyen, H.; Menegali, M.; Jung, J.; Verma, V.; Le, Q
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a740ac3-fda4-4759-89b2-cbc2857280b9 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Training Verifiers to Solve Math Word Problems
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 823dc748-acc7-43ca-9ddc-209620e0dde3 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1682acc7-86f5-49fb-b6bd-999fe06814c6 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Puzzle Solving using Reasoning of Large Language Models: A Survey
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db483bbe-601a-4225-a1e6-25cf69b12a15 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eda3e062-da76-44d8-b78c-2af5ad068bfd · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6dee929-919a-4567-b871-7e1eaa86b420 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Measuring Mathematical Problem Solving With the MATH Dataset
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddb1a080-64b2-4276-aa7c-5464b4f3a081 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0974290-344d-4535-8144-9be15832edb5 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 436dfa2e-86e0-4c0d-ad85-be963254fe7c · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning AIMO-2 Winning Solution: Building State-of-the-Art Mathematical Reasoning Models with OpenMathReasoning dataset
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 013e1b47-08f1-413b-91cb-09753a8c9a6d · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Instruction Tuning with GPT-4
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80a8c158-7aef-491f-89e5-9a5cb7030b09 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 975b78e7-bd62-4b63-be5b-a04e75c8c1de · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning DeepSeek-Prover-V2: Advancing Formal Mathematical Reasoning via Reinforcement Learning for Subgoal Decomposition
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f1a336d-4493-44be-abff-52b47b38eadc · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb642055-069f-49a9-a58c-65798539348c · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf4644b6-9ade-4cd6-be97-82153afcb6d6 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80cd305b-df40-4fd1-a4b1-858eaebb26a4 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 887d9805-7ec3-461b-a203-4cb49fa4379a · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af310fad-6d59-4dd3-bc84-cfa9d1623182 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning Leanabell-Prover: Posttraining Scaling in Formal Reasoning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6724016-5825-423b-99c1-3eb68387ef20 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning , " * write output.state after.block = add.period write newline
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b5b9d75-a6e3-4e33-a920-9d2850aa4bf7 · outbound
LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning write newline
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 623f87b7-4031-403d-98a2-f37cb1f0fdf9 · inbound
Do LLMs Overthink Basic Math Reasoning? Benchmarking the Accuracy-Efficiency Tradeoff in Language Models LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 999d9f8e-21e8-4887-9e79-2dd05598ca63 · inbound
PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 07ecb561-819d-4d51-8b35-05de69aa6164 · inbound
Step-by-Step Optimization-like Reasoning in LLMs over Expanding Search Spaces LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fda495d6-d9a8-4a96-a2ac-0282e47ae368 · inbound
Robots Need More than VLA and World Models LogicPuzzleRL: Cultivating Robust Mathematical Reasoning in LLMs via Reinforcement Learning
Reference 150
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.