Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T01:12:20.864362Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2605.07316.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-11T01:12:20.864362Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T18:39:40.459071Z
A source-named dated measurement, never combined with another source.
Source: cited_works
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 91e96891-59ce-46fb-bd14-4f1f141d98c6 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Chain-of-thought prompting elicits reasoning in large language models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c324769-3218-465d-b791-1f58c8aaf5d0 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dea649c4-6c98-4252-a624-c6b667204ecc · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 382a8b7e-f7e9-42ba-a869-99cea33ba63d · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4c7adc71-39bf-4e3d-b1f4-5aa3f0f4f6c8 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Between Underthinking and Overthinking: An Empirical Study of Reasoning Length and correctness in LLMs
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a17cc85f-a316-4bef-adcc-640fb215f5ba · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Reasoning on a Budget: A Survey of Adaptive and Controllable Test-Time Compute in LLMs
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e516f3c-ad6d-4590-99d6-7dcaf053f4fe · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bc60c29f-e64e-4c1d-8d8c-ddc4ec6841ed · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 523ea300-591f-4fe0-8279-3dd872289080 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 858f7524-ac87-4673-934c-914e639251a6 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Shorterbetter: Guiding reasoning models to find optimal inference length for efficient reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9c3f1357-e3e3-422c-884b-cff07b2a11f8 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Length-unbiased sequence policy optimization: Revealing and controlling response length variation in rlvr.arXiv preprint arXiv:2602.05261
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2bb18992-52ae-4298-8fb8-c4e1a17e6a09 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training On the Optimal Reasoning Length for RL-Trained Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d851b4fc-b9d4-4e7c-8dc5-3e7b345ffe8e · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e29ec674-0fbe-463b-b15c-0e7904a9939a · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training S-GRPO: Early Exit via Reinforcement Learning in Reasoning Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 82a05dd3-6a45-44a9-834b-0ac172f5e82f · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Explore briefly, then decide: Mitigating llm overthinking via cumulative entropy regulation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0809db36-c9f9-448e-83e8-83e98ef3a1c4 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Optimizing Length Compression in Large Reasoning Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 30a50862-f941-4f8c-8585-f242735d6ccf · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill?
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d996b2c1-faa1-4792-bec6-3842afdbf886 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training 2025 , journal =
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e2efe663-dd40-4762-8852-1ebd6706bb78 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 11a22fa5-d1c3-421b-85d2-0dedcfd2c553 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Training language models to reason efficiently
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 621aa321-98df-4cf1-8d43-93fc91d81e1d · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Learn to Reason Efficiently with Adaptive Length-based Reward Shaping
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 902f5ff5-d894-4cdf-8cb7-c4071c5d5742 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e7c4f40f-4e81-4b8c-a167-3a09be53fdc0 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Deepcompress: A dual reward strategy for dynamically exploring and compressing reasoning chains
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5898e4d1-b6b1-4d22-af7f-e5e754e60ac1 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training SmartThinker: Progressive Chain-of-Thought Length Calibration for Efficient Large Language Model Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4369648a-efbc-44ca-b123-55488db6a1dc · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Thinking Fast and Right: Balancing Accuracy and Reasoning Length with Adaptive Rewards
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7aac7b12-34de-4f66-a37a-85817e8be6e5 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Adaptthink: Reasoning models can learn when to think
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 87f66fed-3d7f-449b-8072-97ab65fa8de9 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Chain of Draft: Thinking Faster by Writing Less
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a2038e33-a1d3-4ef7-803f-99fd6fbbb808 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Qwen3 Technical Report
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0e1542cd-3c0e-492b-aef6-270839fab4df · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Easyr1: An efficient, scalable, multi-modality rl training framework
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0307b7ff-7dd8-42d0-9f09-d58c28a74063 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Hybridflow: A flexible and efficient rlhf framework
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e0fa0a3f-c747-47a6-bb6a-3321c8fce8a3 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training MathArena: Evaluating LLMs on Uncontaminated Math Competitions
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6220c4ba-6214-40c5-b81d-b303e91ae715 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Let's Verify Step by Step
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1b513867-1f1a-488a-b858-ed9731639766 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Training Verifiers to Solve Math Word Problems
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d8cef271-bb77-44f9-9d9a-b0937f66f8ac · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fcb8b05d-5457-4aba-a0e5-5f60bc40d167 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Mmlu-pro: A more robust and challenging multi-task language understanding benchmark.Advances in Neural Information Processing Systems, 37:95266–95290
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation af6d637b-0215-42d2-a44e-eda3bd6f84d2 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 25800b19-0182-4413-b403-5c108fc0ac15 · outbound
Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training Guidelines: • The answer [N/A] means that the paper does not involve crowdsourcing nor research with human subjects
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ecf09c03-2927-4a36-9243-d9819033cdab · inbound
Distilled Reinforcement Learning for LLM Post-training Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.