Pith. sign in

Paper Citation Record · LEDGER

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

As of 7 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.04771.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04771 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:07:02.537106Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3f7401b1-e66e-4a35-82f9-b95cfafdcffa · outbound

This paper cites xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.541065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.541065Z digest=sha256:c530ad90234b72c2e4e98313d4445a1cf899248048c63595acf28cc93f500bc2

Observation 23052acf-81c3-4cd9-a423-071ef0845d51 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.657571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.657571Z digest=sha256:0da1bb511287703cdec7a57886d2b32972f6c1e45b93aa1650a51bd252dea317

Observation fc845da2-56de-4b98-bd43-00d02756bb98 · outbound

This paper cites InFindings of the Association for Computational Linguistics: ACL 2025, 21539–21564.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InFindings of the Association for Computational Linguistics: ACL 2025, 21539–21564

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:04.195709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:00.768651Z digest=sha256:3eb446cd60cf9163bc92acf702934f65c1d1be239dbda4a718d8f95bbdec8e25

Observation 2eb70d07-703f-49c1-860c-461436066125 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.917179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.917179Z digest=sha256:5c6d1e5d0e2e7affb2b231c0b8b906c364f856e7939ce8601c9295b4c4ca6c82

Observation a0cfa119-adc9-4e40-ae43-174ab5186bf9 · outbound

This paper cites InFindings of the Asso- ciation for Computational Linguistics: ACL 2026, 17051– 17064.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InFindings of the Asso- ciation for Computational Linguistics: ACL 2026, 17051– 17064

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.795551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:01.067843Z digest=sha256:18d0d54f02c63aafbc84588d89189c8abc54699cb0006e0a710335ce5e0804b9

Observation 49a984e6-6c44-4cab-b23a-d2eec4bfe0fe · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.139839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.139839Z digest=sha256:3f888df5be657478acfb0da231858247cf85c8453acdacf11d597b8e93ac4a3e

Observation 4910617c-f844-4da1-8e5b-cac80449faa1 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.200943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.200943Z digest=sha256:d6d2c77485b66958fea5559c1d00f940738fa91b0c53b9c95f12b13de44c1897

Observation 1c088ec5-3752-46d1-85fc-26178c07292f · outbound

This paper cites Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:03.397949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:01.285318Z digest=sha256:b8264d2915f445731d79850e9f72a757280e5032f52904d38e7f9d4b7c9f57c9

Observation e5d6b42c-067e-41e5-9e1e-efc67620bace · outbound

This paper cites OpenAI o1 System Card.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning OpenAI o1 System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.346455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.346455Z digest=sha256:bc9984381435227018c674439c267aa0d9106b4fde319444beff512cbc7848a4

Observation 4a0002bf-6639-46c2-bd20-44e0d630bdf3 · outbound

This paper cites Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.424636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.424636Z digest=sha256:04ee401db10e61280db35b75df5da9565dfd8f3cc0729366ecc209f3bdd2cd08

Observation 59fc4700-df1c-4469-b7c4-14d4160378d4 · outbound

This paper cites Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.503753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.503753Z digest=sha256:3e024b4b86cc78a3c2cfe393a3f1e874cc62c15e89114efc039df03b42fd2734

Observation db03dc27-dd94-410c-aad5-6827c9ad12df · outbound

This paper cites RouteLLM: Learning to Route LLMs with Preference Data.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning RouteLLM: Learning to Route LLMs with Preference Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.589061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.589061Z digest=sha256:17fb46789e5b185157748cb4dc72ed3a9fc7de9a62f9b300cbc081efc7cbeb6f

Observation ebdb3aa2-b3ba-4ebc-b39e-6308be8b0f6a · outbound

This paper cites InProceedingsofthe2025ConferenceonEmpirical Methods in Natural Language Processing, 8021–8040.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InProceedingsofthe2025ConferenceonEmpirical Methods in Natural Language Processing, 8021–8040

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.673043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.673043Z digest=sha256:d18efbfb18a519a07603716fa0ca48c7bdfcbcee3cb69ff82d4748b139e5cdd6

Observation 805f982d-50e4-4eb7-acbc-6268eb00bdc9 · outbound

This paper cites The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.843310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.843310Z digest=sha256:4ec8c6c0042dfc9adc8299922563f9df9d039c3235bc6b0891e7ca4d040170c0

Observation 3d84b4c5-e307-46b6-ab4b-e2182e7b85a1 · outbound

This paper cites Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.912817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.912817Z digest=sha256:a76148fafdfc9b18254df62ed053c88dd9104ddb7d80a5987a9c783ec90f2ff7

Observation c3d48be3-8fe0-4a3b-af0c-ffec19f407d7 · outbound

This paper cites D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.982112Z digest=sha256:75f200902071a38c8a622831d7a8e270a838c37110553d6d9dc6f160e675ad62

Observation 549eccfd-5e7e-4c97-8cf8-65d628511f39 · outbound

This paper cites Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.070191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.070191Z digest=sha256:65b63f39df6df12593cd02f43c90975d8af6ed6847bd19eb07b47aefc84bd2ef

Observation e0cfc844-8e03-42d8-bbfe-f86823a3b740 · outbound

This paper cites Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.153112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.153112Z digest=sha256:7fb210e41f6874c49273aae835114a69c34f37b7f12dd339b278a4bf6de64165

Observation 93959c88-548e-4514-83e4-ab896f40bf64 · outbound

This paper cites IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:02.935451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:02.216605Z digest=sha256:34ff20d293fd931bb5899388fe18b4263372091c21ca612fb21756960e0300e8

Observation 3f881354-a945-4269-a6ce-f1f8beb34ef0 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.311998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.311998Z digest=sha256:d85d48ca484ccd562012a8b266e6bdc46dbadc38f4739660f6825ca67d9e1bd1

Observation b09b1b72-a8b6-4ee0-b50d-d90f19642436 · outbound

This paper cites TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.376649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.376649Z digest=sha256:0ba58dc7941c52e4fbcff2ef7944ca927570b87510265b7758d5e2956d2c0c6d

Observation 6f0be876-e02f-4e36-a750-4c5dbe3c5d05 · outbound

This paper cites Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:02.699059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:02.442534Z digest=sha256:6fd33efcfde997305c19c45c6a62b8eb0cc51b55a0026a1545874eb7fd2fdaf5

Observation 39209269-5347-455c-a632-467254c6e068 · outbound

This paper cites InForty-firstinternationalconferenceonmachine learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InForty-firstinternationalconferenceonmachine learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.659023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:02.537106Z digest=sha256:f256763cbc177ae0cf13005674fe9dd729ab3478c02291035521f34d006473db

Observation c13a16ea-9bbc-4ca1-8f42-ebcf4ab8af63 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.834739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.834739Z digest=sha256:3b8cb540e10601910842a62e76f9e4fda2bea7cf7f18cc7f993786a959c7e069

Observation a614b709-e4d9-4874-87f2-c23eaaf79da8 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.757657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.757657Z digest=sha256:7213e1d24adbb829b92ba4c41578844965e992d224ee4d130ddc8df3dcd885e9

Observation f19a1cce-233f-4218-b01c-f50a35e21906 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.437098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.437098Z digest=sha256:7299c9823f1216d692e9b60f7beea4f9682f7bc5f845439712fa9305173ab2e4

Observation 7d523cba-690a-46f3-9833-5a0d50c287b0 · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.309946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.309946Z digest=sha256:9bdeb1cc6bc209e2080bcad853035bbe4bd2d08006c3ed14b2eed265c3d6e227

Observation 5db40434-4347-4daa-9515-279f3e5efbb0 · outbound

This paper cites Fu,Y.;Chen,J.;Zhuang,Y.;Fu,Z.;Stoica,I.;andZhang,H.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Fu,Y.;Chen,J.;Zhuang,Y.;Fu,Z.;Stoica,I.;andZhang,H

Reference 2026

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.996112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T17:07:01.005716Z digest=sha256:2e015abd29e00284c52d2136ce637a284bf93e43ea91c2bd48e93a8bceb8f655

Pith citing papers

No inbound Pith citation observations are available.