Pith. sign in

Paper Citation Record · LEDGER

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

As of 7 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.04771.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04771 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:07:02.537106Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3f7401b1-e66e-4a35-82f9-b95cfafdcffa · outbound

This paper cites xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.541065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.541065Z digest=sha256:83f3e8a56cbc212f5d53dad264372a4ca197e14b39bfb6d0ccf48c58190a4e6e

Observation 23052acf-81c3-4cd9-a423-071ef0845d51 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.657571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.657571Z digest=sha256:51a855861c732c50489c27480eb22eaebf89d6af962268cb727a01c476479c14

Observation fc845da2-56de-4b98-bd43-00d02756bb98 · outbound

This paper cites InFindings of the Association for Computational Linguistics: ACL 2025, 21539–21564.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InFindings of the Association for Computational Linguistics: ACL 2025, 21539–21564

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:04.195709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:00.768651Z digest=sha256:4d437e25a75cbfc15e7dbfb0b2da717eff8f9ae3d679073a8c906699e4da77eb

Observation 2eb70d07-703f-49c1-860c-461436066125 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.917179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.917179Z digest=sha256:0d659851a146c22840f21e52956e33af697eb827bf3d8c9b555c442068c5b9e3

Observation a0cfa119-adc9-4e40-ae43-174ab5186bf9 · outbound

This paper cites InFindings of the Asso- ciation for Computational Linguistics: ACL 2026, 17051– 17064.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InFindings of the Asso- ciation for Computational Linguistics: ACL 2026, 17051– 17064

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.795551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:01.067843Z digest=sha256:e8838efa058b9e8acf4f694a47842cb811c29cb4ac73750b96635108532bc0eb

Observation 49a984e6-6c44-4cab-b23a-d2eec4bfe0fe · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.139839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.139839Z digest=sha256:6917dd5223caf7bda226381fa3d6e54b3ef24f72f5d7c91e4c5737e6b748dc3e

Observation 4910617c-f844-4da1-8e5b-cac80449faa1 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.200943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.200943Z digest=sha256:542b4875881310437d4b22d2819f678f728256fc2936ddb12e767375ec7b0cc0

Observation 1c088ec5-3752-46d1-85fc-26178c07292f · outbound

This paper cites Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:03.397949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:01.285318Z digest=sha256:bee6b6ab08b7834b4f790ef895a2c4c8e0886c95095038e8785593501a1c661b

Observation e5d6b42c-067e-41e5-9e1e-efc67620bace · outbound

This paper cites OpenAI o1 System Card.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning OpenAI o1 System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.346455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.346455Z digest=sha256:d652ded3d1f15fa7429036da5d4ae5ede3c181f0131de747c56e83cb585ef0cc

Observation 4a0002bf-6639-46c2-bd20-44e0d630bdf3 · outbound

This paper cites Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.424636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.424636Z digest=sha256:fd797c59df79e23469919104ea7725cc0ec0e057a97be6f1cd4cf9a921e98245

Observation 59fc4700-df1c-4469-b7c4-14d4160378d4 · outbound

This paper cites Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.503753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.503753Z digest=sha256:ec4ca3c38eee10b1117024e1ce1e29239375778beaf80301b3d0a58384ce169a

Observation db03dc27-dd94-410c-aad5-6827c9ad12df · outbound

This paper cites RouteLLM: Learning to Route LLMs with Preference Data.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning RouteLLM: Learning to Route LLMs with Preference Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.589061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.589061Z digest=sha256:6754f7688c29fa48446808e73457146acb9a734394e7c0b331a53f5ea3fe3f2e

Observation ebdb3aa2-b3ba-4ebc-b39e-6308be8b0f6a · outbound

This paper cites InProceedingsofthe2025ConferenceonEmpirical Methods in Natural Language Processing, 8021–8040.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InProceedingsofthe2025ConferenceonEmpirical Methods in Natural Language Processing, 8021–8040

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.673043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.673043Z digest=sha256:4066f5c2e2bde9efe7a2ee3a6a345871e24047619e4241843e2e6bc10aaf33f1

Observation 805f982d-50e4-4eb7-acbc-6268eb00bdc9 · outbound

This paper cites The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.843310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.843310Z digest=sha256:9cb6663c53e6dc35d70654a7852915694b6f92979339bcc135d35401a0cf1995

Observation 3d84b4c5-e307-46b6-ab4b-e2182e7b85a1 · outbound

This paper cites Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.912817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.912817Z digest=sha256:b440aca8df86cce353df8b31ab421969940ff2ff86b9cd2653a835f16f3bbb35

Observation c3d48be3-8fe0-4a3b-af0c-ffec19f407d7 · outbound

This paper cites D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.982112Z digest=sha256:0cc3438e8dbebd46be51a1f3db3eba861162b4136a69ee7a79a99703929ee21f

Observation 549eccfd-5e7e-4c97-8cf8-65d628511f39 · outbound

This paper cites Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.070191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.070191Z digest=sha256:42ce170cdb3e411aecd2f8f9fc1848ff9111d14bc0238a0f1c7b34221f413a4e

Observation e0cfc844-8e03-42d8-bbfe-f86823a3b740 · outbound

This paper cites Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.153112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.153112Z digest=sha256:567d9e09f30e68119b4e4cebd16715cb22dc5e608327680016017460dd1bedc3

Observation 93959c88-548e-4514-83e4-ab896f40bf64 · outbound

This paper cites IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:02.935451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:02.216605Z digest=sha256:cecda8e6fece624372747f0845a52f949cd4ee3729242284699113e3e7c52503

Observation 3f881354-a945-4269-a6ce-f1f8beb34ef0 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.311998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.311998Z digest=sha256:2170c69696389619ded646ae1737af9913fc17e0bb2c6b02af259a167957b005

Observation b09b1b72-a8b6-4ee0-b50d-d90f19642436 · outbound

This paper cites TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.376649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.376649Z digest=sha256:2c02ace9aa95ff75c05c42eb9614825aa38ca93142a221cc201f5d2c6f32b6fb

Observation 6f0be876-e02f-4e36-a750-4c5dbe3c5d05 · outbound

This paper cites Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:02.699059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:02.442534Z digest=sha256:5d8cfb607e876c27b985cff271fe2c830f27c33af40239c062bc160e25fe71e1

Observation 39209269-5347-455c-a632-467254c6e068 · outbound

This paper cites InForty-firstinternationalconferenceonmachine learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InForty-firstinternationalconferenceonmachine learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.659023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:02.537106Z digest=sha256:9f3f6ab084f655d93e407cc1b709c70441164e7fb9f72115f86d3f2c3be0f76e

Observation c13a16ea-9bbc-4ca1-8f42-ebcf4ab8af63 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.834739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.834739Z digest=sha256:c4f2d7a2062e46a0790c1b170dcdb6b3c16981d3c40c1a878531fe3de40d3e68

Observation a614b709-e4d9-4874-87f2-c23eaaf79da8 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.757657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.757657Z digest=sha256:e6256fe5e67e87d64bc9a6bea8e50d61f75222a81725393062e234213d40e2d4

Observation f19a1cce-233f-4218-b01c-f50a35e21906 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.437098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.437098Z digest=sha256:a706aa8c029e1a68ca0df72124543b8b93a6a028a595b558edba7525bb01fc68

Observation 7d523cba-690a-46f3-9833-5a0d50c287b0 · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.309946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.309946Z digest=sha256:21c80d486b43cf2d2e7e4f388229b1143a841865154ed6428d6b5a0af6091996

Observation 5db40434-4347-4daa-9515-279f3e5efbb0 · outbound

This paper cites Fu,Y.;Chen,J.;Zhuang,Y.;Fu,Z.;Stoica,I.;andZhang,H.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Fu,Y.;Chen,J.;Zhuang,Y.;Fu,Z.;Stoica,I.;andZhang,H

Reference 2026

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.996112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T17:07:01.005716Z digest=sha256:f89124d0f300f9e4cbe72c3e3738f8b51719e8ae8c3c4a304623012771edfad5

Pith citing papers

No inbound Pith citation observations are available.