Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T08:58:08.181911Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2607.20507.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T08:58:08.181911Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 90b49b74-0e22-45b6-8ac5-94023f6cc8ee · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Advances in Neural Information Processing Systems , volume =
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16baf39c-1956-4da2-9250-d913920b9975 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Transactions on Machine Learning Research , year =
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 268a6c57-1b20-4e03-acd2-6185d20afc28 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference 2023 , volume =
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24610a3e-bf75-4c98-9182-2942538ff241 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference 2023 , url =
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 516aff8b-85a5-4c06-ac77-23b9d7442bd7 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Proceedings of the 40th International Conference on Machine Learning , pages =
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 675aff8d-7a30-4427-b801-c4a7f47ca0e1 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Accelerating Large Language Model Decoding with Speculative Sampling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e44686-6a96-4b7e-9eaa-42f3d3d44e65 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference 2024 , publisher =
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8659404b-334c-4bd0-9e3a-650210a9c112 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference and Chen, Deming and Dao, Tri , booktitle =
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c987e5ba-0a3b-4b70-8003-9fd8ce8a9ffb · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference 2024 , volume =
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26ce685d-ddb2-48cf-8eb1-4e963efc8dff · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Break the Sequential Dependency of
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 942d0c99-998d-464c-87ca-35504046ffb9 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 452c4d0e-da0a-4830-84e3-f69b0a3a79bf · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Findings of the Association for Computational Linguistics: ACL 2024 , pages =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09be17be-f8f3-44b2-b617-ee20682ec72e · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference 2023 , address =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 068aa273-416e-483d-8142-29866a5d3d24 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Proceedings of the Seventh Annual Conference on Machine Learning and Systems , year =
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d07fdbd-a640-4bb8-975e-5ed2c1cd8115 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference MeanCache: User-Centric Semantic Caching for LLM Web Services
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36b6fbfb-565b-4601-9a2c-2b3cae74170e · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Advances in Neural Information Processing Systems , volume =
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da5eb0de-d486-4885-be94-7280f4fa4f76 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference KVShare: An LLM Service System with Efficient and Effective Multi-Tenant KV Cache Reuse
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b708ea27-3567-42ca-a72d-27a74d2b3016 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Qwen3 Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b169a79f-5a0a-43b9-83c6-e2dbf5aa5980 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing , pages =
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f652980-3369-4f4b-8804-b24b5a33ed3a · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Advances in Neural Information Processing Systems , volume =
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac43dd8a-f21a-47ef-9732-0a01388b6ae8 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Advances in Neural Information Processing Systems , volume =
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e81a133e-06aa-4a36-ad3c-ad4b9f5c0ae2 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f664783c-749e-496c-8d20-92e82c2a7108 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Advances in Neural Information Processing Systems , volume =
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb5594d6-8ec9-416d-a30c-17b0611d0036 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Scaling Laws for Neural Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff9d53e7-7b09-4ae1-bb68-44a35ba17889 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference Advances in Neural Information Processing Systems , volume =
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf4bba0c-d969-4a6a-97fb-4ea8a2a3d06d · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference The Llama 3 Herd of Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5498f6bd-3387-4ca8-a9a4-e225290f5fea · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference DeepSeek-V3 Technical Report
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a653bc0-de73-4415-a241-44863d1ff648 · outbound
MiniCache: Reusable Program Caching with Small Model Interfaces for Efficient LLM Inference FinLoRA: Benchmarking LoRA Methods for Fine-Tuning LLMs on Financial Datasets
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.