Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:45:42.411640Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 16 inbound Pith citation observations for arXiv:2507.07400.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:45:42.411640Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T04:58:39.204380Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
40 of 40 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation ddc67721-7ff7-45d2-a005-486c7bc7f317 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows React: Synergizing reasoning and acting in language models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a60f3c9d-5f10-40f5-95ac-be25f82fab4d · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Reflexion: Language agents with verbal reinforcement learning.Advances in Neural Information Processing Systems, 36:8634–8652, 2023
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55921259-8778-4990-a660-09036a7a2743 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7914f024-1158-4497-b3e8-1fb12c7a18a9 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Camel: Communicative agents for" mind" exploration of large language model society
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d96f7d36-be33-47db-9b86-ca3418007a30 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows PEER: Expertizing Domain-Specific Tasks with a Multi-Agent Framework and Tuning Methods
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 90d6b685-6b40-4beb-9549-a1b1c9b8e75f · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7eb983f-f5fe-46e0-b19b-ab1d3674e4a6 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Gptswarm: Language agents as optimizable graphs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bc9def31-20cf-4e04-a077-65e5c6cd55e1 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows AFlow: Automating Agentic Workflow Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c2ce70a-94c3-4c42-affe-fe97b883e641 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Very Large-Scale Multi-Agent Simulation in AgentScope
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce19083b-5ba9-4f51-b18a-a88d0bb19648 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Cognify: Supercharging Gen-AI Workflows With Hierarchical Autotuning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96a5f89f-c5d8-4ad5-abee-3a93b3096963 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Efficient memory management for large language model serving with pagedattention
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bda1cd92-185a-472c-9c29-957d9edb573a · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Gonzalez, Clark W
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c7c4aa97-e930-41cd-bf27-9bdc2e4bce7c · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows TensorRT-LLM
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 687a0462-43bb-40c4-8c89-d89cf734eba8 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Automatic Prefix Caching
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 74408f15-ecf0-4635-b034-f7e1da15f41e · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb353792-c84a-45fa-98b5-ff09cd5c5a15 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows {Cost-Efficient} large language model serving for multi- turn conversations with {CachedAttention}
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6426a51c-213d-4d06-b5d7-2c56a4c8a539 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Generative agents: Interactive simulacra of human behavior
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96815315-538e-4beb-b77a-fa09f0deda9a · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows A survey on large language model based autonomous agents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cf33364-7348-4f83-9dcf-2c133442380a · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ec7bd6f-7b16-49e6-be9b-08466e82c288 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Code Generation with AlphaCodium: From Prompt Engineering to Flow Engineering
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 007a7681-db51-4dd4-9127-da403a12d622 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efa75a61-7f40-4e16-8f3e-6bc7ec6816df · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows MAGE: A Multi-Agent Engine for Automated RTL Code Generation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0af19f0d-9703-4ccb-8a29-fe0a9137a7e4 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows ChatDev: Communicative Agents for Software Development
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 071ae83c-0559-41d4-9882-9ab8bc5e9250 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Improv- ing factuality and reasoning in language models through multiagent debate
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c39ff46b-ee40-459a-8bda-3d5cb2e53ba3 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Accelerating Large Language Model Decoding with Speculative Sampling
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a0484f4-020a-4727-8d2e-9342846f4c11 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Fast inference from transformers via speculative decoding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ba568b7-6fb8-4c60-afe1-3312768270a2 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Prompt lookup decoding, November 2023
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ef4c5fae-3d8f-48b2-a03c-032fc212b8eb · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Efficient streaming language models with attention sinks
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cf829135-f430-4acc-8754-43b503bed880 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Efficiently Scaling LLM Reasoning with Certaindex
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218cb1c7-c682-42ad-a1f2-aa2bae8309cc · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Orca: A distributed serving system for {Transformer-Based} generative models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fc98626-c13a-466f-aff9-12bf892ad6e8 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Fast Distributed Inference Serving for Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d38722-36b3-4010-bb3f-fde70f928c6d · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Andes: Defining and Enhancing Quality-of-Experience in LLM-Based Text Streaming Services
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39572001-5ffb-47af-b551-e1bd8474c506 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Stateful large language model serving with pensieve
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 519ae05f-2cce-46bd-92ab-866c62479560 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows InferCept: Efficient Intercept Support for Augmented Large Language Model Inference
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f300aba2-02cf-48da-a93b-cbe612a86f26 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Autellix: An Efficient Serving Engine for LLM Agents as General Programs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99c684cf-6b44-4846-b990-b15ee6860f83 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Parrot: Efficient serving of {LLM-based} applications with semantic variable
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 72716aa1-a196-4ab0-9289-d329a85fc8d1 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows LangGraph
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9e15dc4e-6d66-4013-92a9-70b26b4662fe · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Building effective agents
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 13f354ff-166e-4bb5-ace0-ee2b53f11b46 · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows AgentScope: A Flexible yet Robust Multi-Agent Platform
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 887322ec-44c8-4927-a54c-ed5fc257900b · outbound
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows Cut the Crap: An Economical Communication Pipeline for LLM-based Multi-Agent Systems
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26d3d1de-fff5-4a5c-9e1a-38e9331db713 · inbound
Parallelizing Tool Execution and LLM Generation for Low-Latency Agent Serving KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8504fea4-ac1a-4c6f-ad49-df4739e91feb · inbound
ForkKV: Scaling Multi-LoRA Agent Serving via Copy-on-Write Disaggregated KV Cache KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9bd72f02-e53d-4bdd-a705-6e85b90aa4a2 · inbound
Hive: A Multi-Agent Infrastructure for Algorithm- and Task-Level Scaling KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ed36350b-6157-449b-bb3f-bb64896b7591 · inbound
PolyKV: A Shared Asymmetrically-Compressed KV Cache Pool for Multi-Agent LLM Inference KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d380eb5e-e830-48b5-bdc7-32b31d024ce3 · inbound
AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 903503ab-0c69-45a8-bbe7-430b6ddc9c2a · inbound
PRISM: Fast Online LLM Serving via Scheduling-Memory Co-design KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 3d13b212-72be-4504-8766-0d483000950c · inbound
VeriCache: Turning Lossy KV Cache into Lossless LLM Inference KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation d7ec5bef-2697-4fe9-be9a-fa4b5f8e50a3 · inbound
Not All Tokens Are Worth Caching: Learning Semantic-Aware Eviction for LLM Prefix Caches KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 061ca800-78f4-4b34-b6dd-832b565c2b15 · inbound
VineLM: Trie-Based Fine-Grained Control for Agentic Workflows KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22dab550-d3c8-4bb7-9865-251ed0bc1eaf · inbound
Resident KV Claims: A Conformance Contract for Future Reuse under Active KV Pressure KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 57a31b54-9b25-48a8-92a6-c6e5d950bfc7 · inbound
VikingMem: A Memory Base Management System for Stateful LLM-based Applications KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation aec54df1-77a5-4946-b6ae-20d28c31d373 · inbound
Streaming Communication in Multi-Agent Reasoning KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 4ae630f1-8f27-4dba-8663-26525168fd5d · inbound
Streaming Communication in Multi-Agent Reasoning KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8cb190-cd0a-4d88-8fe9-f9124eebd6a0 · inbound
SmoothAgent: Efficient Long-Horizon LLM-Based Agent Serving with Lookahead Context Engineering KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation bd742d16-48fb-4143-9404-57f04d3397a2 · inbound
Workload-Aware Caching for Multi-Agent Systems KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a65ee2b-195e-4223-ba1d-95c47b4382bf · inbound
Rethinking AI Cloud Infrastructure for Agentic Serving Systems with the Aries Experimentation Framework KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.