Pith. sign in

Paper Citation Record · LEDGER

Workload-Aware Caching for Multi-Agent Systems

As of 21 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2607.20495.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.20495 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T11:23:00.475610Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 929236fa-a99b-4695-9162-c38c21a8de79 · outbound

This paper cites SWE-bench: Can language models resolve real- world github issues?.

Workload-Aware Caching for Multi-Agent Systems SWE-bench: Can language models resolve real- world github issues?

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.244902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.244902Z digest=sha256:fffc29bfd136576fd2fe6a437371b0b314bfd96c3f298b38c12ea312d9879050

Observation 7d4107bf-4e3e-48a7-8c5b-39d3d9c2eece · outbound

This paper cites Aiopslab: A holistic framework for evaluating ai agents for enabling autonomous cloud,.

Workload-Aware Caching for Multi-Agent Systems Aiopslab: A holistic framework for evaluating ai agents for enabling autonomous cloud,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.323738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.323738Z digest=sha256:0924955d4f787b36459ef45f367dfa58f6810b3319c51ae85f0d5a82ed098ac0

Observation 5def4b6b-1ca8-4ebe-ab25-d1e3d5cc24f7 · outbound

This paper cites The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery.

Workload-Aware Caching for Multi-Agent Systems The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.388148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.388148Z digest=sha256:26ab08435874e0445928abdefa8b7cf01f406c0c7ef48e0ba594ccee3763aa5d

Observation 50d6676c-1960-47a3-a93b-3a5d5c23f305 · outbound

This paper cites The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search.

Workload-Aware Caching for Multi-Agent Systems The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.497764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.497764Z digest=sha256:c0559414e25408042abc2bfa4ce58f5e41f11e8a14f0b6d3773e7615716cc2c5

Observation 8a747cde-1bdf-4960-8215-318906cf4cd9 · outbound

This paper cites React: Synergizing reasoning and acting in language models,.

Workload-Aware Caching for Multi-Agent Systems React: Synergizing reasoning and acting in language models,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.575077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.575077Z digest=sha256:1133ec8a51f9498a731fc076c5d7848f531c3e4694980f78f1e6f14c1eebece3

Observation 90061d91-f8a1-4ad9-b68b-4397ec455986 · outbound

This paper cites A survey on large language model based autonomous agents,.

Workload-Aware Caching for Multi-Agent Systems A survey on large language model based autonomous agents,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.675764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.675764Z digest=sha256:3f9601e9b9a8cfe38c9c1de10cd1263bfc0d1097b5629118eb11f58f8e1f4da1

Observation 85b5b12b-36e8-4b0b-ab69-02ab238a07b0 · outbound

This paper cites Efficient Inference for Large Reasoning Models: A Survey.

Workload-Aware Caching for Multi-Agent Systems Efficient Inference for Large Reasoning Models: A Survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.753787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.753787Z digest=sha256:ec9b97c5abfd617dad95d8b46ab8b5767b62df1a4cc16cc981bcbdb5cedee355

Observation 3405c3e0-4a79-4b3d-9bf2-ad2f27e3fbb3 · outbound

This paper cites Attention is all you need,.

Workload-Aware Caching for Multi-Agent Systems Attention is all you need,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.818911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.818911Z digest=sha256:1a2f6a213ce11e5f44cbe1031a002760b28f600684ce259a37222db8774fe566

Observation f7a1abed-baae-4dc3-bca2-14da69c7c662 · outbound

This paper cites Efficient memory management for large language model serving with pagedattention,.

Workload-Aware Caching for Multi-Agent Systems Efficient memory management for large language model serving with pagedattention,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.880620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.880620Z digest=sha256:34b2bcc91397b74b91ce4226efca05d68f163d4a335bf692a13141e170843e90

Observation 49f31b2d-98e7-4fac-9407-2043cfe631bd · outbound

This paper cites SGLang: Efficient execution of structured language model programs,.

Workload-Aware Caching for Multi-Agent Systems SGLang: Efficient execution of structured language model programs,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:56.950669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:56.950669Z digest=sha256:01c249d66b339c918d218a036640fd65f7fd196f5dbcf54a5e400977373e74ee

Observation bd742d16-48fb-4143-9404-57f04d3397a2 · outbound

This paper cites KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows.

Workload-Aware Caching for Multi-Agent Systems KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.020958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.020958Z digest=sha256:6e8b0bfa1d4648dedd8cb00d5900803bbb89ba18347fe2132ffae2c59783fd66

Observation 2051d56d-0e40-4490-bb31-e527fbc1032e · outbound

This paper cites KVCOMM: Online cross-context KV-cache communication for efficient LLM-based multi-agent systems,.

Workload-Aware Caching for Multi-Agent Systems KVCOMM: Online cross-context KV-cache communication for efficient LLM-based multi-agent systems,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.090168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.090168Z digest=sha256:9199f741d8631fe058859c270ea646e256f32ac11b0ff4b2cf4802647b173b5e

Observation 4cea82c2-b216-4db5-9c8a-7d1847b8794e · outbound

This paper cites GPTCache: An open-source semantic cache for LLM applications enabling faster answers and cost savings,.

Workload-Aware Caching for Multi-Agent Systems GPTCache: An open-source semantic cache for LLM applications enabling faster answers and cost savings,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.161637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.161637Z digest=sha256:72e1e5aed9f623655f22f4c31c0d149342e59c410d2deb50c602b3fdf5b1a63f

Observation 7aa66d5f-b2f7-4a5a-9352-8a7667b76e68 · outbound

This paper cites Billion-scale similarity search with GPUs,.

Workload-Aware Caching for Multi-Agent Systems Billion-scale similarity search with GPUs,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.219423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.219423Z digest=sha256:b006b282f05a2fae26d28825f822cd1311d4967637526ce5fb1bc6ca3131802f

Observation 32ac36b4-b2b6-40ba-8ce0-2473484ad29f · outbound

This paper cites Cortex: Achieving low-latency, cost-efficient remote data access for llm via semantic-aware knowledge caching,.

Workload-Aware Caching for Multi-Agent Systems Cortex: Achieving low-latency, cost-efficient remote data access for llm via semantic-aware knowledge caching,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.381659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.381659Z digest=sha256:99a4dd86b9aa2ae9d1354c4a92d95b48f99c57f9ba2a4b5d346ab92f9035ea6f

Observation 0d6c4385-9bf7-473c-9e70-4ea0ff94bb21 · outbound

This paper cites Agentic plan caching: Test- time memory for fast and cost-efficient LLM agents,.

Workload-Aware Caching for Multi-Agent Systems Agentic plan caching: Test- time memory for fast and cost-efficient LLM agents,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.453149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.453149Z digest=sha256:58b6ac1a995bc9980b7566c2c67962169aa4c97af8aa8aec4e2a11cea07a67dc

Observation 1e81c84c-8d48-4e90-b085-c25c0b521bbc · outbound

This paper cites Semanticalli: Caching reasoning, not just responses, in agentic systems,.

Workload-Aware Caching for Multi-Agent Systems Semanticalli: Caching reasoning, not just responses, in agentic systems,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.523453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.523453Z digest=sha256:10bfc939d3b67491c6397e12b347e773892ca577291d88b8dc2abac0e6cf9e79

Observation fbbfc3a4-a13a-485c-8f27-5ed4be2e024b · outbound

This paper cites Lrc: Dependency-aware cache management for data analytics clusters,.

Workload-Aware Caching for Multi-Agent Systems Lrc: Dependency-aware cache management for data analytics clusters,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.708196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.708196Z digest=sha256:de9dccff493b54b8f2255f5fb6be8fc75c41b967dcd14ef23b63733bd3e00c7d

Observation ea4c7af5-5500-4db9-aaa6-7002102504cb · outbound

This paper cites Resilient distributed datasets: A {fault-tolerant}abstraction for{in-memory}cluster computing,.

Workload-Aware Caching for Multi-Agent Systems Resilient distributed datasets: A {fault-tolerant}abstraction for{in-memory}cluster computing,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.823282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.823282Z digest=sha256:e56512aecdd92cf999effdd828f07203bd9366abd54ba80282b24adfac01473b

Observation 33598cc5-f8d4-4797-9b12-33ea4034a3aa · outbound

This paper cites Lerc: Coordinated cache management for data-parallel systems,.

Workload-Aware Caching for Multi-Agent Systems Lerc: Coordinated cache management for data-parallel systems,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:57.938531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:57.938531Z digest=sha256:e854037d07bf0c3a9a0acf889051e76632a16d1ca14006b282fadced255a29ce

Observation 4cfa7766-1f1a-45b0-988d-61df84c25e52 · outbound

This paper cites S/c: Speeding up data materialization with bounded memory,.

Workload-Aware Caching for Multi-Agent Systems S/c: Speeding up data materialization with bounded memory,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.063849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.063849Z digest=sha256:2fcff046dd44df35767bcd30244b6a3916e0c8a904d11a8b3cdd61e4ebe64991

Observation 6aaff067-d55c-4781-a05b-0c954a1bd20b · outbound

This paper cites Real-Time Analytics by Coordinating Reuse and Work Sharing.

Workload-Aware Caching for Multi-Agent Systems Real-Time Analytics by Coordinating Reuse and Work Sharing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.148809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.148809Z digest=sha256:f0bdea7451628d5ac52c7d0b8f30fbee5f2e187d407bcbf4abbb252bed5416a5

Observation eedabd05-de26-4cd0-b203-b0ade40d9974 · outbound

This paper cites S-dag: A subject-based directed acyclic graph for multi-agent heterogeneous reasoning,.

Workload-Aware Caching for Multi-Agent Systems S-dag: A subject-based directed acyclic graph for multi-agent heterogeneous reasoning,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.262762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.262762Z digest=sha256:6fdd57db3ed0d825814d381bebaa586e448e0e58aaa01eb504d0198c823b2d82

Observation b3ece365-7ef1-4941-a231-0b5975135091 · outbound

This paper cites Agentnet: Decentralized evolutionary coordination for LLM-based multi-agent systems,.

Workload-Aware Caching for Multi-Agent Systems Agentnet: Decentralized evolutionary coordination for LLM-based multi-agent systems,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.365011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.365011Z digest=sha256:b5e4a875ede8041f0734b67b5937326659f7570382e1e56efd6bc582fed575bf

Observation 74e5828f-e0d4-4f34-ac9b-336eb4aa8105 · outbound

This paper cites Benchmarking agentic workflow generation,.

Workload-Aware Caching for Multi-Agent Systems Benchmarking agentic workflow generation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.483202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.483202Z digest=sha256:c759c9e5fac2ffd8215fb5506bd5f93f53632513375ab191e8fcb5ff055ddbbb

Observation 3509fcfa-6633-41b4-9230-22a4839a9883 · outbound

This paper cites Kairos: Low-latency multi-agent serving with shared llms and excessive loads in the public cloud,.

Workload-Aware Caching for Multi-Agent Systems Kairos: Low-latency multi-agent serving with shared llms and excessive loads in the public cloud,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.628928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.628928Z digest=sha256:f4eb7e3352a12f628c469f0054e898046f7c7a536306eab4f20f65663cf969ab

Observation 00ed3a32-6ad2-4bc9-ac69-e528b27ada66 · outbound

This paper cites Towards end-to-end optimization of llm-based applications with ayo,.

Workload-Aware Caching for Multi-Agent Systems Towards end-to-end optimization of llm-based applications with ayo,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.843180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.843180Z digest=sha256:8e7a3282bf8977cc5a012d67c3b5cc42e5e10cf4feb2eb8ea933ffe2c5c571dc

Observation 26115895-48f4-4560-b761-93f7312ff447 · outbound

This paper cites AFlow: Automating agentic workflow generation,.

Workload-Aware Caching for Multi-Agent Systems AFlow: Automating agentic workflow generation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.949619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.949619Z digest=sha256:2b1092b34dcd0caf6f72e4879b02c1f631e1ed680cd7ab499b9b78a859a835e7

Observation 0a3e0a95-cba2-42ff-9729-71f81d33eff5 · outbound

This paper cites Automated design of agentic systems,.

Workload-Aware Caching for Multi-Agent Systems Automated design of agentic systems,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.010703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.010703Z digest=sha256:80a12b41342c11a1199a3a10fe1b77b8e407f42c05461cc738e2ce26f14fcafc

Observation 14e41f65-5d2b-4709-8cfc-bdcb8e92980e · outbound

This paper cites DynaSaur: Large Language Agents Beyond Predefined Actions.

Workload-Aware Caching for Multi-Agent Systems DynaSaur: Large Language Agents Beyond Predefined Actions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.122377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.122377Z digest=sha256:a6b728d00897ba434adf3fb456b81d57b16d5c39c430b42bd7e5e411f30f0982

Observation 81d7bae8-3ac3-4930-894c-15833ee89db3 · outbound

This paper cites Slidevqa: A dataset for document visual question answering on multiple images,.

Workload-Aware Caching for Multi-Agent Systems Slidevqa: A dataset for document visual question answering on multiple images,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.288250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.288250Z digest=sha256:a0a26e25a49ab8c913d385aef8bbae3227a96bf9d53e031af9da327cabc91794

Observation a5c121cd-a98d-42da-bb1a-c495f9e5f330 · outbound

This paper cites Hierarchical multimodal transformers for Multi-Page DocVQA.

Workload-Aware Caching for Multi-Agent Systems Hierarchical multimodal transformers for Multi-Page DocVQA

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.426670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.426670Z digest=sha256:7c539f9c82292b683a63ec62e84b2994db5f9dc8012ab0d85e834aba81f503f7

Observation 6e0dcf2b-9b32-46d8-873c-7bc6fb172c8b · outbound

This paper cites Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis,.

Workload-Aware Caching for Multi-Agent Systems Video-mme: The first-ever comprehensive evaluation benchmark of multi-modal llms in video analysis,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.590150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.590150Z digest=sha256:04c48f45025d5d28b541ddeb0fd00be15cb54070a38818331d8985a969915915

Observation 5c62be42-8fd7-41a9-a95d-88abb906aec5 · outbound

This paper cites Plan-and-act: Improving planning of agents for long-horizon tasks,.

Workload-Aware Caching for Multi-Agent Systems Plan-and-act: Improving planning of agents for long-horizon tasks,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.700138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.700138Z digest=sha256:8fd58514d69c2adbdbcdedeaaf9fc3ceead59348b3eefa990c20433227c0af17

Observation ada20323-bc55-4679-aec6-8b0e0b1a5360 · outbound

This paper cites ARC: A Self-Tuning, low overhead replacement cache,.

Workload-Aware Caching for Multi-Agent Systems ARC: A Self-Tuning, low overhead replacement cache,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.809932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.809932Z digest=sha256:8730bc6814efd899757896ba84a1804a07d6e60b2c941c2c998f7b3df9d64c86

Observation 319da7b0-28c4-4b75-956f-6677510b5675 · outbound

This paper cites An llm compiler for parallel function calling,.

Workload-Aware Caching for Multi-Agent Systems An llm compiler for parallel function calling,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:59.966611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:59.966611Z digest=sha256:3d02365093e5ebe9a9442eb3dac24bdfe543e637c3d4a19922b8fd7e743eedb7

Observation 262c7428-b251-4ebb-8c8a-ca3f8047d0fb · outbound

This paper cites Qwen2.5 Technical Report.

Workload-Aware Caching for Multi-Agent Systems Qwen2.5 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T11:23:00.070441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:23:00.070441Z digest=sha256:0927b701be1e3b1e29f71ccb65df8cb315b09ecde2b7cc9ecaa385247432087c

Observation aeeb0b4c-e08d-4d81-9138-afe8f276eee5 · outbound

This paper cites Ollama: Run AI models locally,.

Workload-Aware Caching for Multi-Agent Systems Ollama: Run AI models locally,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T11:23:00.195827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:23:00.195827Z digest=sha256:042510619a472f7de46065792ab4ec6be8e021ffbb3c9be6759225ebecced5a6

Observation 5aad65d3-e372-4ba7-b045-4e4fdebed334 · outbound

This paper cites Can Increasing the Hit Ratio Hurt Cache Throughput? (Long Version).

Workload-Aware Caching for Multi-Agent Systems Can Increasing the Hit Ratio Hurt Cache Throughput? (Long Version)

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T11:23:00.299874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:23:00.299874Z digest=sha256:37ebccc955123895072b229e7118f23ad68f5b58ff899f6b708c1f6c1a971e7e

Observation 413ff1bb-9d1c-4b29-9fc2-1c85d56c5a60 · outbound

This paper cites Mistral Small 3.2,.

Workload-Aware Caching for Multi-Agent Systems Mistral Small 3.2,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T11:23:00.475610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:23:00.475610Z digest=sha256:9201a6d88acc903d01e1aa9fd460364bb7bedb02c9043fee84ee21f0178d239f

Observation f7ab6e29-14f5-4dec-a243-44f72d744e44 · outbound

This paper cites Kairos: Low-latency Multi-Agent Serving with Shared LLMs and Excessive Loads in the Public Cloud.

Workload-Aware Caching for Multi-Agent Systems Kairos: Low-latency Multi-Agent Serving with Shared LLMs and Excessive Loads in the Public Cloud

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T11:22:58.727229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T11:22:58.727229Z digest=sha256:b715166ec0e83186bf4e7dae16993fcd70e5c009270a922192f6bf4eb8ef3108

Pith citing papers

No inbound Pith citation observations are available.