Pith. sign in

Paper Citation Record · LEDGER

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache

As of 16 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2505.10951.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10951 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:05:40.210126Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T18:48:03.257015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-09T06:10:42.888719Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved34
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ded81b9f-31f4-448f-a80e-0925d3907834 · outbound

This paper cites GPT-4 Technical Report.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.913974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.913974Z digest=sha256:9f0ab61c94f75a467b79e333eafa799ecb50dac98e1e2d63603a0948635a2144

Observation fa539310-8d35-4155-aa95-22eddf9766b8 · outbound

This paper cites Self-rag: Learn- ing to retrieve, generate, and critique through self-reflection.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Self-rag: Learn- ing to retrieve, generate, and critique through self-reflection

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.921411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.921411Z digest=sha256:454f33a0128df49ec7ee09c3644b3da71714e4f5013bb0861430c5af5d585b4b

Observation 338093e6-692a-4595-8ca4-23269946cf67 · outbound

This paper cites Improving language models by retrieving from trillions of tokens.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Improving language models by retrieving from trillions of tokens

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.930206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.930206Z digest=sha256:760e8ab93d917eadf0bd39b6ed84b147374efba870b6b9dd49f8cd71b6d9a4fe

Observation 7390e84e-b285-4011-9795-23188f3ef2a7 · outbound

This paper cites Hint on steroids: Batch query processing for interval data.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Hint on steroids: Batch query processing for interval data

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:41.047661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:39.940620Z digest=sha256:361f82e1973b858cea338db5cddda96490abd7f1e50f2e9300bfd965d69acc58

Observation 27d052fd-d155-4eb1-b79b-ef0528738962 · outbound

This paper cites Batch processing of top-k spatial-textual queries.ACM Transactions on Spatial Algorithms and Systems (TSAS), 3 (4):1–40, 2018.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Batch processing of top-k spatial-textual queries.ACM Transactions on Spatial Algorithms and Systems (TSAS), 3 (4):1–40, 2018

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:41.028978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:39.952008Z digest=sha256:08771b3d75308935ef359840aa5917285282afd5873e31deac49cd0b6852eb4a

Observation aaf5f625-e40c-46e3-8652-6502abbaf145 · outbound

This paper cites Palm: Scaling language modeling with pathways.Journal of Machine Learning Research, 24(240): 1–113, 2023.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Palm: Scaling language modeling with pathways.Journal of Machine Learning Research, 24(240): 1–113, 2023

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.964166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.964166Z digest=sha256:eb038171dd9b2612ed0408c8b3d95257e1cd640551b1ebaff15472682a0a3555

Observation fcc1ddaa-1076-49d9-800d-6cdd3d028ed1 · outbound

This paper cites Batch query processing for web search engines.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Batch query processing for web search engines

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.999525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:39.972250Z digest=sha256:01b05805cf7b52c300108f70a12740f74e74e3ec66bcc4736739e7a87fb286e9

Observation a5a3384f-58d3-4bc7-b80b-b38a8cd26384 · outbound

This paper cites From Local to Global: A Graph RAG Approach to Query-Focused Summarization.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache From Local to Global: A Graph RAG Approach to Query-Focused Summarization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.978823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.978823Z digest=sha256:4f2994d63b68fc6e8a5f659dd6d552b02c95c442dfecbfb15c162df64ea977b1

Observation 0516231f-37c8-46ee-9709-dea2b4168e25 · outbound

This paper cites A survey on rag meeting llms: Towards retrieval-augmented large language models.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache A survey on rag meeting llms: Towards retrieval-augmented large language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.984460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.984460Z digest=sha256:63bc07128c17d425cb3ee97af7dec1fc213167e3dad53bef8b7925dd45a038c2

Observation a3015774-f020-42b0-b9c7-c2219321f5cd · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.990943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.990943Z digest=sha256:49c993869e9f9981ed5dec0572ee0f4032b14d3f1636001c8fe83fdda5392200

Observation 3ea73ea1-8e8a-4677-accc-c196132f571f · outbound

This paper cites Prompt cache: Modular attention reuse for low-latency inference.Proceedings of Machine Learning and Systems, 6:325–338, 2024.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Prompt cache: Modular attention reuse for low-latency inference.Proceedings of Machine Learning and Systems, 6:325–338, 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:39.997566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:39.997566Z digest=sha256:06c7cb9e0e47ecaaf024d872c425d6602f41d95b66ce96d42884983d5898f599

Observation 2355bbf1-8c98-422f-81bb-99e2d63ad157 · outbound

This paper cites The Llama 3 Herd of Models.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.004342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.004342Z digest=sha256:5a19e18bd7036dd3e3d36759a1f1bfda5c097858e894043151191c73fc95cfb6

Observation 415fd710-4811-4472-9904-442ed2b7f627 · outbound

This paper cites Lightrag: Simple and fast retrieval-augmented generation.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Lightrag: Simple and fast retrieval-augmented generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.010606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.010606Z digest=sha256:8101310cef70f3119d502c592a976eb36cfc7449210088b0621070ebf14394c9

Observation 67d002f8-f6ce-44f6-8034-cdd7170b9787 · outbound

This paper cites Retrieval-Augmented Generation with Graphs (GraphRAG).

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Retrieval-Augmented Generation with Graphs (GraphRAG)

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.016023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.016023Z digest=sha256:dfacad808d3ac10750fd28a5ac7f11601da525d600f446a35b382cf20ae07636

Observation f4a68cd8-61ab-4971-81ba-24f221f054c5 · outbound

This paper cites G-retriever: Retrieval-augmented generation for textual graph understanding and question answering.Advances in Neural Information Processing Systems, 37:132876–132907, 2024.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache G-retriever: Retrieval-augmented generation for textual graph understanding and question answering.Advances in Neural Information Processing Systems, 37:132876–132907, 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.945687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.021819Z digest=sha256:3999ecda4ce912fa819b9195ff3cf94787bbe426f1a8b34886882c063fe538ae

Observation 0cb45d0f-1b42-4655-9b66-b94cde207428 · outbound

This paper cites RAG and RAU: A Survey on Retrieval-Augmented Language Model in Natural Language Processing.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache RAG and RAU: A Survey on Retrieval-Augmented Language Model in Natural Language Processing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.027436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.027436Z digest=sha256:f6085e530b88a8ac57b6db408836b4f52dcb6bd693a525ccb76cb5455b92d949

Observation bcb29c82-0f45-4778-85de-56f40d729b5b · outbound

This paper cites GRAG: Graph Retrieval-Augmented Generation.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache GRAG: Graph Retrieval-Augmented Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.033240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.033240Z digest=sha256:0937cbbb86074869bc78f01f1be5f934541ee785100376846c948166e394791b

Observation 210b45a2-b2f6-42a2-97d4-180bbdbd20dc · outbound

This paper cites A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions.ACM Transactions on Information Systems, 43(2):1–55, 2025.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache A survey on hallucination in large language models: Principles, taxonomy, challenges, and open questions.ACM Transactions on Information Systems, 43(2):1–55, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.040376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.040376Z digest=sha256:9bc877d6e974040fa34a2d39cd4a2086fab72eee3a6947fe8b47895e001f2681

Observation 6f6f3bdd-2e30-42a3-b430-3b07fe186382 · outbound

This paper cites Large language models on graphs: A comprehensive survey.IEEE Transactions on Knowledge and Data Engineering, 2024.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Large language models on graphs: A comprehensive survey.IEEE Transactions on Knowledge and Data Engineering, 2024

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.046628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.046628Z digest=sha256:9fb6b5eeb2f280c63fb50c396c788faa68944429d8deabd2fa86bf46e82079c3

Observation ea1b1d68-2841-461c-85f0-335452ef9fee · outbound

This paper cites RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.052335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.052335Z digest=sha256:d546d348f095ec0ac3f2ff50f2d3921066bbbfd6525ccc831726f6e2ffbc2aaf

Observation 413e8b20-0ae5-4c53-8f06-6aefecaa8415 · outbound

This paper cites Compute Or Load KV Cache? Why Not Both?.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Compute Or Load KV Cache? Why Not Both?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.059155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.059155Z digest=sha256:4fce01f0f5610c88dd5d737eacb1f6ad6fa496f1109b89e2e99d9c15b6b2d25d

Observation 1fba6331-989d-44a4-87a8-679101f2c911 · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks.Advances in neural information processing systems, 33:9459–9474, 2020.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Retrieval-augmented generation for knowledge-intensive nlp tasks.Advances in neural information processing systems, 33:9459–9474, 2020

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.065121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.065121Z digest=sha256:fcc0065f3b0fccc6d5961ec12393d03b3b276e66d55f3eec6fedfaa91002201e

Observation b2e7eb95-1421-4eca-822b-1d440e0a14a6 · outbound

This paper cites Sharedcontextbench: Evaluating long-context methods in kv cache reuse.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Sharedcontextbench: Evaluating long-context methods in kv cache reuse

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.892470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.070824Z digest=sha256:3c27a7e0edea8714e299a3e032c53856920bd413adc308320bd1c64b8326e4c1

Observation dbc7b100-7c69-486f-b884-7a3aaf4ef43b · outbound

This paper cites A Survey of Graph Meets Large Language Model: Progress and Future Directions.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache A Survey of Graph Meets Large Language Model: Progress and Future Directions

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.077890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.077890Z digest=sha256:3d465b204be271fc913963a8f90d1bbb4bdb06b4186382081c56eeb1df508298

Observation 8c955975-017c-4fac-8421-6c3a6b9b4deb · outbound

This paper cites Decoupled Weight Decay Regularization.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Decoupled Weight Decay Regularization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.083386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.083386Z digest=sha256:ad6bb52bb8dce1d187864775c51782f2a77e0aa507fad0ed59099a90d3000862

Observation f0b880c0-14f1-4c5a-b64d-f981166ec08f · outbound

This paper cites TurboRAG: Accelerating Retrieval-Augmented Generation with Precomputed KV Caches for Chunked Text.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache TurboRAG: Accelerating Retrieval-Augmented Generation with Precomputed KV Caches for Chunked Text

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.088826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.088826Z digest=sha256:8c16841065307ff4ed68328bd705f212cdd73135ea2bfbb90727038563e4b7bf

Observation 3746f130-4b87-4be0-ba08-5dffd1d23e01 · outbound

This paper cites The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache The RefinedWeb Dataset for Falcon LLM: Outperforming Curated Corpora with Web Data, and Web Data Only

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.094911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.094911Z digest=sha256:a73aed20a4b035ddfe6c4b288aa236ad672c9e3b276f3f4ccf80d15e9af28996

Observation 324d6300-15e7-4dc9-b3e5-809bc0ecab24 · outbound

This paper cites In-context retrieval-augmented language models.Transactions of the Association for Computational Linguistics, 11:1316–1331, 2023.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache In-context retrieval-augmented language models.Transactions of the Association for Computational Linguistics, 11:1316–1331, 2023

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.101743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.101743Z digest=sha256:d17fea4d893d49d441a8682b230281941e2e312361e8d2daee86c527edf36731

Observation d5dec647-4fe5-49bc-bfd3-86404247e1be · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.107024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.107024Z digest=sha256:7c03123931e020407b5761711838bf422fa06701210b9cd23a72f1d6ff2c5517

Observation 072be4c3-29c5-4267-911b-a2e42d984f47 · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.Advances in Neural Information Processing Systems, 36: 68539–68551, 2023.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Toolformer: Language models can teach themselves to use tools.Advances in Neural Information Processing Systems, 36: 68539–68551, 2023

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.113426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.113426Z digest=sha256:a0230f03528a936e0353fe87a7b0499ad0b84a9654b0b7296d1f8908263f1fc0

Observation a524f094-b4b2-4773-bec3-0d26fb58918c · outbound

This paper cites Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Masked Label Prediction: Unified Message Passing Model for Semi-Supervised Classification

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.120053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.120053Z digest=sha256:aac327cbd915e77eef99befdab417ff5455589c3712bcb3f7afed4258b282d56

Observation 4d7ada2f-4fac-49e2-9fec-5b8d2b9b9222 · outbound

This paper cites an unresolved cited work.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.129578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.129578Z digest=sha256:5aaeedd39caf22684864bb66f1115fdef1d1ad477d120a093f92a97ae10d1589

Observation 1b1238f8-b2de-4011-a948-f839c2133b9f · outbound

This paper cites Introducing mpt-7b: A new standard for open-source, commercially usable llms.DataBricks (May, 2023) www.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Introducing mpt-7b: A new standard for open-source, commercially usable llms.DataBricks (May, 2023) www

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.839268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.137053Z digest=sha256:439af9614103c9f7eea3646dc9b88bf038e38cfa763e1b6ecd2bd7c3c5e875ab

Observation 361a678b-7da0-4ce2-90ae-f793073dbac7 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.142994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.142994Z digest=sha256:ded808cd1c1da2a1dea84a1c2441a2c4aec95700b178c7f525dc060153987fd6

Observation 40985cbe-1468-4cae-9cd6-46c4d19b1e79 · outbound

This paper cites Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.148297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.148297Z digest=sha256:a9ece3cc0158e5ed5d5d84b8632e055621f45fb2210542453b49a353d9d659a2

Observation e1dad676-8c34-4e8b-a171-77861ec2b219 · outbound

This paper cites Graph Attention Networks.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Graph Attention Networks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.153919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.153919Z digest=sha256:1342de79369c9868a741916dfe2dc8a32106fd096c5fba7ec6389af936ecf31e

Observation 7532619a-7267-49ec-bea1-0b4900804a50 · outbound

This paper cites Can language models solve graph problems in natural language?Advances in Neural Informa- tion Processing Systems, 36:30840–30861, 2023.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Can language models solve graph problems in natural language?Advances in Neural Informa- tion Processing Systems, 36:30840–30861, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.160608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.160608Z digest=sha256:7c3d829b01e4d911978d45790e2e6ccf888f30c5bf8010f083fe9e14ac723c36

Observation fb6fef8d-706f-44ed-9f17-4cf2e70b0187 · outbound

This paper cites Cacheblend: Fast large language model serving for rag with cached knowledge fusion.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Cacheblend: Fast large language model serving for rag with cached knowledge fusion

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.810295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.166266Z digest=sha256:6c90072cd8447c06d8acfe14dfb3894f1edc2d78db783a5d02c29039d0d6555c

Observation ccf74f51-df73-4e96-afe9-7dde2a386974 · outbound

This paper cites Evaluation of retrieval-augmented generation: A survey.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Evaluation of retrieval-augmented generation: A survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.172226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.172226Z digest=sha256:7e4c0902df58d2225f80036688388b3b25b6e06fa4b0a184d7b069afce46a318

Observation 82e2e1ca-835d-451d-8941-485223cdca00 · outbound

This paper cites Prompting large language model for machine translation: A case study.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Prompting large language model for machine translation: A case study

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.776150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.177239Z digest=sha256:3ae01997e3a25f225f2e982af6b99597e5e1cd9bb9c3aa00af4a81154180d53c

Observation 175894d1-a3c8-49ed-9293-90eaa9b3894f · outbound

This paper cites Oag: Toward linking large-scale heterogeneous entity graphs.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Oag: Toward linking large-scale heterogeneous entity graphs

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.756033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.182406Z digest=sha256:9294867a258a9b5c074cc8425931593458c01f6289c714609f694e6bf06b474d

Observation c8970d8f-bdfa-4ce7-9488-dd37a6dccc07 · outbound

This paper cites An enhanced batch query architecture in real-time recommendation.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache An enhanced batch query architecture in real-time recommendation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.731988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.187686Z digest=sha256:3ea12930d9fe0761c6dd227fd14c61bfea767570139af4abea8687d9463987a5

Observation 731559f5-3c9f-4fc6-b789-a65fab3f0d83 · outbound

This paper cites Benchmarking large language models for news summarization.Transactions of the Association for Computational Linguistics, 12:39–57, 2024.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Benchmarking large language models for news summarization.Transactions of the Association for Computational Linguistics, 12:39–57, 2024

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.193110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.193110Z digest=sha256:8d9deed1143b4dec38e77d2146bc38f8c67df338b3e39192051d11befd763038

Observation b69bd7a1-c18f-40d8-94d7-b4fb0ccde35c · outbound

This paper cites Retrieval-Augmented Generation for AI-Generated Content: A Survey.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Retrieval-Augmented Generation for AI-Generated Content: A Survey

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:05:40.198793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:05:40.198793Z digest=sha256:75a94803919bf4c26b40cbc629031e75aabb2f9b01d339654f064c077d251633

Observation 30851db3-fe16-4a5a-a9e1-2078f1e24aae · outbound

This paper cites Efficiently programming large language models using sglang.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache Efficiently programming large language models using sglang

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:05:40.700569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.204566Z digest=sha256:98d7ffb7a3d36776c5c9b8227ecac8e47e168ff367da88987b5f27f842baf649

Observation d508513b-4cd6-44b7-824b-e3a6013445a3 · outbound

This paper cites HierPromptLM: A Pure PLM-based Framework for Representation Learning on Heterogeneous Text-rich Networks.

SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache HierPromptLM: A Pure PLM-based Framework for Representation Learning on Heterogeneous Text-rich Networks

Reference 46

Resolution
malformed identifier
local_arxiv, observed 2026-08-15T21:05:40.272767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T21:05:40.210126Z digest=sha256:7a953e5dae3e1c33e24bd6e92bef1433f7edad1f380127b60e1874b1a64f4993

Pith citing papers

Observation a4675863-8932-448e-929d-1c25b99f95d5 · inbound

Position: How can Graphs Help Large Language Models? cites this paper.

Position: How can Graphs Help Large Language Models? SubGCache: Accelerating Graph-based RAG with Subgraph-level KV Cache

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:10:42.890422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T18:48:03.257015Z digest=sha256:5f09c1cbd5b27a9008d5f5a17fd32ef39bbbb4bc67f66073947a48fc0a2a1efc