Pith. sign in

Paper Citation Record · LEDGER

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents

As of 7 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2607.24748.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.24748 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T13:26:31.987399Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 13ca8334-0769-46b6-bae4-6a11d6ec6ba6 · outbound

This paper cites - Start with meta search using document titles and metadata for fast candidate selection.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Start with meta search using document titles and metadata for fast candidate selection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.449554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.449554Z digest=sha256:b1b1db804d8fc2c17ddb99139ded41926485d84c8a5fc669c30c38130713b7dd

Observation cd5af9bd-2028-41e2-a89c-c26eeef662ca · outbound

This paper cites - Perform semantic search using HyDE embeddings for conceptual similarity.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Perform semantic search using HyDE embeddings for conceptual similarity

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.606217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.606217Z digest=sha256:5d72e27573c52954daf52b7cfea15841a15e683466286109d5b17317944ae799

Observation 29cb9329-9009-4bca-847a-5ee70afa8bad · outbound

This paper cites Hu, A., Xu, H., Zhang, L., Ye, J., Yan, M., Zhang, J., Jin, Q., Huang, F., and Zhou, J.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Hu, A., Xu, H., Zhang, L., Ye, J., Yan, M., Zhang, J., Jin, Q., Huang, F., and Zhou, J

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.676884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.676884Z digest=sha256:93846cbb171602d26be46d593638d998e7dcdf6760db7c0ba4a6c3e3fb73d6f9

Observation 8ea05033-810f-4109-8777-0960d41aebdd · outbound

This paper cites MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents MRAG-Bench: Vision-Centric Evaluation for Retrieval-Augmented Multimodal Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.810681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.810681Z digest=sha256:65f637bf5217ba40ea7c18703a243ba52b703bc1aaaea878735a3ea2499c12e1

Observation d4a5edc3-07b1-4e7d-8d4c-eefd0038f00a · outbound

This paper cites - Include document metadata and chunk positions for traceability.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Include document metadata and chunk positions for traceability

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.028796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.028796Z digest=sha256:04dd7d0553cf7fe1f89bc0c8a335d041eb5a100013f73f67156bee520c266f98

Observation fe02b393-dc29-4800-9b77-4e56c2c67881 · outbound

This paper cites Shi, Y ., Wang, J., Shan, Z., Peng, D., Lin, Z., and Jin, L.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Shi, Y ., Wang, J., Shan, Z., Peng, D., Lin, Z., and Jin, L

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.159099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.159099Z digest=sha256:b4e6c3aeef49d048bf0a9820382364ce86b6fa692c4a31df44bfc45b00295279

Observation f9d19562-1b51-4950-8752-e07749f432ce · outbound

This paper cites - Generate evidence from the final chunks with proper source attribution.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Generate evidence from the final chunks with proper source attribution

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.789058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.789058Z digest=sha256:41ee71a6afb7864cfccbad53d57d3a580a27386c84145ee138ed590a5e5a27f8

Observation d7cd0f3d-5b61-4276-a5eb-2163ad1a64d3 · outbound

This paper cites - Do not repeat the same failed search parameters.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Do not repeat the same failed search parameters

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.937442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.937442Z digest=sha256:b76f426b2355dc192e257746f0c80b1dec2d3703dcc9c3ace01e211d5f089a3d

Observation e4c3dd45-edc8-4a4a-b65a-b2bdc5af6f9e · outbound

This paper cites - Identify restrictive qualifiers in the original query.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Identify restrictive qualifiers in the original query

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.192196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.192196Z digest=sha256:fecabb3862174256c29d5557c93da20af3fbc6369eff9c7ef2a7721ca287eb05

Observation f9d1d3b8-7012-4545-8b20-2b108d8301c5 · outbound

This paper cites - Remove numerical constraints if not essential.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Remove numerical constraints if not essential

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.350618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.350618Z digest=sha256:95ce1b8af57d4940cc0fd47b766c6197656450c65e85bfb4c483ace92e4c5c03

Observation ded97a5b-c8b6-48eb-a6b7-62b5b4c2b584 · outbound

This paper cites Generate a HyDE (Hypothetical Document Embedding) query for semantic search.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Generate a HyDE (Hypothetical Document Embedding) query for semantic search

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.446244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.446244Z digest=sha256:95d8056812fd6b5890495eb42bec815d1a85cefec2446f2f3eb8ce48016d2358

Observation 9e490cef-901d-487c-99a9-44839ced6d46 · outbound

This paper cites - Generate 1-3 sentences, approximately 100-250 tokens.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Generate 1-3 sentences, approximately 100-250 tokens

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.602117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.602117Z digest=sha256:5ced9e0b10c079e19eb0150a32041bf6f4c18c641b37f9a375cc4615e84280b0

Observation 74003f01-ffff-4efc-8131-fa9a564f10f8 · outbound

This paper cites - Describe the type of document that would contain the answer.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Describe the type of document that would contain the answer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.743963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.743963Z digest=sha256:51200e4f381836c280068ab9ea20bed865d1659a77eda12de5bdacfc25d6b539

Observation 8119981a-6a10-4934-819a-217060ac9883 · outbound

This paper cites - Include related concepts and context that might appear in relevant documents.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Include related concepts and context that might appear in relevant documents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.868542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.868542Z digest=sha256:bd2990cb865bcbc3f19be34b4d4d0118f7bacc02eae6251533431ac8131910de

Observation 3741cb58-0fa7-4d35-ac0b-9d27198684c0 · outbound

This paper cites - Do NOT use question format.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents - Do NOT use question format

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:31.987399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:31.987399Z digest=sha256:0048fe738cf8f54611326c45df3e43c65018c8dd4021911f86fd3b01a97c2d58

Observation 5ad0e590-b36d-4834-b898-3d7f258b82e3 · outbound

This paper cites emnlp-main.311/.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents emnlp-main.311/

Reference 311

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.000582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.000582Z digest=sha256:af76ba09b3dbc1b224919ed81f7b4c39bf77ee7bc239c3a5430487352162dca6

Observation 90e48d1b-acc6-4ed1-8d1d-13fa881ad2bb · outbound

This paper cites Retrieval-Augmented Generation with Graphs (GraphRAG).

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents Retrieval-Augmented Generation with Graphs (GraphRAG)

Reference 735

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.578873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.578873Z digest=sha256:6b1e19917fb355cfaabd0c5e4d9264d0f2c9e9c218b86981b091ba5c555ee347

Observation c5b73bbf-e22c-4165-9cb5-f1312655973b · outbound

This paper cites VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents VisRAG2.0: Mitigating Visual Hallucinations via Evidence-Guided Multi-Image Reasoning in Visual Retrieval-Augmented Generation

Reference 1166

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:30.298761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:30.298761Z digest=sha256:92bcedee72904a81b3a861aceb6ae33e1e81bfec78c40ffb66bc6ca2d76aaebc

Observation 10edcf7f-3bcc-42de-9138-811916a638e0 · outbound

This paper cites ISBN 979-8-89176-251-0.

VLD-RAG: Agentic Vision-Language Retrieval-Augmented Generation for Long, Visually-Rich Multi-Page Documents ISBN 979-8-89176-251-0

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T13:26:29.528679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:26:29.528679Z digest=sha256:fb9b40d360030e4ccdb0f47a9d62a7694ab995dc422860c936f1a1ca12b7b497

Pith citing papers

No inbound Pith citation observations are available.