Pith. sign in

Paper Citation Record · LEDGER

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark

As of 7 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.15882.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.15882 v2

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:12:03.811310Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bdec5a75-e740-4419-9381-c9d9a27ecc92 · outbound

This paper cites Improving language understanding by generative pre-training.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Improving language understanding by generative pre-training

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:09.403705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:11:59.365713Z digest=sha256:ef379ddf341d822ffd5fba309e2e592387ba253356e0912b94b1ed492bd81062

Observation 15ff5d91-7beb-4982-a532-37f0e1536a26 · outbound

This paper cites Palm: Scaling language modeling with pathways.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Palm: Scaling language modeling with pathways

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.401910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.401910Z digest=sha256:ee9cbf1e39d776268cafdeaf150f04e5aab3dc2da0ea38935fd7588feb6a1798

Observation d7bf8b15-b176-45bb-9de9-0b3e549a0122 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark LLaMA: Open and Efficient Foundation Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.430437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.430437Z digest=sha256:f9be32a8d39bcca6ab6d66f10a48ed7193a73b63d2b1cd8e06713b18aa452cd9

Observation 5ffc3dc8-c6b8-4fe4-822d-8208fa21d627 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.455039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.455039Z digest=sha256:99b52027461b465ae28ddaa467ee64adc4e5dcb7de7edd980b6fe5a70cfa5634

Observation 055f1b8a-079e-4295-a6e0-373421acef82 · outbound

This paper cites GPT-4 Technical Report.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark GPT-4 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.493405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.493405Z digest=sha256:d3bc327edd62adcb233ec114b54ded6235df8154d6a12f6aed7368b2f92ba34c

Observation afb4f16d-a539-4702-bdc4-b88b33ff2f43 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Gemini: A Family of Highly Capable Multimodal Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.549630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.549630Z digest=sha256:47c9d8c48f59b97bde3f50bc1cefcc9d49b89763dacdb46ddf3ec775b206f9be

Observation 746a3555-eb67-43d1-bb57-19cc4bcab9f6 · outbound

This paper cites The Llama 3 Herd of Models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark The Llama 3 Herd of Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:11:59.677645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:11:59.677645Z digest=sha256:862a0b374e88ca8041ab36dc48ca503f892bc540891a6dce59cd2cb2fe78a338

Observation dd1fa778-bdde-4df4-b501-cc0a34073135 · outbound

This paper cites Introducing the next generation of claude: Claude 3 model family.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Introducing the next generation of claude: Claude 3 model family

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:09.150309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:11:59.812322Z digest=sha256:7f471f4438816509d24c3032a967bb06f8dc3b330433625dc2bcff0d43297ed2

Observation 6029eb52-ba7e-4834-b8a9-72a7cc99e65f · outbound

This paper cites The amazon nova family of models: Technical report and model card.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark The amazon nova family of models: Technical report and model card

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.923684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:11:59.945149Z digest=sha256:9342bf682b4b3d69b86d620e2ce2e265cc3cc51b6f439447b813b5397ae00cbc

Observation f74e43bc-4295-4db0-a940-1a8ed08f3396 · outbound

This paper cites Visionllm: Large language model is also an open-ended decoder for vision-centric tasks.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Visionllm: Large language model is also an open-ended decoder for vision-centric tasks

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.718134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.043620Z digest=sha256:b422cf32bc8fe33aed08a2078283df730e27391918b375007ef964309ead741a

Observation 48e270f1-ef79-48ee-9c86-fb54bdc34c40 · outbound

This paper cites Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:00.161768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:00.161768Z digest=sha256:6ad3fecb74d5c232364a1c178fccedc0482a74628b7048be0eb0570b63c71722

Observation f70a710e-72fa-49f5-9473-03ad86e397d1 · outbound

This paper cites Zero-resource speech translation and recognition with llms.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Zero-resource speech translation and recognition with llms

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.478594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.359797Z digest=sha256:c40376e5a055b4534ab6730d5c8ec8b0750c6d1147181289169c340eecbf0f5a

Observation 53e7e97b-432f-487e-9f15-55cc23f44d33 · outbound

This paper cites Legal- bench: A collaboratively built benchmark for measuring legal reasoning in large language models.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Legal- bench: A collaboratively built benchmark for measuring legal reasoning in large language models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:08.206903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.538822Z digest=sha256:20c7c13172289981a63b90e3736301e55a6ccf2820eb4fbe1682348519c2962a

Observation 691f8977-ccee-4b33-91cf-ed45f729094a · outbound

This paper cites Chatlaw: Open-source legal large language model with integrated external knowledge bases.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Chatlaw: Open-source legal large language model with integrated external knowledge bases

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.973577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.658961Z digest=sha256:fb015b1543ab765507e3f2941bd473b348cd42af57eb1d9e01998d0a6dd54164

Observation cfd8f6f6-3ff4-4515-a935-f83a85ea8b94 · outbound

This paper cites A study of generative large language model for medical research and healthcare.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark A study of generative large language model for medical research and healthcare

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.652417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.724539Z digest=sha256:ebc418adbadd08253b4158c70fc18f7c33650e5ec99f6729bb7ae0b9c15b6f55

Observation 5709f843-c7c6-4453-b980-c1b77cd516df · outbound

This paper cites Large language models in medicine.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Large language models in medicine

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.486410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.817655Z digest=sha256:66886f8a7aabe7dbf62a1e4cd1462c0775a47396534643d9bad8d8ed00d0993b

Observation 2d67e545-e78e-4c93-93af-b62101901c99 · outbound

This paper cites Large language models in finance: A survey.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Large language models in finance: A survey

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:07.265020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:00.931170Z digest=sha256:68de03e6ebbcd740f49baaee15d8999aee29603f9dde9bfc9edf195ae1b0ed30

Observation 01eccb1b-fd52-43fb-a708-227884a5c295 · outbound

This paper cites A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark A Survey on Large Language Models for Critical Societal Domains: Finance, Healthcare, and Law

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:01.088100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:01.088100Z digest=sha256:d20f3b828aee0c427f7bff0c7ed2b10b288fc04a8c44be5337f006402c323f9d

Observation be36be46-14c7-442b-9148-bdf8a426487c · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:01.261807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:01.261807Z digest=sha256:625a7f004f9364357386e1ac1cf86b613932b01b8912541ce18884dacb6d5e82

Observation dadc202f-5e84-45c4-a3ff-3e8eceeda7ff · outbound

This paper cites Superglue: A stickier benchmark for general- purpose language understanding systems.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Superglue: A stickier benchmark for general- purpose language understanding systems

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.917425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:01.394095Z digest=sha256:831a5a9242681c1fb671c85b00ea96b42ad4beda6e10eb21aaaba1efa27cb437

Observation 05a1a7e5-78ad-45cf-b5b1-2b7ce80cbff0 · outbound

This paper cites Needle in a haystack-pressure testing llms.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Needle in a haystack-pressure testing llms

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.668443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:01.566528Z digest=sha256:48c3ba5ae27c1a00507e6a9abf871d2b79021081aab802d57cdc0c2f29d12756

Observation b703fee8-e780-40c9-a86a-53f5c6c9b6a9 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:01.712774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:01.712774Z digest=sha256:ff393a33f62307e0450e1f188495491b853c968bde478fa4ca82d5ac6f314e6b

Observation e8e3a648-420a-4b26-b5f5-3e32071c7ed3 · outbound

This paper cites Vqa: Visual question answering.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Vqa: Visual question answering

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.382209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:01.876492Z digest=sha256:7c54c8e51b2114157ee0a7f773a14aa2c822a1513ae589c363e2b57563780c58

Observation f4e19d24-0934-444b-8fab-2888b4389a5b · outbound

This paper cites A Corpus for Reasoning About Natural Language Grounded in Photographs.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark A Corpus for Reasoning About Natural Language Grounded in Photographs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.036747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.036747Z digest=sha256:8110e009d23997fc478b7d98da95ca543b8d991f9b1f3e4b6661d6c7409a5481

Observation 2c6eae63-2f9c-449f-b856-6080ae010be5 · outbound

This paper cites MileBench: Benchmarking MLLMs in Long Context.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark MileBench: Benchmarking MLLMs in Long Context

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.171862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.171862Z digest=sha256:c0c4dd221d3cc5a037799764b2cab256e4c537b4739c47d2730c0c332e1a09a1

Observation 1e246951-2bed-4f69-9193-0ad1ef4dd56f · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.326329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.326329Z digest=sha256:b5acc47770d25d00109e8c8a0e83c8ad44d422ab001b96d302b30435b3dc67ca

Observation 76c5ce3d-5e6d-464e-a5d8-2572687aa250 · outbound

This paper cites Docvqa: A dataset for vqa on document images.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Docvqa: A dataset for vqa on document images

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:06.071480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:02.479204Z digest=sha256:9aca9f1f818b7f298a0ce374367dff3b2ba81c83ec026f9a4480ed22a02cf562

Observation 91332362-0274-45a4-a58f-852ee01988dd · outbound

This paper cites Document understanding dataset and evalua- tion (dude).

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Document understanding dataset and evalua- tion (dude)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:05.884146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:02.641880Z digest=sha256:8e48d00c43a9e46621c21bb3f7c560dc5d24ad1457fbdc65cabe386f1edfed3b

Observation eaa94a9d-bd12-43eb-bc08-6fa51c8b11c3 · outbound

This paper cites Leave no document behind: Benchmarking long-context llms with extended multi-doc qa.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Leave no document behind: Benchmarking long-context llms with extended multi-doc qa

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:05.594366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:02.766229Z digest=sha256:cbdfb5f3251c017863c72fb580d2a076d69876f09f6318ec0446ca552349ed0d

Observation e8b89465-2f2f-4727-9996-0e454d557092 · outbound

This paper cites Slidevqa: A dataset for document visual question answering on multiple images.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Slidevqa: A dataset for document visual question answering on multiple images

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:05.270470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:02.916848Z digest=sha256:bce4a9d406bfc87d586bf97a13b3e572260619f5ba51c9ec9f36467434557bb7

Observation d4acb8a5-5d1b-43b1-94a9-f204c153c49a · outbound

This paper cites MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.996244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.996244Z digest=sha256:cfb62e2cbe33c044a7fa0f771a06a805c3c57140661a708504309c243bb5e694

Observation 527e72e5-0a7d-4e03-9788-624ac5fab665 · outbound

This paper cites Needle In A Multimodal Haystack.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Needle In A Multimodal Haystack

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:03.116125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:03.116125Z digest=sha256:1b254a11a0274689c9616325ccf9bdcab81379f68a622e4a1776c4a241397989

Observation 40d0d1e0-ba4b-4b0a-abe3-874870c8707a · outbound

This paper cites M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:03.272897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:03.272897Z digest=sha256:e42d3890374fb07199c5ad4f4b0a54ca9ec592e97481f3f6d401c6db4b079597

Observation c64c4e5b-afd7-4f9b-a92e-1e04eed10b3e · outbound

This paper cites Pixtral 12B.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Pixtral 12B

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:03.418422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:03.418422Z digest=sha256:9361a5cfcbfa472144a01be382a8d805ea79287ee23399d7aa85cfc7d0a22136

Observation 07604ab5-94a6-416c-8a5b-08e471b4db48 · outbound

This paper cites Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:12:04.176645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:03.638856Z digest=sha256:27f06f9cd82b661ba3cf08549b560aa8cc9b9f641870b8a2a3f06d9055c7374f

Observation 8d68d363-5b69-4a39-86db-2c1b14cf346d · outbound

This paper cites Lost in the middle: How language models use long contexts.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark Lost in the middle: How language models use long contexts

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:12:04.935596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T16:12:03.811310Z digest=sha256:9df1ce14897249b269b0bd83b2b5f86c8cf85e1ab2e8d2e5a9f84648b2953940

Pith citing papers

No inbound Pith citation observations are available.