Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:34:38.340134Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2506.21316.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:34:38.340134Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 08222e7e-2dc9-4fb1-a287-4bf071b47021 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Manmatha
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 818156ef-effa-486f-b649-57812d72f569 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Paddleocr, awesome multilingual ocr toolkits based on paddlepaddle
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 734429a8-cdfb-4d64-a25d-c84fc13fdbfa · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Bharatgen unveils patram: India’s pioneer- ing vision-language foundation model for document intelli- gence, 2025
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41ddf928-70c8-4b8e-a667-2cb316beb88d · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Molmo and pixmo: Open weights and open data for state- of-the-art vision-language models, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b842fa82-5fc4-4a65-8a4d-9ea92d83c69f · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents The Llama 3 Herd of Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5c9bf1e-a4e8-4fb4-b9f3-a137212ed0c2 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents mPLUG-DocOwl2: High-resolution Compressing for OCR-free Multi-page Document Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9551178-2d6b-4390-8a77-4060d5b445cb · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Layoutlmv3: Pre-training for document ai with uni- fied text and image masking
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ff6eef45-1be0-4eb1-a914-190b55fc339d · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a3633cfb-e759-4e92-b1b0-df3cb5bc3d99 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Towards visual text grounding of multimodal large language model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f288daed-560b-4aaa-8e87-ae9c25f7473c · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Layoutllm: Layout instruction tun- ing with large language models for document understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d8a63d0-4dfb-4ce0-9cfa-fa3d159a3e8e · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Manmatha, and C
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b8237044-855f-4bfa-9c99-a5fc9f649726 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents doctr: Document text recognition
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd85ff38-1840-4c48-99a1-e8d27e393c66 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Nagaraja, Vlad I
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7293f1f5-efa8-4bdb-b161-831c66a9849a · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Surya: A lightweight document ocr and analysis toolkit
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3ed71f0-a2e2-4a52-9176-52848ac7601f · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Going Full-TILT Boogie on Document Understanding with Text-Image-Layout Transformer
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03772cf5-7c6c-4ee2-879c-a242821c1315 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Grounding of textual phrases in images by reconstruction
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6297c7f0-0d49-47ab-ad74-2602afa5bd94 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Unifying vision, text, and layout for universal document processing
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 182a84dd-83fc-4dbc-8716-a71a0b4d09d8 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Gpt-4 technical report
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 887b9de5-1192-48e0-ac6c-e7577ae81940 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Hierarchical multimodal transformers for Multi-Page DocVQA
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53c3c62c-21e8-45b2-908f-e25224daeeaa · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Docllm: A layout-aware gener- ative language model for multimodal document understand- ing
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69761ad4-aa54-4e5e-a872-fc183c7c1275 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1703ad8c-9734-4663-ade3-b85c758fef5a · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Towards Improving Document Understanding: An Exploration on Text-Grounding via MLLMs
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e77bf65-a5ce-412f-873c-d7c3da2d502a · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents Toward Visual Grounding: A Survey
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83781985-7053-40ae-a698-8e9887e53953 · outbound
DRISHTIKON: Visual Grounding at Multiple Granularities in Documents DOGR: Towards Versatile Visual Document Grounding and Referring
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.