Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T10:38:13.638865Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 2 inbound Pith citation observations for arXiv:2412.16420.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T10:38:13.638865Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:09:13.307268Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T15:48:52.562583Z
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ad7f72ba-a4e4-402d-963a-1a5b83cae79d · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Llama 3.2: Vision at the edge with mobile devices
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 368ab642-513a-48f9-bb3a-cbf5c55546f7 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Introducing claude 3.5 sonnet, 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d49b3753-18a3-41d5-b6b8-654fbb811489 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Disentangling Knowledge-based and Visual Reasoning by Question Decomposition in KB-VQA
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d6661da2-246c-42f8-99c9-3ce0323167fa · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Modularized zero-shot VQA with pre-trained models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4277e715-9ce3-452a-a7d9-76f440b5a08b · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding The Llama 3 Herd of Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3a49f58-010b-4fb4-b064-985d72e79c62 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Exploring question decomposition for zero-shot vqa
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 81ed77fe-6323-4081-b5d6-73f8fd9526b3 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Llava-next: Improved reasoning, ocr, and world knowledge, January 2024
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 073f6a9d-1df5-41ea-9497-72e3531543f2 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Iconqa: A new benchmark for abstract diagram understanding and visual language reasoning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fc5d96f9-82bb-445d-9f75-c081e179efc9 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Hello gpt-4o, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation a76b5c08-c337-4da2-ac1b-e6b8b9454b59 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding FlowLearn: Evaluating Large Vision-Language Models on Flowchart Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00c1fba5-9401-4404-87b8-82a9cad67a99 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding F low VQA : Mapping multimodal logic in visual question answering with flowcharts
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae5b42dd-aef2-4c36-b9a6-ee1d4ea67ba7 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Feighelstein, Jasmina Bogojeska, Joseph Shtok, Assaf Arbelle, Peter W
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 5e6a8077-0786-4f75-be2a-47188452bfbd · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Mixtral 8x22b: Cheaper, better, faster, stronger, 2024 a
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation d305a17b-54e1-40fb-80e4-d8175ffdf348 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Qwen2.5: A party of foundation models, September 2024 b
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a187496-ae91-41a6-b584-7d68cb25a39b · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Discover the new multi-lingual, high-quality phi-3.5 slms, 2024
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0618911a-367a-4201-af64-81e52c18f49e · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ef79cf3-4abd-4921-9ae1-1c4160a53d39 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Ayyubi, Kai-Wei Chang, and Shih-Fu Chang
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6b0d8317-1199-4acc-a88d-4391c16bc1ac · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding write newline
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba12dd5b-5a3a-46b5-8ad4-9f57605b181e · outbound
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b5ec85b-01e4-47b7-a97c-07ab6d2ca185 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 175cdee9-cb93-4670-b43c-429b2f305391 · outbound
Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding blue The master students are still working on the annotation and will get the results for those empty entries by today
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7206a5bf-b538-4c66-aee6-8b6a5350f007 · inbound
Overcoming Vision Language Model Challenges in Diagram Understanding: A Proof-of-Concept with XML-Driven Large Language Models Solutions Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0320b4fa-3aa1-4bd2-b947-5371c8880304 · inbound
Survey of GenAI for Automotive Software Development: From Requirements to Executable Code Beyond End-to-End VLMs: Leveraging Intermediate Text Representations for Superior Flowchart Understanding
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.