Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:38:24.547522Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2501.15144.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T14:38:24.547522Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b9600f08-70cc-46cc-8d16-9e5bd6b977ae · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a930986-248c-450f-9d8e-60d6af521c45 · outbound
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fe7bcbf2-fa40-4db2-a84a-a1d8bbfed217 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 940a74c5-8ba4-48df-a740-0a32275092f2 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models PaliGemma: A versatile 3B VLM for transfer
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10576d6b-f8c2-49ee-b744-fed08389c296 · outbound
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3c1a8817-28ba-4336-8eb7-c38a2dcd6382 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4dceec40-7eab-41d6-9659-523db1b678d9 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models The Llama 3 Herd of Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 238f5e4a-4b42-4556-a9d0-95e355648d00 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6b657000-0ba8-412d-aa29-55ab72f5b5b0 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bec82237-c7d1-4272-81e2-49e398be6a08 · outbound
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0977a93c-da7a-4db5-8d65-abb39d5a0576 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5d547e25-110d-46d1-9a00-7089337b122f · outbound
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b1114157-1104-4bb7-a156-a3ff464717d8 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Prism: A Framework for Decoupling and Assessing the Capabilities of VLMs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d6e0c9c3-502f-4b1d-8630-292561fcc54d · outbound
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4492f9b0-c03f-4c34-9738-dd928133df44 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Vision language models are blind: Failing to translate detailed visual features into words
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 235333c6-971a-48ab-9389-ae5ccdd6eb78 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c63b0f55-e4ae-4d82-a09a-99c4ce743a26 · outbound
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 96a05e5f-1e9b-4574-be11-c733b22f262a · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6eef9f09-924d-460d-8b05-be5df45f10d9 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1cdf967-880c-4933-8791-c87a16da6ac7 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Guiding Vision-Language Model Selection for Visual Question-Answering Across Tasks, Domains, and Knowledge Types
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d08b5476-9cec-49ff-883a-52c4d94ca8a0 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cb8647d-f12e-44d0-bd35-61754c089210 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Qwen2 Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a700a717-608f-4fbb-b18d-4bd2b1c7d44b · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eea5803-ea97-434c-ace7-79e3569581ad · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Specifically, the shape limit is increased to 5–6 shapes per image, and the occlusion limit is also raised to 5–6 shapes
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 06aa8683-e30e-4912-a802-ed735ad8f655 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models This setup tests the model’s ability to detect and attribute more shapes in configurations that were not present in the training data
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 628f8bf5-7b8d-4947-b61f-b63869b57ebf · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models This scenario assesses the model’s ability to accurately detect and attribute shapes under previously unseen levels of occlusion
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c0a8e2e7-c228-495c-b296-9ff3fdc85bc8 · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models The ability to generalize to these out-of-domain rotations Model OD Comp
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1387a2fc-f641-4791-98f9-ac2311053ecc · outbound
Exploring Primitive Visual Measurement Understanding and the Role of Output Format in Learning in Vision-Language Models Sentence Format Tuple Format MiniCPM-V2.5 Figure S4. Sentence Vs Tuple Output Comparison for MiniCPM-V2.5 Models for Validation Dataset
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
No inbound Pith citation observations are available.