Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-12T22:34:11.603463Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2604.10528.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-12T22:34:11.603463Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5b645d1a-c999-44bf-822b-432385986a75 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Pixtral 12B
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e94dedf-4a22-4964-89e5-d00accf5e328 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Ovis2.5: Structural embedding alignment for multimodal large language model, 2025
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13114dde-0758-4625-a9ed-9d070867ac1e · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Claude 4.5 model card, 2025
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f485e2b-12d1-4ba5-bb58-3637d83fb8c3 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4c820be-b9f2-42b3-a7b5-3f95dfa0d930 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Re:Verse - Can Your VLM Read a Manga? InProceedings of the IEEE/CVF Inter- national Conference on Computer Vision, pages 3761–3771,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4648a73-0c01-4bd0-b476-3fd020ba58a4 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Florence-vl: Enhancing vision-language models with generative vision encoder and depth-breadth fusion, 2024
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22ea92c8-83e6-4c7b-9fe8-8949a65996a4 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs The pascal visual object classes (voc) challenge.IJCV, 2010
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e01b4e4-9dc9-404c-a990-c82b9280962e · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Large-scale unsupervised semantic seg- mentation.IEEE TPAMI, 2022
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9c50db6-5c30-4e78-9e41-f4bc5ef33af9 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Are vision language models texture or shape biased and can we steer them? InMMFM Workshop @ CVPR, 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e750d369-54cd-49c9-b5c2-b7961b001680 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Imagenet-trained cnns are biased to- wards texture; increasing shape bias improves accuracy and robustness
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbf4aa75-f6f3-4431-9a27-f1b5f062ce50 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e1bb70-3f0d-4960-847c-b999349da81d · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Paligemma: A versatile 3b vlm for transfer, 2024
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37f57ba2-bc06-40d1-bf59-235f01e1438c · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs The origins and prevalence of tex- ture bias in convolutional neural networks
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37abc0a1-da38-428a-acae-d88be9cb5122 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Scaling up visual and vision-language rep- resentation learning with noisy text supervision
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f7295d2-0886-44eb-99ec-2c8fe24ca07a · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs LLaVA-OneVision: Easy Visual Task Transfer
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54b9f056-af28-4e13-877e-c38e5da08bd0 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Seed-bench: Benchmarking multimodal llms with generative comprehension
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0983e2ea-2123-4f9f-8a57-de8c9e4a96f2 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Deep learning for thin object segmenta- tion
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cfa1c32-2f7e-4734-a723-77fc2ba52f66 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Visual instruction tuning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b462110a-a552-4d00-abd8-ef223407bdb6 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f85996-6221-417d-9d1c-52de878c9b2a · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs SmolVLM: Redefining small and efficient multimodal models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11cd585e-75da-4f05-936a-57e66565fb4d · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Phi-3 vision 128k instruct, 2024
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c078542f-b371-4b1a-87a1-b309e091b85e · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Verification Learning: Make Unsupervised Neuro-Symbolic System Feasible
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1d4d114-b130-4118-b605-31ac9495f4aa · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Internvl2.5 pretrained models, 2024
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80735fb8-fbac-4506-a77b-a67f5672a02a · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Robust onion: Peeling Open V ocab Object Detectors Under Noise
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e61e22ab-d32d-4ad5-bf9f-a8be7c8c0456 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Highly accurate dichotomous image seg- mentation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2fd4b0d-4233-4954-824c-007aead7330b · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Qwen2.5-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e95c2ff0-0cea-424d-8ac3-90eeace477d7 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Learning transferable visual models from natural language supervision
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5babb2ba-bd7b-4109-a9a8-5cf317b53757 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Eyes wide shut? exploring the visual shortcomings of multimodal llms
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62df8977-9da1-467e-8634-15088da1f8f4 · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs The caltech-ucsd birds-200-2011 dataset
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bce14a2-83ae-4496-8e73-f2e4c25667bb · outbound
BareBones: Benchmarking Zero-Shot Geometric Comprehension in VLMs Who’s That Pok´emon?
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.