Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:50:01.430675Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 2 inbound Pith citation observations for arXiv:2412.02172.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T23:50:01.430675Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:39:59.412070Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T18:51:06.705524Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4b2c1e37-acc1-4e64-83c7-569df24b5b6e · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Tallyqa: Answering complex counting questions
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 914b30c6-9a1c-4547-a249-29f52d74a8fd · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning GPT-4 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a40dc8b7-f1ac-4d2e-bc7d-a1d1e9169a12 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Understanding the limits of vision language models through the lens of the binding problem
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5c967b40-98ce-423d-a258-835f3835bb1d · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mllm-as-a-judge: Assessing mul- timodal llm-as-a-judge with vision-language benchmark
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 41623b8f-6a04-4e28-857a-68ec08d740ba · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Teaching large language models to self-debug
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6a8abd41-55e6-4213-9206-a5b1721374f8 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb1b79ae-a6d3-4136-bb2a-4990f4963bd9 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7f00301-093d-465b-8ac1-4cd83b039fca · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Nvlm: Open frontier-class multimodal llms, 2024
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2af7b839-4a10-45a2-a7b1-874b4f5eba93 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df8e62ee-fcaa-48d3-8780-2f3c1434a1de · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e0471a3-f957-41c4-8605-6ca4578ed7fe · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e9ab90fe-def2-4bae-bee7-ad07affd5d94 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Detecting and preventing hallucinations in large vision language mod- els
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 16b19b72-7192-4672-bad6-c20b2cf6e47f · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff6816fc-c358-4491-9015-66beb0f00ec9 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 01378e97-0cd0-4abb-8aea-3f2b81159ed7 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Prometheus: Induc- ing fine-grained evaluation capability in language mod- els
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e3bd6d9b-b6cd-4f2f-914e-7e940c6556e9 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Criticeval: Evaluating large language model as critic, 2024
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f1311552-7d65-45d6-aed8-73b6c272972f · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Evaluating object hallucination in large vision-language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3ca17eba-71c1-463b-a7ac-68a90deb7b18 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e320964-10a1-4e97-bd25-4fc8e42274bd · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning CriticBench: Benchmarking LLMs for critique-correct reasoning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 74e9f7ee-5c98-482f-9844-10ec85db8349 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning A review of feedback models and theories: Descriptions, definitions, and conclusions
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e2d80857-d4ea-468f-970c-83f819e67b7e · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mitigating hallucination in large multi-modal models via robust instruction tun- ing
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 15578dbd-a851-4d6d-bfbc-08b1f7faac21 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual spatial reasoning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 287c2df2-1bdc-419e-ac8e-d25f345302d5 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual instruction tuning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 048c6073-7b70-4163-8e3d-71da8631ad0b · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning G-eval: Nlg evaluation using gpt-4 with better human alignment
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c529c18e-e9da-48dc-b009-f7231aa7fe71 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15e597ab-3e6e-413a-984c-a4165215ce81 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f72533be-0a1b-4e4a-9f70-3e6bf50dcd26 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mathvista: Evaluating mathematical reasoning of foundation models in visual con- texts
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8b41e188-7c91-424f-84c1-57385e539c05 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Critique ability of large lan- guage models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 295ed2a0-8791-4bca-a44a-7a93d0994266 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Self-refine: It- erative refinement with self-feedback
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d0ee6a8a-43c3-4cad-929f-38d5699f49ef · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d0db1b-4a48-4267-ad83-70b100bcb11c · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Docvqa: A dataset for vqa on document images
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a7e04bbe-4173-470d-b52b-934e79b78f98 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Teaching clip to count to ten
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 13451074-2a35-41d4-bb9b-ac59fd2721c8 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f64a04d-0c6f-471e-b7ab-6a52a5da28f9 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Reflexion: Language agents with verbal reinforcement learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 4d9bd026-2547-4cbf-a1bb-c340674fa09b · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Towards vqa models that can read
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2e5908c1-0acc-4ce4-81aa-c4328dae202e · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning The critique of critique
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5ace5606-4cd8-4ee6-9d25-f975a7f27fa8 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning V oyager: An open-ended embodied agent with large language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0adbef5c-b4ce-458d-b6ac-65f388403b1a · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Measuring multimodal mathemat- ical reasoning with math-vision dataset, 2024
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cabf7503-6a20-49cf-ae6d-a0cda7d4b7b4 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d30d8293-00ff-4c4c-b6e8-13062621e2a1 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 368581c2-16d2-4cac-a399-0af9728c28d4 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec563ee-04b7-467e-8554-70ffaec55c47 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning LLaVA-Critic: Learning to Evaluate Multimodal Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20efb07-160b-47ae-8c55-dc6ccb1b621b · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f890a9cc-27c9-4399-bace-76ea6ca2d0d9 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd676919-e4e8-43f3-b461-a1fae30b1154 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 797cbb9e-1926-4fff-b7c8-2b222324a8b4 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Attention Prompting on Image for Large Vision-Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02bcea21-f9af-4bcd-88bf-20ca402493a6 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6b7fee0c-10f1-4fc0-96a3-63ac76c253bf · outbound
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 0a48c339-e45c-40e3-ba31-97bfebcbaad5 · outbound
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning The final answer is: MODEL RESPONSE ANSWER - Critique: CRITIQUE FOR ANSWER Figure 28
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a43603fd-a2b5-49e8-a1eb-5e56b51e8161 · inbound
MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2430184c-629d-4eec-8bc3-f9736569496e · inbound
MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.