Pith. sign in

Paper Citation Record · LEDGER

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning

As of 22 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 2 inbound Pith citation observations for arXiv:2412.02172.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02172 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:50:01.430675Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:39:59.412070Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T18:51:06.705524Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved19
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4b2c1e37-acc1-4e64-83c7-569df24b5b6e · outbound

This paper cites Tallyqa: Answering complex counting questions.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Tallyqa: Answering complex counting questions

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.195529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.097880Z digest=sha256:380a85930ee77072dd6e6b29ed8001eb8b94510da85e132b861e6b98413b40be

Observation 914b30c6-9a1c-4547-a249-29f52d74a8fd · outbound

This paper cites GPT-4 Technical Report.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.102790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.102790Z digest=sha256:b98d67041e484e1dea9f10583c5585b2e96428acb861a763eab4af6144ed8128

Observation a40dc8b7-f1ac-4d2e-bc7d-a1d1e9169a12 · outbound

This paper cites Understanding the limits of vision language models through the lens of the binding problem.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Understanding the limits of vision language models through the lens of the binding problem

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.180255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.107625Z digest=sha256:c6329e2f7c0224a9c0ea7d302c031360f9fed8d398250044c2ec6f6f942a7283

Observation 5c967b40-98ce-423d-a258-835f3835bb1d · outbound

This paper cites Mllm-as-a-judge: Assessing mul- timodal llm-as-a-judge with vision-language benchmark.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mllm-as-a-judge: Assessing mul- timodal llm-as-a-judge with vision-language benchmark

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.166529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.111951Z digest=sha256:a1686d7d4cd677142115e2e69fa221587c545f3326489f2f4bddfa5a32c8ad7a

Observation 41623b8f-6a04-4e28-857a-68ec08d740ba · outbound

This paper cites Teaching large language models to self-debug.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Teaching large language models to self-debug

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.152875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.225476Z digest=sha256:3d84759d2672cb983a9bf03ff538ee5717df826c53abb2783defd94ed36a70c8

Observation 6a8abd41-55e6-4213-9206-a5b1721374f8 · outbound

This paper cites Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Measuring and Improving Chain-of-Thought Reasoning in Vision-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.230550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.230550Z digest=sha256:3abd0dc2a76c07051f7a607f28c60740f0c88d4498c4dd72a286d1c7e84c88c8

Observation eb1b79ae-a6d3-4136-bb2a-4990f4963bd9 · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.235860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.235860Z digest=sha256:90ae6254e371261420ab488946d59d07654bcc60b84a249942bfff020189bb9a

Observation f7f00301-093d-465b-8ac1-4cd83b039fca · outbound

This paper cites Nvlm: Open frontier-class multimodal llms, 2024.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Nvlm: Open frontier-class multimodal llms, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.138372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.240481Z digest=sha256:a93eb116bdea4cda750d0a7d7e18bdc5dd434b9a9370a6aafa3bf983efa3361c

Observation 2af7b839-4a10-45a2-a7b1-874b4f5eba93 · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.244758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.244758Z digest=sha256:442fe2f81e8764b9af100a5b21e083667794d84c77fbcb7c21fb3210fe2edc4f

Observation df8e62ee-fcaa-48d3-8780-2f3c1434a1de · outbound

This paper cites EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.249846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.249846Z digest=sha256:db1af1bff680f6241ee40062c665a6157a440930520ac0a27d2aeac4e10e1ef0

Observation 0e0471a3-f957-41c4-8605-6ca4578ed7fe · outbound

This paper cites Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.123924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.254670Z digest=sha256:05fabbbeff2c5be7c27a7021feac753ad4515d1d0071e7a292540528ff9b08ba

Observation e9ab90fe-def2-4bae-bee7-ad07affd5d94 · outbound

This paper cites Detecting and preventing hallucinations in large vision language mod- els.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Detecting and preventing hallucinations in large vision language mod- els

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.108688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.259117Z digest=sha256:cb15fa5e0c6b0b0c9f43d57a0eec6ad53fda488c2dd7c51194f18dd9fac1febe

Observation 16b19b72-7192-4672-bad6-c20b2cf6e47f · outbound

This paper cites Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.263191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.263191Z digest=sha256:a3aa9e36744340a224f323a8c9b2466f0a29ffc7ece14be6c4c7b98cd3339a86

Observation ff6816fc-c358-4491-9015-66beb0f00ec9 · outbound

This paper cites What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning What’s “up” with vision-language models? investigating their strug- gle with spatial reasoning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.092753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.267690Z digest=sha256:5b440144adadf91c3eedd5a1485e092ffb3e887f25fd82a61ca9c946849634ac

Observation 01378e97-0cd0-4abb-8aea-3f2b81159ed7 · outbound

This paper cites Prometheus: Induc- ing fine-grained evaluation capability in language mod- els.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Prometheus: Induc- ing fine-grained evaluation capability in language mod- els

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.077701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.272137Z digest=sha256:06d07a4b2425c3e013095ee8247c0a3eade60df25582e85d803ebcfd14477a3d

Observation e3bd6d9b-b6cd-4f2f-914e-7e940c6556e9 · outbound

This paper cites Criticeval: Evaluating large language model as critic, 2024.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Criticeval: Evaluating large language model as critic, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.062407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.276326Z digest=sha256:8cfaeaea96382384a9244c5a62a1ecd7096bcd9b346b89623adfa8d04cc9771f

Observation f1311552-7d65-45d6-aed8-73b6c272972f · outbound

This paper cites Evaluating object hallucination in large vision-language models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Evaluating object hallucination in large vision-language models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.047452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.280487Z digest=sha256:1fd6d3bcbec7d52a34d9c71a1e3909225786aa7e703ae308076a44dead603eda

Observation 3ca17eba-71c1-463b-a7ac-68a90deb7b18 · outbound

This paper cites SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning SceMQA: A Scientific College Entrance Level Multimodal Question Answering Benchmark

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.284610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.284610Z digest=sha256:403da011d5e540f1cc5173c2ace80487140dba0d7fbed248d12f508992c07900

Observation 7e320964-10a1-4e97-bd25-4fc8e42274bd · outbound

This paper cites CriticBench: Benchmarking LLMs for critique-correct reasoning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning CriticBench: Benchmarking LLMs for critique-correct reasoning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.032810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.288777Z digest=sha256:a7994d7dba663c0c28cb76e6f314f823263dd5d76f86af38ab75de7af248c1bd

Observation 74e9f7ee-5c98-482f-9844-10ec85db8349 · outbound

This paper cites A review of feedback models and theories: Descriptions, definitions, and conclusions.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning A review of feedback models and theories: Descriptions, definitions, and conclusions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.018355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.292830Z digest=sha256:43cd98227a23e7a5aff4c0e60d1a4d6c00b12243c8ff0e1c04b52e38bf8bc043

Observation e2d80857-d4ea-468f-970c-83f819e67b7e · outbound

This paper cites Mitigating hallucination in large multi-modal models via robust instruction tun- ing.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mitigating hallucination in large multi-modal models via robust instruction tun- ing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:02.004920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.296757Z digest=sha256:021052d8cda28bbb50cabf7afb957777908cf937f463ecdffb2d95817dfcba70

Observation 15578dbd-a851-4d6d-bfbc-08b1f7faac21 · outbound

This paper cites Visual spatial reasoning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual spatial reasoning

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.988665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.300766Z digest=sha256:b146d60cef58154520e94b18f76be30a50a1dca398049f96adee0e6093fdebb3

Observation 287c2df2-1bdc-419e-ac8e-d25f345302d5 · outbound

This paper cites Visual instruction tuning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual instruction tuning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.974796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.305095Z digest=sha256:f53be611fd749f9eeecc029c6c0efae60097953c1ab2ca0b773cfb8c5d2f0dbe

Observation 048c6073-7b70-4163-8e3d-71da8631ad0b · outbound

This paper cites G-eval: Nlg evaluation using gpt-4 with better human alignment.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning G-eval: Nlg evaluation using gpt-4 with better human alignment

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.958593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.308955Z digest=sha256:2cbfcaf3d67b8dc1ad6f361bb0bbc0d683de2a3f24efaab4e8d8462ba9ac3b66

Observation c529c18e-e9da-48dc-b009-f7231aa7fe71 · outbound

This paper cites Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.313217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.313217Z digest=sha256:f996f2a386917991de535b35c9bac2884fdc30c34b3da07a24972528c3617477

Observation 15e597ab-3e6e-413a-984c-a4165215ce81 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.318500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.318500Z digest=sha256:22f4bb15a4ba419f4bf4df78ad86cd816a616416f89f9b8b11febf6b78ea0d8c

Observation f72533be-0a1b-4e4a-9f70-3e6bf50dcd26 · outbound

This paper cites Mathvista: Evaluating mathematical reasoning of foundation models in visual con- texts.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mathvista: Evaluating mathematical reasoning of foundation models in visual con- texts

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.938980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.324543Z digest=sha256:f0b24e387946b0eebbe3d4252f8bab7df7c59c53e7abf9ad52f3097099d26704

Observation 8b41e188-7c91-424f-84c1-57385e539c05 · outbound

This paper cites Critique ability of large lan- guage models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Critique ability of large lan- guage models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.921691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.329019Z digest=sha256:513ac3f53b16cde48cbcb12f3505b7a73fa6a36be64a53de2660b11983d8bbd6

Observation 295ed2a0-8791-4bca-a44a-7a93d0994266 · outbound

This paper cites Self-refine: It- erative refinement with self-feedback.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Self-refine: It- erative refinement with self-feedback

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.906622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.333372Z digest=sha256:f3100044d366ed5ef699908631ccd9295034d2786eeee2d7655fc79fbfea0ff3

Observation d0ee6a8a-43c3-4cad-929f-38d5699f49ef · outbound

This paper cites VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning VURF: A General-purpose Reasoning and Self-refinement Framework for Video Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.337445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.337445Z digest=sha256:a449927c363f3c4ffb002432375d6abdf6ae17b7db344eb1734fd3832cdeb743

Observation 91d0db1b-4a48-4267-ad83-70b100bcb11c · outbound

This paper cites Docvqa: A dataset for vqa on document images.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Docvqa: A dataset for vqa on document images

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.890380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.341791Z digest=sha256:17916580f60883b767d9b65c0642da144586beb99742c1a7216990f744680599

Observation a7e04bbe-4173-470d-b52b-934e79b78f98 · outbound

This paper cites Teaching clip to count to ten.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Teaching clip to count to ten

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.871000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.346088Z digest=sha256:8e4830b2de4b64ff1ffb3c4dcbcf7738fe5a83488944079b5e57666c61613f47

Observation 13451074-2a35-41d4-bb9b-ac59fd2721c8 · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.350264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.350264Z digest=sha256:ba21915ce068ea926cf7d5b8a72c0b2731af25f9844ab69908978be653ac35d1

Observation 2f64a04d-0c6f-471e-b7ab-6a52a5da28f9 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Reflexion: Language agents with verbal reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.855269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.356781Z digest=sha256:a85a71d95d9d33110479097d8d2163f7e14d85583108ccc4d23000e7a3b50536

Observation 4d9bd026-2547-4cbf-a1bb-c340674fa09b · outbound

This paper cites Towards vqa models that can read.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Towards vqa models that can read

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.840709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.363023Z digest=sha256:a599d3b20e87bc2254fe832477dbf16c911aaa7b7e6c00da370c7d39216487ef

Observation 2e5908c1-0acc-4ce4-81aa-c4328dae202e · outbound

This paper cites The critique of critique.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning The critique of critique

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.825638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.367561Z digest=sha256:4760cc6a6902a5dbb3502ed999d7bd35df98d4df1a6fc5f5a4cbd0aeec1690b1

Observation 5ace5606-4cd8-4ee6-9d25-f975a7f27fa8 · outbound

This paper cites V oyager: An open-ended embodied agent with large language models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning V oyager: An open-ended embodied agent with large language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.810760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.371834Z digest=sha256:9b78165a79c01f0676803a51f7db5cf4476ebe811344fe340ac3cdb216656a79

Observation 0adbef5c-b4ce-458d-b6ac-65f388403b1a · outbound

This paper cites Measuring multimodal mathemat- ical reasoning with math-vision dataset, 2024.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Measuring multimodal mathemat- ical reasoning with math-vision dataset, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.796244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.376648Z digest=sha256:19df8d78e02128cee4caedafcfbb458e144c27bdaec0f1795faacfecf9fb25ae

Observation cabf7503-6a20-49cf-ae6d-a0cda7d4b7b4 · outbound

This paper cites Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.386240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.386240Z digest=sha256:2e61e0cb3d909a6041737d327eb715904286c0f80b4d65ff249a83d417ea0b4e

Observation d30d8293-00ff-4c4c-b6e8-13062621e2a1 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.390763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.390763Z digest=sha256:b987f74d31e750fd062ac4d34c2df78afca3fb03ca1a57e2247579ee169772dc

Observation 368581c2-16d2-4cac-a399-0af9728c28d4 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Chain-of-thought prompting elicits reasoning in large lan- guage models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.395285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.395285Z digest=sha256:53994f303844411709b856cfbe60698404678321591cf7652d05b3cc48f140e0

Observation 3ec563ee-04b7-467e-8554-70ffaec55c47 · outbound

This paper cites LLaVA-Critic: Learning to Evaluate Multimodal Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning LLaVA-Critic: Learning to Evaluate Multimodal Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.399332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.399332Z digest=sha256:c27d7cebfff879ffaf16f544283167eff3642a9dbe8c1c1ed34221af92964d5b

Observation d20efb07-160b-47ae-8c55-dc6ccb1b621b · outbound

This paper cites The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision).

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.404124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.404124Z digest=sha256:60630790a6c48cf2edf8233bad8c1c5d250eee776c7190794ff843b1a03aaaaf

Observation f890a9cc-27c9-4399-bace-76ea6ca2d0d9 · outbound

This paper cites Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Idea2Img: Iterative Self-Refinement with GPT-4V(ision) for Automatic Image Design and Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.408742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.408742Z digest=sha256:4f1df8e743b72d73b98a271ebd1e64371c76770a388bde569e5b217191942795

Observation dd676919-e4e8-43f3-b461-a1fae30b1154 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.413464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.413464Z digest=sha256:e97e9e274b015f9a40a1049783cd1c167a7f52444a0ba59cab945c56f3806bd1

Observation 797cbb9e-1926-4fff-b7c8-2b222324a8b4 · outbound

This paper cites Attention Prompting on Image for Large Vision-Language Models.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Attention Prompting on Image for Large Vision-Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:01.417826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:01.417826Z digest=sha256:24ef976522b77bc4501d5fee59e78c6a74935cdba943f1578bf2ac2f81a925c4

Observation 02bcea21-f9af-4bcd-88bf-20ca402493a6 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for ex- pert agi

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.763309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.422119Z digest=sha256:2d9f73a391cafbbfcfac93cb21263fac847cbf19abfcc02cd002e6cea538dd82

Observation 6b7fee0c-10f1-4fc0-96a3-63ac76c253bf · outbound

This paper cites reasoning.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning reasoning

Reference 48

Resolution
malformed identifier
raw_fallback, observed 2026-08-11T23:50:01.749783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.426118Z digest=sha256:072c28e2ab5391c41f0c80faf68f64a3782588ec0f56068d3acef41c50afa10b

Observation 0a48c339-e45c-40e3-ba31-97bfebcbaad5 · outbound

This paper cites The final answer is: MODEL RESPONSE ANSWER - Critique: CRITIQUE FOR ANSWER Figure 28.

VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning The final answer is: MODEL RESPONSE ANSWER - Critique: CRITIQUE FOR ANSWER Figure 28

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:01.733663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T23:50:01.430675Z digest=sha256:fb8f8a088fdd5344e521025b27f11217050fedaeeafa7c42fc0466c888fcbe76

Pith citing papers

Observation a43603fd-a2b5-49e8-a1eb-5e56b51e8161 · inbound

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models cites this paper.

MMRefine: Unveiling the Obstacles to Robust Refinement in Multimodal Large Language Models VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:39:59.412070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:39:59.412070Z digest=sha256:eda4fffe665271b4d5e1f183f3f27417f37b47c1ffdb0872fb5ba01d5d66b0e9

Observation 2430184c-629d-4eec-8bc3-f9736569496e · inbound

MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning cites this paper.

MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:51:06.738821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-06T18:51:05.455071Z digest=sha256:d11f43afb43f0ef3e24c771a27ff2517bb8a3ab675705198089bbb070be8b707