Pith. sign in

Paper Citation Record · LEDGER

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT

As of 8 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 2 inbound Pith citation observations for arXiv:2505.24182.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24182 v1

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:36:26.033113Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T23:15:38.962013Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T15:57:06.751718Z

Reference resolution

100 of 118 outbound references displayed

  • verified exact1
  • verified fuzzy8
  • unresolved89
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3cd4d0b4-5021-4b0b-b540-c9ce30af7584 · outbound

This paper cites Origins of physical knowledge.Psychological Review, 99(4):605–632, 1992.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Origins of physical knowledge.Psychological Review, 99(4):605–632, 1992

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:17.832939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:17.832939Z digest=sha256:e3486d43f554aab741590c0a70a705016937ce6e57b25bb3b01ee76c15100c0e

Observation 13c7ee59-cebb-44aa-b4b2-d3e4e902b873 · outbound

This paper cites Infants’ physical world.Current Directions in Psychological Science, 13(3):89–94, 2004.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Infants’ physical world.Current Directions in Psychological Science, 13(3):89–94, 2004

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:17.919876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:17.919876Z digest=sha256:d2f6b6463945775b2909f0b4c6332931fc66f2ca0c999867e7e4e2d6e8eb080d

Observation 995443c2-dab8-49d0-a6b3-c5c29f4736a0 · outbound

This paper cites A theory of causal learning in children: Causal maps and bayes nets.Psychological Review, 111(1):3–32, 2004.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT A theory of causal learning in children: Causal maps and bayes nets.Psychological Review, 111(1):3–32, 2004

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.059860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.059860Z digest=sha256:7751678d71cf0a1e5d3c30d9bd54b03edeaf72601d5458609ac0e7296d748641

Observation a44fd7cb-17b2-4bab-b091-4b6ab85da1e6 · outbound

This paper cites Building machines that learn and think like people.Behavioral and Brain Sciences, 40, 2017.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Building machines that learn and think like people.Behavioral and Brain Sciences, 40, 2017

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.157018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.157018Z digest=sha256:39bacd41f4c5d144b89a56f514009d50b79689f01e8efe627e7bddf84a8df1db

Observation 776a1c1e-6dda-4d08-8e04-6c6cb76f62f2 · outbound

This paper cites GPT o3.https://chatgpt.com/?model=o3, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT GPT o3.https://chatgpt.com/?model=o3, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.222324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.222324Z digest=sha256:9d2c0affb76399278ee4e25f0fd344b09af102091e9ca07d144a52721b0158c4

Observation 33fe6ec1-bae3-45d3-83a0-e8015b2aa7a8 · outbound

This paper cites Hello GPT-4o.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Hello GPT-4o

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.298510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.298510Z digest=sha256:2a2a626366423492db4e4401faa80daa017ac37b19c310a5b4af242fb585d898

Observation 396f73ce-a2d5-4881-95b4-2ab07a32adcf · outbound

This paper cites Gemini: A family of highly capable multimodal models, 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Gemini: A family of highly capable multimodal models, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.372868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.372868Z digest=sha256:f02607ef637cd024261c72e58430e30cb0bec58310e0dd2e712279a33a24d98d

Observation 01b07d1d-4d82-4309-8863-116f46653402 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.458818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.458818Z digest=sha256:a62d6d93472c6cf19f487782928df95d752b90df8bcc0a41c50dd1f4d2aba069

Observation 663549a3-a626-4a5a-8f3a-5289fad1bbea · outbound

This paper cites Kimi k1.5: Scaling reinforcement learning with llms, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Kimi k1.5: Scaling reinforcement learning with llms, 2025

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.521583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.521583Z digest=sha256:9767339af868a0b0ed48e19a184c916988ea8c021bd02fcd5c026a529e7bc70d

Observation 3eaaac8c-bd53-4cf9-a801-8daa0779959d · outbound

This paper cites R1-v: Reinforcing super generalization ability in vision-language models with less than $3.https://github.com/Deep-Agent/R1-V, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT R1-v: Reinforcing super generalization ability in vision-language models with less than $3.https://github.com/Deep-Agent/R1-V, 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.576525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.576525Z digest=sha256:3f4e62ba1245817171f3587d9bae65bedcf41f803f836a989fa53dedb22ffd9f

Observation 6f2bb7d3-6f3a-4e60-b81b-95b08d2efed1 · outbound

This paper cites Easyr1: An efficient, scalable, multi-modality rl training framework.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Easyr1: An efficient, scalable, multi-modality rl training framework

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.668837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.668837Z digest=sha256:304d92f7c1c181b6827cfcadbae6ff5ae67c8784651ee241b1964c56d85c1e60

Observation 23031486-063f-4993-8aa9-75308f6d03df · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.739712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.739712Z digest=sha256:94a523d67d8b08b7fac12cd17cbbcaef6fd2ca278350bdec9b8205181e30cb22

Observation c34fc8e9-6d24-4fa0-a12e-972dcae22fbb · outbound

This paper cites Can we generate images with cot? let’s verify and reinforce image generation step by step, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Can we generate images with cot? let’s verify and reinforce image generation step by step, 2025

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.817836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.817836Z digest=sha256:54d48cc0340d49dfb6b4bb9144f5d4b75559fe5674183089cb489d91940f5320

Observation 03e02f1f-4c35-46fe-ad00-18c1b505ff88 · outbound

This paper cites Imagine while reasoning in space: Multimodal visualization-of-thought, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Imagine while reasoning in space: Multimodal visualization-of-thought, 2025

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.900204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.900204Z digest=sha256:3dbdeea8f88dd7e1b046fe67abc113b71e7a834a9ae196a56692dccdb05fd269

Observation 413a23f5-fecc-46b0-8e56-781aef5a4c35 · outbound

This paper cites Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Mm-spatial: Exploring 3d spatial understanding in multimodal llms, 2025

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:18.968292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:18.968292Z digest=sha256:90400f4c669d32a1b1cd440c069363620ba7db9329f2e6790ab51b1fe02e35c4

Observation f6977751-e323-4e53-9119-d5009e4ec2d1 · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.034689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.034689Z digest=sha256:993ad9bc565fb47fa73aec16acacac9614de726604b29cc796dd3737f978450f

Observation 99a5dc0e-e9aa-441d-a29e-a7e888159747 · outbound

This paper cites Mllm-sul: Multimodal large language model for semantic scene understanding and localization in traffic scenarios, 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Mllm-sul: Multimodal large language model for semantic scene understanding and localization in traffic scenarios, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.068799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.068799Z digest=sha256:701b742789f22912c3a2d009b827fc12bd3326c25569bbd76c4356688f5ae7df

Observation 15f5e482-fba9-4813-8c9b-f432791eeeb3 · outbound

This paper cites The second half.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT The second half

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.125517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.125517Z digest=sha256:bf449715210068bb44fd2ee520ef0d1a7b34d7f6d8f6f2b5a4f5b86d0bad6b37

Observation 0485d055-943c-46a3-a326-8f9d22948a46 · outbound

This paper cites How Far is Video Generation from World Model: A Physical Law Perspective.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT How Far is Video Generation from World Model: A Physical Law Perspective

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.219155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.219155Z digest=sha256:d516b49ef66093113d82b358a92cc68565d4dfe1809b232fd63deb8e33683aaf

Observation 967eb7ff-d878-4022-b98b-1174ccc27284 · outbound

This paper cites Contphy: Continuum physical concept learning and reasoning from videos.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Contphy: Continuum physical concept learning and reasoning from videos

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.278900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.278900Z digest=sha256:4a27361617e382d2fef197351d22b7569fb374519b1cc304812b6edda457651d

Observation ecfa2f21-7a48-4473-b7f1-aefb7dd0253e · outbound

This paper cites Mdk12-bench: A multi-discipline benchmark for evaluating reasoning in multimodal large language models, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Mdk12-bench: A multi-discipline benchmark for evaluating reasoning in multimodal large language models, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.363631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.363631Z digest=sha256:2c0f3b8956f12045a5fbba0284dcbf774b396545902def6f4791df48cc3cf287

Observation 4fba22f8-6f71-4112-887c-6250eface497 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi, 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi, 2024

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.454947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.454947Z digest=sha256:c90203c69fc63d4cf0968b0587a4620d66288578e8f3da5a4dbb5cc507b58f94

Observation 297d6855-afaa-4ef4-821a-3b05c75a24fc · outbound

This paper cites Vlind-bench: Measuring language priors in large vision-language models, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Vlind-bench: Measuring language priors in large vision-language models, 2025

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.547264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.547264Z digest=sha256:072a0e09f01428d2dc91642cffbc725ba2115922bd735b652a3ea7f9cb576bff

Observation cd3064e3-fe71-4bab-b0a8-eada1baeb521 · outbound

This paper cites MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT MME-CoT: Benchmarking Chain-of-Thought in Large Multimodal Models for Reasoning Quality, Robustness, and Efficiency

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.638964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.638964Z digest=sha256:6210a1b28f1e2429028152670da940af58477c627ae687783a2d7e049e991753

Observation 6b0461fd-5910-4188-ab00-2abd10dca60d · outbound

This paper cites MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.731991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.731991Z digest=sha256:cfbace406e9d2629be627e404c87a7682c898a3a1ea60945fa7a20c556cf634d

Observation 95c3523d-ead9-4f75-98b5-12ad20117606 · outbound

This paper cites MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.820855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.820855Z digest=sha256:4d7386d4ab0e6ef8dfdbd5e50eac3c5f264e0f0205356000bd6b68984d02d215

Observation c35339db-39e9-442c-84a5-24d7b06a90b9 · outbound

This paper cites PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT PhyGrasp: Generalizing Robotic Grasping with Physics-informed Large Multimodal Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:19.942520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:19.942520Z digest=sha256:5f5c63eec836d6bcf56333fa0e570adab230a8d5bf160e82be0cb496a1d2e37a

Observation e603c6ff-3b77-4693-aa7e-e097e3d5c118 · outbound

This paper cites Evaluating Multiview Object Consistency in Humans and Image Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Evaluating Multiview Object Consistency in Humans and Image Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.033626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.033626Z digest=sha256:527c1fb74f45631062c6603115d251655372b2af475d4b4dbafd29391884bd88

Observation 82cf3d16-46a1-4666-86ca-38edf60f9724 · outbound

This paper cites Open Eyes, Then Reason: Fine-grained Visual Mathematical Understanding in MLLMs.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Open Eyes, Then Reason: Fine-grained Visual Mathematical Understanding in MLLMs

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:36:28.351368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:20.108953Z digest=sha256:ea9ad5d8025cf13cc493cfe6e2dca89632b62fcdf23d6598d753f4dfb0f464c7

Observation 44b7c249-c56e-4905-bf2e-019f1cca6ebd · outbound

This paper cites V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT V-MAGE: A Game Evaluation Framework for Assessing Vision-Centric Capabilities in Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.191213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.191213Z digest=sha256:c451d7481374fcc83e1d1055dafae5a3c3f263619940d101daa407f6c30944f2

Observation cc6d883d-c664-4434-ba4b-aeca9bc13c0f · outbound

This paper cites VERIFY: A Benchmark of Visual Explanation and Reasoning for Investigating Multimodal Reasoning Fidelity.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT VERIFY: A Benchmark of Visual Explanation and Reasoning for Investigating Multimodal Reasoning Fidelity

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.262944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.262944Z digest=sha256:db2fd349e39a8c01a6fb07336fd78c6b6e4160ee77290234a8b6757be45176d1

Observation e2e65cdc-7675-4ef8-b54b-3b7d1475aaf4 · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.318015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.318015Z digest=sha256:19ddff4e4416dd1a7fd40571e062e2983d389cad6044cb68177061edd5b53e1b

Observation 908e0f4b-ba88-4a97-9a71-89a183d4a691 · outbound

This paper cites CLEVRER: CoLlision Events for Video REpresentation and Reasoning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT CLEVRER: CoLlision Events for Video REpresentation and Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.404754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.404754Z digest=sha256:67f15ead63e69b5a065eb72a6f046776d51c727cf33b6ada19ac853598dee8b5

Observation b167a712-eb8e-46d2-8573-c78590142f6d · outbound

This paper cites Physion: Evaluating Physical Prediction from Vision in Humans and Machines.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Physion: Evaluating Physical Prediction from Vision in Humans and Machines

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.520292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.520292Z digest=sha256:d8486867e794feaf80afeea9839901dcf653beb18fa3c1fb3821d8a92593293e

Observation ae3560fa-e2b2-468a-8aab-1a7ca3bbf029 · outbound

This paper cites NEWTON: Are Large Language Models Capable of Physical Reasoning?.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT NEWTON: Are Large Language Models Capable of Physical Reasoning?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.594399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.594399Z digest=sha256:b0dff2bb524c07ea099c668406cd9007cc11e0d245568163703fe76b174c7035

Observation ab2bb033-976e-482b-8352-b7c889977d56 · outbound

This paper cites Krishnan.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Krishnan

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.684237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.684237Z digest=sha256:b7a6d9f7de67ca3346090916dd0874040e1c8b66580f85e2d8c77c9dd52b8a61

Observation ecee111c-99a8-4f8e-82a8-7493c4edd222 · outbound

This paper cites SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT SpatialVLM: Endowing Vision-Language Models with Spatial Reasoning Capabilities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.778339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.778339Z digest=sha256:39fc0b13274b483b6908cbd8f91758b06897f99da04d54513ddfad9a22221ee0

Observation 8774458e-4ede-49ad-bdfb-5ade160e8838 · outbound

This paper cites Benchmarking Sequential Visual Input Reasoning and Prediction in Multimodal Large Language Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Benchmarking Sequential Visual Input Reasoning and Prediction in Multimodal Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.868896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.868896Z digest=sha256:a737289af1bb4154ab3c2ce345e426c903b9dcd3a5848043c1fe508888292432

Observation 67a775f1-c542-4221-ae98-dab8d9f561d7 · outbound

This paper cites Physion++: Evaluating Physical Scene Understanding that Requires Online Inference of Different Physical Properties.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Physion++: Evaluating Physical Scene Understanding that Requires Online Inference of Different Physical Properties

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:20.960291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:20.960291Z digest=sha256:00ae44d4d24a4025e2c6b9be4cf4e0aa3e1fb1e8572b70052fd568dfbb3ea26c

Observation 822f327a-b428-42cd-8f6c-795c96b062b0 · outbound

This paper cites IntPhys: A Framework and Benchmark for Visual Intuitive Physics Reasoning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT IntPhys: A Framework and Benchmark for Visual Intuitive Physics Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.046870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.046870Z digest=sha256:9c1fecdfdc55eab3e238816008f612787520edbdb8412314d911608e96881b29

Observation 2b7f97da-6412-476c-a8ab-1630c8cc4cc2 · outbound

This paper cites VisScience: An Extensive Benchmark for Evaluating K12 Educational Multi-modal Scientific Reasoning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT VisScience: An Extensive Benchmark for Evaluating K12 Educational Multi-modal Scientific Reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.116573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.116573Z digest=sha256:6c1b00adbabb130c8aaf999e3fccec0cea659711f6d7224327c4d9b420c9463a

Observation c463ac39-5abf-407c-a49b-30cfbfc8bd6f · outbound

This paper cites Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.188252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.188252Z digest=sha256:f23be0f1126ea00de3d5f180d1112c1356b429cc935e6bd839c51a9de26a00a2

Observation a5964cce-ce42-4bca-9940-f8b194a075a7 · outbound

This paper cites Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.244871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.244871Z digest=sha256:83619df43fc214d75631dd66ec19d6f29146b395beb860bb53979130de509439

Observation 5c35bf45-4fb4-40b1-8e39-8b6d0e7510ea · outbound

This paper cites PhysReason: A Comprehensive Benchmark towards Physics-Based Reasoning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT PhysReason: A Comprehensive Benchmark towards Physics-Based Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.317514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.317514Z digest=sha256:a54cd19f6bf89c9efb962d876d5819b9e64d3a5a5c4eb3b2ef9a51cf32271e71

Observation ad5b2f0c-bb69-4e52-85e7-1ab692f38bba · outbound

This paper cites EmbodiedScan: A Holistic Multi-Modal 3D Perception Suite Towards Embodied AI.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT EmbodiedScan: A Holistic Multi-Modal 3D Perception Suite Towards Embodied AI

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.373436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.373436Z digest=sha256:5c948f7e39b221309b5d1d04d7254502ba71fbe0b9a6cde5dcce4e2894477e9f

Observation 63384adf-d6bd-4088-81a7-c43adc1af628 · outbound

This paper cites Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.417316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.417316Z digest=sha256:0cee6ff3e540f1968bc8c5c0a189e7293abe64ee4642b72e2ed6432ca3343112

Observation a4532fb0-a63d-47eb-bfe9-83c99d75f682 · outbound

This paper cites An Empirical Analysis on Spatial Reasoning Capabilities of Large Multimodal Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT An Empirical Analysis on Spatial Reasoning Capabilities of Large Multimodal Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.486935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.486935Z digest=sha256:e16217a495638ad4ed724ab968b9e1808d8e7daaa9af591b99baea6f89771092

Observation 7c262a1a-3ac7-447e-961c-842df6a37134 · outbound

This paper cites Proximity QA: Unleashing the Power of Multi-Modal Large Language Models for Spatial Proximity Analysis.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Proximity QA: Unleashing the Power of Multi-Modal Large Language Models for Spatial Proximity Analysis

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.555751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.555751Z digest=sha256:a63d8d8c6919c01e827f0476bf7b73af146f9d5bffc669044d23a0447788a52a

Observation 3d74019f-fe34-4f14-a255-a411d2c9ff04 · outbound

This paper cites PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT PhysBench: Benchmarking and Enhancing Vision-Language Models for Physical World Understanding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.624537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.624537Z digest=sha256:c2b356b280e71e066dea05b8de46fd1ebeaf004b8018e3224456453a668d0a61

Observation 90c1b8a9-7127-4475-8b08-8073d63496b9 · outbound

This paper cites GPT-4 Technical Report.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT GPT-4 Technical Report

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.680691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.680691Z digest=sha256:8ee744522b090381ccba0295e7b87de226bbe5d3977e84d43855d75162528a4d

Observation 46b9bca8-e09f-48f8-8f36-3c31042c8101 · outbound

This paper cites Mixtral of Experts.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Mixtral of Experts

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.723524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.723524Z digest=sha256:bcd14e8381f03aec91fcecbaf506e0844931c170f92f5e3251f3641353cdd482

Observation ea89ced4-9a2c-4a8a-b32b-ea61316a9c9a · outbound

This paper cites Segment Anything.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Segment Anything

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.790062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.790062Z digest=sha256:ef6322e6aa1e9e7e484147bcef9f1df456d89957306b1789ba2cc2a83a179b05

Observation 9c6c0ba1-be81-4917-b8b0-dfa1ce47731e · outbound

This paper cites Personalize Segment Anything Model with One Shot.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Personalize Segment Anything Model with One Shot

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.875766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.875766Z digest=sha256:302c109e9623f671e3a1e0fc58e3a5f68d770e01c8caaac462bb3b955b21a032

Observation 60c21250-ccb4-4717-a05e-19930cb6a5c7 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Flamingo: a visual language model for few-shot learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:21.950051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:21.950051Z digest=sha256:14d95d731733b60dfce55de51f33bd1a33f10e1926fbe2a6f1124bc6d9887d24

Observation 260abcfe-bd4f-44d9-826f-6b836f423142 · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Instructblip: Towards general-purpose vision-language models with instruction tuning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.016171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.016171Z digest=sha256:06e6ec938569a78eb1ac1970aac9dd489e371681a8a12492710d43c1afc6ec82

Observation 6d2a1a43-faca-45d4-8932-d7947869b798 · outbound

This paper cites 3d-llm: Injecting the 3d world into large language models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT 3d-llm: Injecting the 3d world into large language models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.092255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.092255Z digest=sha256:6349c36ab5816e450c94bbc3d59709abe0d5816c76a271435a676455d7dc9284

Observation 66df359e-d2d6-48f5-b4a9-24bd298106d0 · outbound

This paper cites Pandagpt: One model to instruction- follow them all, 2023.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Pandagpt: One model to instruction- follow them all, 2023

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.160080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.160080Z digest=sha256:d80a8fe92d714980fb4e08357b55da64b253ec2dfc3da571404d2b880e920350

Observation 1660023d-958b-4890-88a7-2a40bb1ad2aa · outbound

This paper cites VITA: Towards Open-Source Interactive Omni Multimodal LLM.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT VITA: Towards Open-Source Interactive Omni Multimodal LLM

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.224140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.224140Z digest=sha256:a5f05220ab62a63c22b10453e61df2f935def17546ce7fb2b7135e972ab95d2e

Observation 69a0ebaa-6384-4a20-844d-d20a6dce74e1 · outbound

This paper cites VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.280641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.280641Z digest=sha256:a8aa13f55d2c15c546fc7229c5a34642dc930c643d3fbf092c09c05ed27465e3

Observation 6a9541b2-ccdc-4f3a-a939-c42387392cec · outbound

This paper cites Vita-audio: Fast interleaved cross-modal token generation for efficient large speech-language model, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Vita-audio: Fast interleaved cross-modal token generation for efficient large speech-language model, 2025

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.347165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.347165Z digest=sha256:462e1960a239659530ac351aa673757cb2ac8297710afe0ebfc95e37b5f48cd3

Observation c913f7a9-5fba-4d67-9171-8d4f16a12530 · outbound

This paper cites Video-llama: An instruction-tuned audio-visual language model for video understanding.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Video-llama: An instruction-tuned audio-visual language model for video understanding

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.411001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.411001Z digest=sha256:628784f03a91fbb653322a49add97d9845f75f9179f1f319e9efbfc12e5ab820

Observation f0a2e4ae-da02-4a79-a055-26a160efb31b · outbound

This paper cites Spacevllm: Endowing multimodal large language model with spatio-temporal video grounding capability, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Spacevllm: Endowing multimodal large language model with spatio-temporal video grounding capability, 2025

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.455709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.455709Z digest=sha256:eb4df667751caa0de9519b86d04e41aa845e3753d32005cfacb33199a9e54073

Observation c38a4c4e-a0af-437b-87f9-6d37036f53d4 · outbound

This paper cites LLaMA-adapter: Efficient fine-tuning of large language models with zero-initialized attention.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT LLaMA-adapter: Efficient fine-tuning of large language models with zero-initialized attention

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.458775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.458775Z digest=sha256:ce732887dedef6dcb01914f4eb1220e73cceb99e37e9173dcc57cbcc3948df15

Observation 5039bd49-cdd8-4f8f-acd0-96f3f80d5ec1 · outbound

This paper cites Visual instruction tuning, 2023.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Visual instruction tuning, 2023

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.500101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.500101Z digest=sha256:7cf08515b3b8d7c74a3ebaeb81954d0725734ecea2cdcabd947973d219537c20

Observation a0910820-c1db-4199-8789-e26c10fe9133 · outbound

This paper cites Improved baselines with visual instruction tuning, 2023.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Improved baselines with visual instruction tuning, 2023

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.652265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.652265Z digest=sha256:acafa73181b98f5a4ba9f85bc5aed8bcaafdd355278948934c48af16e924dc7b

Observation ec98eba8-2dbb-41a3-9766-8566ca7da7d6 · outbound

This paper cites Llava-next: Improved reasoning, ocr, and world knowledge, January 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Llava-next: Improved reasoning, ocr, and world knowledge, January 2024

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.725970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.725970Z digest=sha256:284dc03cc62ea05a63cd8ebe2e17bb4e35fd511be5a84fee39b6488f25403ce6

Observation f454c4e6-b036-4bd8-841f-70ce1d281766 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.768465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.768465Z digest=sha256:5650a600500aa0ee54322ab185bef3bc7cea37caa8b102bfedfebee3baaabd76

Observation 799b57d4-03be-4d41-97ff-ac0b46c843a6 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.815837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.815837Z digest=sha256:517b9d47da1f72348375522f3d183d193b6200163e61bf3477b96b37ac7f28e0

Observation df89751e-4604-4235-91ea-81e75435bd92 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Learning transferable visual models from natural language supervision, 2021

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:22.899311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:22.899311Z digest=sha256:b87a27f0b98ce4aba7b0b1a1dd973cd9e927d37dae32bd7fcb4ec24a0c3c8bed

Observation 063ee0c4-8367-4705-b771-9415224150c9 · outbound

This paper cites mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT mplug-owl3: Towards long image-sequence understanding in multi-modal large language models, 2024

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.010251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.010251Z digest=sha256:28ab5fac467dbd4269d710ce7f621c3983c33a9ff11f973663f5bcd4aa9e985d

Observation 0829e115-4579-4cab-b61f-6d19f0feaf4d · outbound

This paper cites Qwen2.5-VL Technical Report.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Qwen2.5-VL Technical Report

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.072424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.072424Z digest=sha256:987ae8750dfdc5c78fe4ee2c9fb7fa329566f33c5e448a2db195459d975d544c

Observation 7a499321-9580-4479-a01d-f51d21fdf84b · outbound

This paper cites Deepseek-vl2: Mixture-of-experts vision-language models for advanced multimodal understanding, 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Deepseek-vl2: Mixture-of-experts vision-language models for advanced multimodal understanding, 2024

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.154071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.154071Z digest=sha256:8bee3ad7ac4d74f5ecf5f0cd3daea4e2829e127002f44ba96d3bfc4efa73594b

Observation fc6cd84c-e397-4d50-8e6a-f4c4ad5091a6 · outbound

This paper cites Kimi-vl technical report, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Kimi-vl technical report, 2025

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.217279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.217279Z digest=sha256:9f9a459157434833dfa27804fd12e62f00231addd6b16ffe7810216d6f65c6a9

Observation c45492da-7ca0-454f-b188-c47c75f7fe41 · outbound

This paper cites Seed1.5-vl technical report, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Seed1.5-vl technical report, 2025

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.280415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.280415Z digest=sha256:37b9000601a3196b8f8633cc51cab95f61bd00655dce3efa06e800226e67db16

Observation 7cc228e6-1605-4f62-abf4-ecd5acaa5b31 · outbound

This paper cites Skywork r1v2: Multimodal hybrid reinforcement learning for reasoning, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Skywork r1v2: Multimodal hybrid reinforcement learning for reasoning, 2025

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.402170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.402170Z digest=sha256:c9f3b8c39f5c12106b0d1b515f75e44302a72fb30157ecaa21b09199fa333a6c

Observation 89ccdf80-22eb-491d-8121-9d1f4642d94a · outbound

This paper cites PhysGame: Uncovering Physical Commonsense Violations in Gameplay Videos.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT PhysGame: Uncovering Physical Commonsense Violations in Gameplay Videos

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.575813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.575813Z digest=sha256:855eb831f56739f8a2f53d95811a196b33558bb0f785da6b35c9bdc3a11f753f

Observation b0d84298-b701-490d-9004-cafeff56dfbc · outbound

This paper cites Embspatial-bench: Benchmarking spatial understanding for embodied tasks with large vision-language models, 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Embspatial-bench: Benchmarking spatial understanding for embodied tasks with large vision-language models, 2024

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.749834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.749834Z digest=sha256:d9df0a122bdf7c17273cdbe7bca6e027419dce307a335fbb6098b37d2d7fc232

Observation e0e312f6-0e91-4162-ad4c-8b73e58f1836 · outbound

This paper cites LLaV A-onevision: Easy visual task transfer.Transactions on Machine Learning Research, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT LLaV A-onevision: Easy visual task transfer.Transactions on Machine Learning Research, 2025

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:34.765235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:23.835605Z digest=sha256:24d129754e76d0439608eb771af9a2cd422102efafa3b75e95dbb014492db688

Observation e654eca7-a0bb-4e73-b503-0d0c8e84f0dd · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:23.942050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:23.942050Z digest=sha256:03852983235a6c56b1c0979fbf1268dff3fd6611bb0fccbdacbfadb643993475

Observation 24f747fc-509d-416c-80e3-8e1d9704fa43 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:24.031956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:24.031956Z digest=sha256:32f2e39c55c764bf8e18af9cfeb6d8af27ace4f2b5dbb98332520bfc426c9568

Observation b3d77043-991d-4814-96cc-777a89c0029a · outbound

This paper cites Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:24.114630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:24.114630Z digest=sha256:ab33b5e2938067c6cc7a1c89f9e9a096bdf7e4289d293d5b85fc5dc3ec1b6680

Observation 5287267f-c1fb-4c71-bd5f-a630bed012fa · outbound

This paper cites QVQ: To See the World with Wisdom, December 2024.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT QVQ: To See the World with Wisdom, December 2024

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:34.558651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:24.200139Z digest=sha256:3ef1245809520a3396804c7a8ee09ee6212dc876bd38ddb1d298eaebdb6d9de7

Observation 51437547-c18b-432c-bb4a-c2b70deba3b6 · outbound

This paper cites Claude 3.7 sonnet.https://claude.ai/new, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Claude 3.7 sonnet.https://claude.ai/new, 2025

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:34.368817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:24.300733Z digest=sha256:364f9c9d6554d8cd73f89e2379cb41a784d7b3b799ef62df9aef4389a1d5cb21

Observation 4d2394ab-553a-4aac-93ea-025791a148fd · outbound

This paper cites Grok 3.https://grok.com, 2025.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Grok 3.https://grok.com, 2025

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:34.160168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:24.375824Z digest=sha256:9b97c61896f38bf22ff0baed0dccd138e2e12a81f5d68a26ee6469f6c43b23b4

Observation ea1e89af-7f03-487e-85f2-927165fd8616 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:24.505754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:24.505754Z digest=sha256:9810c389bbf7846dfb48f6b1c626721d8edfcd12be6ad98c0b5050cd306b7613

Observation 50f62abb-79ea-4f38-a606-848e71f52a4a · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:24.617986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:24.617986Z digest=sha256:9d9ff4a1d6c0643877b9271d5ffb1bc7cf0c849df11478aed0ab60278a17caab

Observation dab57297-c8b5-4a3c-8834-853a5d28a0ae · outbound

This paper cites R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T12:36:24.686510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:36:24.686510Z digest=sha256:093c5e3c39eaad1317fd5f21293981295f9abd76085a57ff58feb7e55c73a4c0

Observation 61e06049-98a8-406e-8436-d662c5ea8d02 · outbound

This paper cites Lighthouse Laboratory.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Lighthouse Laboratory

Reference 88

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T12:36:33.993893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:24.794031Z digest=sha256:91e64e37b1d19d491f08ae4a40da36cc557ddbd6fd6404f8b6c7ec8a2943db8e

Observation 6eeef0b3-d4b8-4833-a11e-7b52f54d542e · outbound

This paper cites How to accelerate the separation of B peels?.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT How to accelerate the separation of B peels?

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:33.827548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:24.862619Z digest=sha256:b0b75d342815582fe69a57f5f9d53922e83ff6ef2412f767708fdefc62c761df

Observation 27d068b8-2425-4329-a5b6-587f40f4b301 · outbound

This paper cites Match": Aligns with ground truth -.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Match": Aligns with ground truth -

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:33.349581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.116132Z digest=sha256:7ee2b2418f898e08b31d842ab81f7dbd9a4302781716dee2f4e36d04c4de6068

Observation d1cff303-65e1-4032-8fb2-3e94366bda11 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:33.161374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.210835Z digest=sha256:f7f7ade612c68825181d997fc89f3b8c02b9059a47f3a60e807e933c33d71801

Observation 46994528-e7f1-4df4-9cc5-0f478736c365 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:32.975801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.270152Z digest=sha256:f5bd6ead45311b2b37609c58faff703ed385d6a823dc6ab1c9ca12b0dad04334

Observation 4bd75fcd-905f-471e-9ccb-710fa60f1e58 · outbound

This paper cites step_type.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT step_type

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:32.788316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.327962Z digest=sha256:cf533ceddd30f0e3c38c31017c10f27df29f4e8f8434527baad8ff79b7595f18

Observation c66b620a-fc0a-4b2c-b5e3-ac3a7dec9427 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:32.636472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.377531Z digest=sha256:2146bb9a88f7d24685427dadc3a1bdf6bffceb77f5c8f94c97c8b26c4d430414

Observation 922b58cc-43e4-4bd1-8e6d-61ed70f34379 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 97

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T12:36:32.446079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.437494Z digest=sha256:802b69ebb97dc329c01b3528b0b68c4af640949df4e1564debbfc026305109f2

Observation bd4f9cf0-0f23-4be6-8ff7-5ebbe941f9fe · outbound

This paper cites step_index.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT step_index

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:36:32.256146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.529787Z digest=sha256:96d37e6249afea707343b97a1583f17c5923ce9f33dbb69dc97c928ff0e7e757

Observation f6e2c853-6c79-4e77-b795-20ce8555daf2 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:32.013586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.673411Z digest=sha256:1aeaccd11c721be943f542e189cb0d07bff359dce30ce74e7371c8a153a58cad

Observation bdad26ad-30d8-437d-8bc8-bb1be0eca3bd · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 100

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:31.807240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.792224Z digest=sha256:600c6bb6f61255e79c52f7747b66de194ff6dac9a0f4dc0e9c142158c9cd93e1

Observation 2d06af11-b0a8-43e7-abe8-135072fcb745 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 101

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:31.634903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:25.911829Z digest=sha256:fd831d5abcb4060b2434062ba3693972ff284f37f2bc82a2266a6bb7177e83b6

Observation 157d4246-ebd0-4774-a782-365a36baeb31 · outbound

This paper cites an unresolved cited work.

Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Unresolved cited work

Reference 102

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:36:31.403584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T12:36:26.033113Z digest=sha256:02d10cb407380f943835de6ccdfed463692400105bb27a79291a12693fc02d29

Pith citing papers

Observation 90ae6915-c9f9-44fc-92cc-1bf860ead61a · inbound

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning cites this paper.

Enhancing MLLM Spatial Understanding via Active 3D Scene Exploration for Multi-Perspective Reasoning Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:15:47.826680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:16:46.753641Z digest=sha256:e6c8a9a502a0fe0925799a384ac451e0e213b8f9f32bed3693ae371d83e9bb28

Observation bdd9a7af-2d5f-44c5-961e-ebf3968c4329 · inbound

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs cites this paper.

Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:57:06.752989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T23:15:38.962013Z digest=sha256:a4f7e72ea2f98ef474c9d9afbea36aa52136db1de3fdb3cf4ebf404832a70b43