Pith. sign in

Paper Citation Record · LEDGER

VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2504.07956.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07956 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:20.019923Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T05:49:36.609405Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 674d1b29-563d-4b34-8aa5-498a40a647db · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:20.019923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:20.019923Z digest=sha256:b476f4ea2e7bd86a39396b54551f1e4b69b34482b7d783e0566e6829149e2de0

Observation 07105e7d-bf71-4a16-868a-f72ca768e4c7 · inbound

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? cites this paper.

Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning? VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:40:56.094429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T05:40:55.944288Z digest=sha256:638701f17f5aded13e5ce4b1c4f877245c40750c523fe678e6b96c9c317878f0

Observation feb29216-6bdf-4d59-a73f-aa1be882fccc · inbound

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos cites this paper.

CausalStep: A Benchmark for Explicit Stepwise Causal Reasoning in Videos VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:17.390535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:17.390535Z digest=sha256:d6303386abfa3f774f377abbe63eae704b481becbfa8352b3ebb2ee339175760

Observation 4bb1c417-c44c-49eb-b3bd-7193a9df94e0 · inbound

ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts cites this paper.

ARC-Hunyuan-Video-7B: Structured Video Comprehension of Real-World Shorts VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T13:12:40.063831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:12:40.063831Z digest=sha256:4bf30b68242c7ebcc26619d8b10bfe6d9d9d90b05f54a1366f8c56d643f4ebd0

Observation b4171d3f-c6fa-41b1-aa78-1e9f60fb1f48 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:06.647166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:06.647166Z digest=sha256:dfae781c1c6bacc1e08f47bec9c5ee82cfd04106485a3202f4fb128129560465

Observation 0808bec6-7856-4334-8b37-7b84434757b4 · inbound

VIDEOP2R: Video Understanding from Perception to Reasoning cites this paper.

VIDEOP2R: Video Understanding from Perception to Reasoning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:25:22.616824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T22:24:41.760120Z digest=sha256:5e5083b29b81b84b5381f39c0b0be418a5f84be3eec365ddb3f0f0cfe0af6dd4

Observation 6fadac5a-4015-41a3-bd48-0e9be9b5318f · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.376813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:36963706b53fefb333427395785b0089fce37570c64cf9cafcf2ae96020ca0f3

Observation c016cd63-2897-4f16-92a7-f84038a7cae7 · inbound

Video-Oasis: Rethinking Evaluation of Video Understanding cites this paper.

Video-Oasis: Rethinking Evaluation of Video Understanding VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-13T15:38:41.390945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:38:41.390945Z digest=sha256:3c50bf11a34355193ee40e72c13464a517562ecfaafb11cea633efce9f4a2441

Observation 78654518-e85c-4b52-81aa-8c35bb74e250 · inbound

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding cites this paper.

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:51.277971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:47:32.778695Z digest=sha256:289aa9481a26d04bbb980717b5cbe09557567d73272c77850c45879b84800ca9

Observation f9f80566-ca13-45fd-997a-e3bacfff75b5 · inbound

EasyVideoR1: Easier RL for Video Understanding cites this paper.

EasyVideoR1: Easier RL for Video Understanding VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:47:12.639121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T07:41:27.231098Z digest=sha256:939798f2b09ad2cd392f8f817fcd34444403325f4649496c5bc3cbbc4615af63

Observation 664a807d-d961-4046-9b91-da6887caa58a · inbound

Act2See: Emergent Active Visual Perception for Video Reasoning cites this paper.

Act2See: Emergent Active Visual Perception for Video Reasoning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:45:22.906813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T19:34:53.683729Z digest=sha256:eb3145b178d8ce32f066ad2fd162ca3241bd51eb0f41a008c1e2c5e75f8a909b

Observation cbd3bf73-c4af-40b5-9364-d0063e8024bc · inbound

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing cites this paper.

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:13.408349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T01:30:19.531699Z digest=sha256:efa81459c81797e5a30b6f2be0bd5b88f879a17fdea015b9a72a0ec93165b38f

Observation eb9960d6-39dc-45a2-9a1d-54c5718145b7 · inbound

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing cites this paper.

VEBench:Benchmarking Large Multimodal Models for Real-World Video Editing VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-12T01:46:13.974098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T01:43:34.898639Z digest=sha256:15fb18cd00b0d957dd16b62b3f9ba682a04ab7f314db698c341cf4d051d2593e

Observation d8b2736d-495a-4c49-935c-a4edf52757d8 · inbound

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation cites this paper.

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:38:56.199595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T19:37:09.244578Z digest=sha256:d55eba732010fd8d1e93ebccbe038115cf6fdefc8c8e2df5b8a096734894a5ee

Observation 7a0ee72f-3988-4eb9-9088-b2e897cb7516 · inbound

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning cites this paper.

ConSteer-RL: Steering Reasoning Capabilities in Large Language Models via Confidence-Aware Reinforcement Learning VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:07:23.872673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T19:56:09.820812Z digest=sha256:db000a6cba17ee37caf7d5adc141b9a8c06b3d48045bfd3dabdde77dc3b2a6a0

Observation 29b40ea8-f33a-4fff-82bb-8af022e8d9f0 · inbound

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval cites this paper.

ELVA: Exploring Ranking-Driven Universal Multimodal Retrieval VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T05:49:36.611612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T15:34:55.062016Z digest=sha256:886535439f1a860422830c2a15a5bbb2963cbc2d904890b6f0a6bd2256e963cf

Observation 175b037c-ead7-402c-ad2c-8772d7a5252d · inbound

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis cites this paper.

RoboGaze: Evaluating Robot World Models via Structured Vision-Language Analysis VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:54:36.512557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T10:49:58.019238Z digest=sha256:a7a3f141d619e6fc3be547a2305ddcb9222a9c10d4f1eb09a511d98e2aa79e8e

Observation acc800f7-de2f-424e-aaed-24d5c04790d9 · inbound

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search cites this paper.

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:41.066581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T06:02:48.532478Z digest=sha256:077dc31b8cb8f31d505efdae989311e48680253d9cc34552c21dc883ea8f1a54

Observation 3b8658af-e80d-41ee-954c-72225ab03aae · inbound

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent cites this paper.

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent VCR-Bench: A Comprehensive Evaluation Framework for Video Chain-of-Thought Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T04:44:27.371880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:44:27.371880Z digest=sha256:0ddd7b9cb54628526e0317e1e6d7405d3c71fda188609f50e1c68532b170a057