Pith. sign in

Paper Citation Record · LEDGER

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models

As of 23 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 5 inbound Pith citation observations for arXiv:2411.09105.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.09105 v2

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T21:07:09.837694Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:24:53.867176Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T07:14:42.853373Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved64
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1bdb651-79a7-46ed-9674-93d15a01f126 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.376101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.376101Z digest=sha256:49f0ea79bea7b91f0884fc14979d9663461158dcf536d86c49977fb4cc5a4f2b

Observation a720ad0d-0558-4cfa-a758-0af1f16aed42 · outbound

This paper cites write newline.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.381970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.381970Z digest=sha256:6f2ca96a8f26cbe64f016f0b33eb1e26355aa52fa9e2f01ff9f1dffb6190f037

Observation 3357dec2-8a46-4f80-810d-3de64d27cc96 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:11.051842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.388025Z digest=sha256:683ee749da25da9be82efdbc0bd1f686b7f1b7fff1daf79fcd06b22cf2568073

Observation 917cb33d-f78f-4e41-883f-1c9b2eb49ee1 · outbound

This paper cites D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models D.; Dhariwal, P.; Neelakantan, A.; Shyam, P.; Sastry, G.; Askell, A.; et al

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.393094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.393094Z digest=sha256:75cc8c9db40340685ddd82aebd762536db4f1794f8fd59049190160d7add2145

Observation 1512597d-7a36-4870-8895-35fd4e258513 · outbound

This paper cites TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models TemporalBench: Benchmarking Fine-grained Temporal Understanding for Multimodal Video Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.398031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.398031Z digest=sha256:86cc7b0b62780cf82480c9fdb31a1d5655bc5266452cfa84263e3ec8093c12e2

Observation 080bc55c-8a69-4a73-ba6f-4cce3e691b1d · outbound

This paper cites PCA-Bench: Evaluating Multimodal Large Language Models in Perception-Cognition-Action Chain.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models PCA-Bench: Evaluating Multimodal Large Language Models in Perception-Cognition-Action Chain

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.403391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.403391Z digest=sha256:e25942e93fe15c4886772d6be824985d224e019383d987169990b1f71484d1bf

Observation d9eba277-63c9-4c93-94c0-364ff809b4c0 · outbound

This paper cites AutoEval-Video: An Automatic Benchmark for Assessing Large Vision Language Models in Open-Ended Video Question Answering.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models AutoEval-Video: An Automatic Benchmark for Assessing Large Vision Language Models in Open-Ended Video Question Answering

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.408657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.408657Z digest=sha256:8c20e15f8e19971f889e5d9218a9e16f32def885a3afae602cf24888e4d86136

Observation ad00d8e1-6121-4d96-919f-394f222c3f1a · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.413879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.413879Z digest=sha256:655a52112384361ac240cb9921d85565cc75b1278cf85174efb70d2d78c8e3b5

Observation bbd45266-383d-44be-b227-af87489deb28 · outbound

This paper cites PuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Abstract Visual Patterns.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models PuzzleVQA: Diagnosing Multimodal Reasoning Challenges of Language Models with Abstract Visual Patterns

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.420518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.420518Z digest=sha256:2ec835dedcd70ba9227fbb466eb09d5a0f7b5f9ec7ed4b835043c5e5aff06040

Observation 4ed089a6-92d1-49ad-9322-caabd366aa80 · outbound

This paper cites On the Measure of Intelligence.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models On the Measure of Intelligence

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.425709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.425709Z digest=sha256:d78f4540431e734a96f235e9b7834a159e31f714c2fda097072c7b1de7a2d6f2

Observation a5a4d92e-344e-4c89-a34b-07fb45d3282a · outbound

This paper cites TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.430884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.430884Z digest=sha256:23771b7d9d66c161e749de816206200b3de0a27c5296cda00943b15107a2f865

Observation a608ca77-cfc2-44d4-a7bf-e483c8ba218f · outbound

This paper cites CogBench: a large language model walks into a psychology lab.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models CogBench: a large language model walks into a psychology lab

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.436948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.436948Z digest=sha256:9b0bf187b11f3e7e3986acb426d5edb7d30105e97ae4b6ac8a297b52eb77031c

Observation baeeebfe-8bd6-4f47-8e6f-e87d0a690af9 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:11.023912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.442172Z digest=sha256:fc5e4b878d524f87ea1e46c22286970c43d40d0c8fd4763ee8717b57c73a372e

Observation d7487486-8b20-42e7-96d7-c58ac5d1533d · outbound

This paper cites MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models MMBench-Video: A Long-Form Multi-Shot Benchmark for Holistic Video Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.446997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.446997Z digest=sha256:8d8d8050351e42137a2d74e63b41d5f7e6bf824780b6c6bdc74b410ce7203cf7

Observation 21db5bb0-2e97-4efa-a790-8c6953049514 · outbound

This paper cites Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Test of Time: A Benchmark for Evaluating LLMs on Temporal Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.451997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.451997Z digest=sha256:dac795e41f79d012ccd126d98a9748b3e86b4178ed66cc4b7a24a7e67c4ff2af

Observation 68c32a3e-396b-4089-95a5-3891d0bfabf7 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.457055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.457055Z digest=sha256:9d1ae0d1cca93c9e5d3bf5fcc35aeab412c71114c5a3984ed13cfbaf59e3fb0f

Observation f292ca44-60d5-42d4-8745-548452642129 · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.462244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.462244Z digest=sha256:bf552bf844a5b596754c12967dd4570da6dcad2619b0a07497617aec75f7b491

Observation 85907818-1608-44f8-b939-c07d911e3541 · outbound

This paper cites CATER: A diagnostic dataset for Compositional Actions and TEmporal Reasoning.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models CATER: A diagnostic dataset for Compositional Actions and TEmporal Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.467237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.467237Z digest=sha256:d51cc640aff7342627ac12796028d7855aa1539ce77a3b14ef7f521a6888c587

Observation 94b06391-f27a-48eb-9d7c-27e565ff3a5b · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.472277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.472277Z digest=sha256:a252522f4f1e6193005fcdd61e5b40ff302c5e38ee686dd79aa2e8e09a09ef6e

Observation 3f6f5d99-86c7-43d4-91d8-0ec1c45ea9a3 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.994972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.477304Z digest=sha256:38a32711691bb0754f62cf74981de17113bc2ca6e652d9e8cced768938ceecd1

Observation 46879492-8c03-4a69-8c6e-9b624dbdd0c1 · outbound

This paper cites A.; and Manning, C.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models A.; and Manning, C

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.482925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.482925Z digest=sha256:5d8bdd3c1c7b7e45be7661d2db3cad9124a476c94c0c38b0fb49c8f33d4c9ee5

Observation b3d4ec45-210d-454d-aa92-1c4999e0bed2 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.487759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.487759Z digest=sha256:02f7ba3bb035161736d35e7a0acb31983830544c23666cb046eec1b69d5d2517

Observation ac2db6cd-a4ab-4f68-aed5-db2aa46ad9ce · outbound

This paper cites S.; DelColle, J.; and Sch \"o ner, G.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models S.; DelColle, J.; and Sch \"o ner, G

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:07:10.954389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.492325Z digest=sha256:03ae22d5ecabdbb7b307438363f49aa373685447c7e0b31bcd95a40bf40c7af8

Observation 6cff3290-27c2-4879-9f6f-f9deb9e5da27 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.496923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.496923Z digest=sha256:4a8930b1222ea105361e48944ea11f320093843e27cde128b72eaaf9823ba2c9

Observation 7a632651-159e-4ded-88d7-975ce08104d9 · outbound

This paper cites VideoChat: Chat-Centric Video Understanding.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models VideoChat: Chat-Centric Video Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.502122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.502122Z digest=sha256:185439acbe48beb2e08e6716da559cac2c14247d1bf8fa01eed8d90c8cdb1740

Observation d8ef572f-8d29-457a-855d-9a0754a8e208 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.936187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.507086Z digest=sha256:beb917a0d75c487599d90d5eb46e79e834f65d6a1df98a2a6df102cdf6b059ec

Observation 08a72715-60ec-46a6-abf1-d16612d8939f · outbound

This paper cites VideoVista: A Versatile Benchmark for Video Understanding and Reasoning.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models VideoVista: A Versatile Benchmark for Video Understanding and Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.511757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.511757Z digest=sha256:46374432a179b868a429be505010435d133a40b44452f777e7cbf8875d842aad

Observation 54e4f9be-ed6a-4720-85fc-d917d3f1a685 · outbound

This paper cites LLMs Meet Long Video: Advancing Long Video Question Answering with An Interactive Visual Adapter in LLMs.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models LLMs Meet Long Video: Advancing Long Video Question Answering with An Interactive Visual Adapter in LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.516657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.516657Z digest=sha256:a00ba1fead4be37bbd9d2b1988a2af67210d0bd550349e780c38ffedfb0bd87a

Observation 4ee44efc-f12a-47e8-9b97-1e8eda21f8ee · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.522525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.522525Z digest=sha256:e4002c47a92968db9bd735613206af5b8dfc435e92487ac598deb483ce01efe7

Observation f0256118-2a9b-4386-815c-ab983a873a59 · outbound

This paper cites Video-LLaVA: Learning United Visual Representation by Alignment Before Projection.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Video-LLaVA: Learning United Visual Representation by Alignment Before Projection

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.527241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.527241Z digest=sha256:fc99a2c356a16d8d24b5502f3bc86617934d58dafed11ec9061518c2da83b548

Observation 42c9d412-8ec0-4ce2-9ffc-7687c580ca4e · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.532063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.532063Z digest=sha256:a0728d507363955ea363716ac5c446050a106a07679cba1789ab5e659421cc78

Observation 1e5cf068-6aaa-4e5b-8259-d334693f4461 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.536625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.536625Z digest=sha256:165472997784a1dcf7e85459b8f5a2ae48df63ec2b9a8c070d303f73eb8dea11

Observation 9f60b207-8e8f-4495-a49b-05844b9bac81 · outbound

This paper cites Efficacy of Synthetic Data as a Benchmark.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Efficacy of Synthetic Data as a Benchmark

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.541917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.541917Z digest=sha256:c5d7944760fec61c74f269c0b1a77672bc86ff3b3ac437335b4f96601a75be03

Observation 949845b2-0bfa-4975-a618-adef139acd77 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.893395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.547098Z digest=sha256:edc04bf0365cc013c7dafb729ba3777e58d2f0b8046a574380b1e6481400cf26

Observation b6b4e4a2-ae7d-461a-9333-58291d36358b · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.876878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.551710Z digest=sha256:ef34a1e0972b5ab61bee4286f6003173a132815604aea5c4c8b82ee48ae86c8f

Observation e5477cd8-9c92-4497-8925-47646329c903 · outbound

This paper cites Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Video-Bench: A Comprehensive Benchmark and Toolkit for Evaluating Video-based Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.556330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.556330Z digest=sha256:68f80ad85170ac7f065cc8fe2c74647ef5c1ab8a5c478b6c0c09e7179fa182d4

Observation 7f9ed7c4-ff04-4eb4-a80f-442a372b2095 · outbound

This paper cites C.; and Patterson, M.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models C.; and Patterson, M

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T21:07:10.860983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.561350Z digest=sha256:46b289c6d1d198c1581a673e596e7902d7a9295d6655032f41ed396c7efd040c

Observation 9cc484da-e42d-458d-8e18-446241c58428 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.844055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.565962Z digest=sha256:da2f8cd0f14b5013da68c86825841bf3e20ba166e101ad40fe3002a048ed0571

Observation a6cade7a-6a50-4f9b-b56a-191a400c4c5a · outbound

This paper cites W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.571121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.571121Z digest=sha256:2b86f4cc710f3b4ecdf44167e4dd0bd803c8695f923f2225e95e674fdb6786cc

Observation f7895394-712d-401f-9b96-9fb37b02e581 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.816182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.575815Z digest=sha256:87cf324b81e42156d2391a6035557433c1f6afec5b26d32d6ed08b9d16cfbe55

Observation f187e768-fe91-4bd5-95f2-4dd7e172bb0a · outbound

This paper cites M3GIA: A Cognition Inspired Multilingual and Multimodal General Intelligence Ability Benchmark.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models M3GIA: A Cognition Inspired Multilingual and Multimodal General Intelligence Ability Benchmark

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.580866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.580866Z digest=sha256:bac77b253b9b5c89a29722a46f447580c84f36997665723755b5e8056672ff09

Observation 3629f069-0da4-46cf-b10d-d31563959724 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.800213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.585992Z digest=sha256:64c7ee6300aa43ef899624324a5553355e42bdf2c4dd9fda81cc5bceb99589ed

Observation eedc3ade-408e-4380-b59c-fbef4059985a · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.783116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.590834Z digest=sha256:28eabea8842403b8bd8b72d99b293afd12ff346a56a29e9d7331ae7bb3f48f13

Observation c5d6c871-2d21-431e-be7a-085394175121 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.765737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.595666Z digest=sha256:27811699e96dc9f49cd638f05c8965fc1ad7044e98cd8a9df32feda26fab815d

Observation e2f3bc07-bcb6-4bad-9c11-2ffacc987b27 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.602078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.602078Z digest=sha256:7afe625b7f93f114d98c7c8ace5f1b710a8b76bc728ca6bfafb48375d92e91c9

Observation 75befaa9-c1e8-4a71-808e-685ae0678850 · outbound

This paper cites GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models GRASP: A Grid-Based Benchmark for Evaluating Commonsense Spatial Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.606990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.606990Z digest=sha256:0f19b2fe559cb383101e781551bb555fde65952eeeae05f39112be2742799a4d

Observation b355ab5f-66e8-49ae-b4f6-c361919d3535 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.749823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.612120Z digest=sha256:896a95a5377203478e7f8097f66cb958b2f8eeed0014761979f949872a503ae6

Observation 170f29cb-6092-44ad-b741-81ee34101ba9 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.733450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.617145Z digest=sha256:b7f46707c2a8608da79b6822a802cf4aecde16c108179961e867e347f45fc4da

Observation 8338a222-219c-412c-9307-ef4b26e7953a · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.716696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.622220Z digest=sha256:56dd2d7e0c8bdc357543bfa7a22ae65869f3e5fcbe4b95e01cb44ba8e14e82b0

Observation 68b48b43-e03d-41e8-9b26-71609b8ff68d · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.627012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.627012Z digest=sha256:f86b18ecdbb4e28b23175d1c193c8bcd1a032cab30c057ea80c1a3c1daa7790f

Observation 2e5dbd48-71ad-454d-9af2-607ea06c9ced · outbound

This paper cites InternVideo2: Scaling Foundation Models for Multimodal Video Understanding.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models InternVideo2: Scaling Foundation Models for Multimodal Video Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.632537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.632537Z digest=sha256:157d943be9db6e7061a25083e61d23ed1340d1a0d079cf764ec74b4d92171c93

Observation 86e7cf85-9c5d-4019-b1f2-7dd6ca2338ed · outbound

This paper cites Q-Bench: A Benchmark for General-Purpose Foundation Models on Low-level Vision.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Q-Bench: A Benchmark for General-Purpose Foundation Models on Low-level Vision

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.637648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.637648Z digest=sha256:4e6c1d449e0a73e55643875e466e30eedbab048f38394771af6773ed1c7c4539

Observation 0044c37b-b761-4c17-922a-f350461adaa7 · outbound

This paper cites Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Mind's Eye of LLMs: Visualization-of-Thought Elicits Spatial Reasoning in Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.643393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.643393Z digest=sha256:dbcba7b09602640eb1db23dbd34470bdaabed55b57bdac8959ce409dd276d762

Observation 0f6ca42a-c266-406b-8868-9bf44571fc3b · outbound

This paper cites SmartPlay: A Benchmark for LLMs as Intelligent Agents.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models SmartPlay: A Benchmark for LLMs as Intelligent Agents

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.648291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.648291Z digest=sha256:e31e94962f1a45b1809a1fb9bb2fd3d38ad5fa653821daed7b6312857778a57d

Observation 57cb7372-eca5-4aa4-999e-9095cb26859a · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.653340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.653340Z digest=sha256:70922a18d4722b9013c7163ae00cc433b453ed1aa1fc8f743e475915dfbc7579

Observation e6a2f3b7-9e67-434d-96b4-8ebd827c55a1 · outbound

This paper cites Benchmarking Benchmark Leakage in Large Language Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Benchmarking Benchmark Leakage in Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.658023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.658023Z digest=sha256:e84bcac51f983aa30f5f667eff3f3f2365d2928eb789461b97dcdc956a85dbb8

Observation 8fb67b1b-4ac0-4d0d-b40f-dd2c72c3077d · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.666187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.666187Z digest=sha256:0cf5fc2899d0dd492c2b7ca10c750514e1c65b39b46c5fb2655c9a5e6c686442

Observation dfe7488c-1925-449b-bec5-65aa8d029d79 · outbound

This paper cites mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.672200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.672200Z digest=sha256:526f996f7eaa416fa62731dbc97eb7654499568df14e8bed0818a8d53c5754c1

Observation dfa55c7f-583c-4b79-8d20-77922b9aea4d · outbound

This paper cites CLEVRER: CoLlision Events for Video REpresentation and Reasoning.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models CLEVRER: CoLlision Events for Video REpresentation and Reasoning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.679386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.679386Z digest=sha256:192c1446e45babd5ef0818194e814ea07ca03bac1f13abf6677b30fe76948e2b

Observation 7e200888-8cb4-4918-9ff1-e9ab74149b24 · outbound

This paper cites an unresolved cited work.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-12T21:07:10.700673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-12T21:07:09.685279Z digest=sha256:1d1f0dc584b83477845024634054f8cd25831a09abc145b574c6d43ca34d4ac5

Observation aa05bc51-012c-4354-ac54-bd002a452672 · outbound

This paper cites Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.690669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.690669Z digest=sha256:f0d72def0e2127794515ef509d117f3f4ef3b69842a8f0d5a050dacfd6cf1801

Observation 0aac9e5e-0f66-4bf0-a1a7-02d79bdb414e · outbound

This paper cites InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.695865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.695865Z digest=sha256:65bd95a838ad663ce03ae3f6f6e99075bb2aff3be075831de78ac176e4dafb2f

Observation eb796474-3953-4ff2-8cb8-dc6755805c17 · outbound

This paper cites InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.700807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.700807Z digest=sha256:602242e42c69ef7592656b1575f4d83f75244654124ed9fd6f708a740ce7d3ef

Observation 1934594f-908d-424f-9f67-dc307956d68e · outbound

This paper cites j.; Gui, L.; Fu, D.; Feng, J.; Liu, Z.; and Li, C.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models j.; Gui, L.; Fu, D.; Feng, J.; Liu, Z.; and Li, C

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.705943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.705943Z digest=sha256:f64c2ea14638def98c03155ee9a0a1cb1fc37e77392e84e0ede93db2e53500ba

Observation 481d6df3-0466-4a50-8f54-abb673b320a3 · outbound

This paper cites Q-Bench+: A Benchmark for Multi-modal Foundation Models on Low-level Vision from Single Images to Pairs.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Q-Bench+: A Benchmark for Multi-modal Foundation Models on Low-level Vision from Single Images to Pairs

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.831974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.831974Z digest=sha256:cffd1795486334cb94908d2a5728f4d47f9ee6b075cb8c6fbb2df908011351a4

Observation 34f92016-d8aa-4c88-8589-6eb863ce04cb · outbound

This paper cites Needle In A Video Haystack: A Scalable Synthetic Evaluator for Video MLLMs.

VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models Needle In A Video Haystack: A Scalable Synthetic Evaluator for Video MLLMs

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T21:07:09.837694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:07:09.837694Z digest=sha256:c9c63d4b48ec7fbf8ae94277e954d363ea6879f08372710e274626727bc91527

Pith citing papers

Observation 646c4465-2b89-473c-baf8-2a16e1d685de · inbound

VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning cites this paper.

VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:17:51.910722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T12:17:42.135851Z digest=sha256:a0a46262a7366d555ee0be6878f8df3a2162f217ab887e5767193b5e179be91e

Observation 51bd2479-57cc-4e29-b0a8-764a3edb7cef · inbound

Air-Know: Arbiter-Calibrated Knowledge-Internalizing Robust Network for Composed Image Retrieval cites this paper.

Air-Know: Arbiter-Calibrated Knowledge-Internalizing Robust Network for Composed Image Retrieval VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:04.498239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T02:08:38.289340Z digest=sha256:04851721fa2f811e99ff3783b380a52bd8594a9c048f4eac740e3a57c58378de

Observation 551c5ae9-8135-43df-bba8-853d44d382ba · inbound

ConeSep: Cone-based Robust Noise-Unlearning Compositional Network for Composed Image Retrieval cites this paper.

ConeSep: Cone-based Robust Noise-Unlearning Compositional Network for Composed Image Retrieval VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:49:48.499885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T00:47:18.249980Z digest=sha256:e6c0983c47362d26afbafa9229850568ea8396689608b7d7a41a6b2b8fd14a52

Observation d54fa693-f376-44c6-b6de-0231f38a8dac · inbound

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis cites this paper.

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:14:42.855986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-22T07:12:02.612292Z digest=sha256:7acca59d2b61e20c3ebf0ddeb3ed556f2121f3b12e0ccf8d8830a1068a9ba231

Observation 81537453-421d-4dd5-add1-b1ca663a9126 · inbound

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping cites this paper.

The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping VideoCogQA: A Controllable Benchmark for Evaluating Cognitive Abilities in Video-Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:24:53.867176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:24:53.867176Z digest=sha256:5e8d19ac6bfc29f6d7919d3e9567a7747a3fee600a70d564fd6a8d182fec4400