Pith. sign in

Paper Citation Record · LEDGER

Lost in Time: A New Temporal Benchmark for VideoLLMs

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2410.07752.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.07752 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:19:36.820475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T19:07:17.473192Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f39631f1-dd69-4503-b0e7-984f504b983c · inbound

Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No! cites this paper.

Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No! Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T19:06:46.161997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:06:46.161997Z digest=sha256:b52b157b747adf8ddf7a940206711c7a01e32b6d9c5ef350bfd5299bb29020c1

Observation 0f9ee443-d1d2-492a-bd74-25d790104880 · inbound

HD-EPIC: A Highly-Detailed Egocentric Video Dataset cites this paper.

HD-EPIC: A Highly-Detailed Egocentric Video Dataset Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T23:26:20.299457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:26:20.299457Z digest=sha256:651f623833c9671eba03dc4b67c2077961b9d5205508132879c2392dbdacc01d

Observation eb4cdb90-95e4-4514-9f5a-af4e447517b5 · inbound

PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding cites this paper.

PerceptionLM: Open-Access Data and Models for Detailed Visual Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-16T12:19:36.820475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:19:36.820475Z digest=sha256:94e139ab1128aaec4fb409a9425d3ba44472acef9bbeded5f479333922d296c4

Observation af4f657b-d3a6-443a-b8cc-74356a31ff63 · inbound

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark cites this paper.

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:40.452655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:46:40.452655Z digest=sha256:f2534613fe2ccda85cfeb839e6a6e7fe3fdc2801ec315cc97621d7a9b36794a0

Observation 8a4a9d7e-edc7-43bb-92fd-b818933fd4d9 · inbound

MINERVA: Evaluating Complex Video Reasoning cites this paper.

MINERVA: Evaluating Complex Video Reasoning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:40:19.041344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:40:19.041344Z digest=sha256:69ca92e4d8e242738e580361e2e40a72346e0e1523296dae741e61d5f2a0c85d

Observation 769e071d-a57c-4a42-a863-a23d22964c86 · inbound

Seed1.5-VL Technical Report cites this paper.

Seed1.5-VL Technical Report Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:26:06.098537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T05:26:04.960844Z digest=sha256:c8737aa9bc343f84458b4b37a5ea3b0562cc5d8b7c229b9819ec476da2a6c502

Observation e75fe3f7-da28-4ef7-8baf-7a88741e82dc · inbound

VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models cites this paper.

VCRBench: Exploring Long-form Causal Reasoning Capabilities of Large Video Language Models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:59:33.668469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:59:33.668469Z digest=sha256:a9bd13512cb954edc238a3b633b2c6263ecb0693b9fe6a5bfcc67fe0e50b6ac0

Observation ec8e0ee3-de3c-4682-8d6c-becc925384ba · inbound

Fostering Video Reasoning via Next-Event Prediction cites this paper.

Fostering Video Reasoning via Next-Event Prediction Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:40.921513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:10:40.921513Z digest=sha256:3b3ecafebedaccf5c1be3c97c76c1649955e1666d8305bf5e3de1fa0df8b4350

Observation 0773ce32-d25b-4c9a-a11e-7d9cc1e2eb03 · inbound

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models cites this paper.

EPFL-Smart-Kitchen-30: Densely annotated cooking dataset with 3D kinematics to challenge video and language models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:42:34.028134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:42:34.028134Z digest=sha256:843107791d8d408435afa1c739bc745469eeafe654d14decfb5d0feef204666e

Observation 03cff194-f5a0-4074-90c0-a2701db95f6b · inbound

How Important are Videos for Training Video LLMs? cites this paper.

How Important are Videos for Training Video LLMs? Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:51:07.790019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:51:07.790019Z digest=sha256:3b0e35c157e0e87710e8abb6560fa9f3c1c239c7611ee1c10e81114a6ca59917

Observation 2ddcf280-f912-4b04-a12e-5e84b7c485e2 · inbound

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning cites this paper.

V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:33:50.755434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T00:33:50.471804Z digest=sha256:b73245771d40a33e164f0759b754dd60f2f938a12c788aa838b38118301649d8

Observation 3f7009cb-616d-44a6-bd44-2b6e59c6e59e · inbound

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos cites this paper.

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:22:55.823070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:22:55.823070Z digest=sha256:39a50c7a35a62dcc509740982b1804657627b28829fca544f062e475bd6779be

Observation 0b62f425-e764-4385-a3c0-7c7e35137384 · inbound

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models cites this paper.

MESH -- Understanding Videos Like Human: Measuring Hallucinations in Large Video Models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T20:32:56.477678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:32:56.477678Z digest=sha256:66a16dcdded32a0fb0f841a93f7d83ac44813469f49d6ff3b31c9112d531cfbe

Observation 8bda65d5-9efd-44e8-89ea-24f8062f3b04 · inbound

EgoExo-Con: Exploring View-Invariant Video Temporal Understanding cites this paper.

EgoExo-Con: Exploring View-Invariant Video Temporal Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T07:23:03.592902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:23:03.592902Z digest=sha256:897e105eaec5f87a691bc79279aa502177cc2892461ad49b21484b0f1e753981

Observation 916a81a1-e40b-4e6d-a595-405f1fcbb1e5 · inbound

Adapting MLLMs for Nuanced Video Retrieval cites this paper.

Adapting MLLMs for Nuanced Video Retrieval Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:21:18.917253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T22:20:09.051957Z digest=sha256:044d2c53618a592a94e94fc9d429a892e7e3d664e9cc438dcdb6454575cca13c

Observation 8e55d311-7132-4f32-8d5c-1edf6eb53b06 · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.381093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:dc7ca0e41106abdb65e66add14ab0c1b3e666b746cdb2712ad14fc72d31398f3

Observation 97dcf00a-7928-4d4c-98d7-667d7095f848 · inbound

Lost in Adaptation: Layer-Selective Recovery of Temporal Reasoning in Video-Language Models cites this paper.

Lost in Adaptation: Layer-Selective Recovery of Temporal Reasoning in Video-Language Models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:06.215848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T15:42:43.948462Z digest=sha256:1c28304651a0f357d32648b0512915aad8ce83c525236c36d556ae7b6a9537d1

Observation 2c98fe22-2bb7-43dc-9a7b-2927aa9238db · inbound

PushupBench: Your VLM is not good at counting pushups cites this paper.

PushupBench: Your VLM is not good at counting pushups Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:08.857453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-08T08:32:47.354181Z digest=sha256:2c6f8f728be400530254750697f41170452ba6b0578a8fca0fe89057faea478b

Observation 8fa4fcd6-dd2b-4686-abb4-51a589436198 · inbound

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs cites this paper.

Tracing the Arrow of Time: Diagnosing Temporal Information Flow in Video-LLMs Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:05:53.171872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T03:04:27.841522Z digest=sha256:0c381ee331ef18935d5d7f67ee6597281e67d88d78ebf4b3017612049a3ad054

Observation eb6ab733-4622-4b3b-85ec-36cb1c12f2cc · inbound

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding cites this paper.

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T16:03:08.026421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T16:02:53.887605Z digest=sha256:233451c2b6537ad7a91a2c3c8b592a0a58f4e83da476e80e714822f6e9b7998d

Observation e8bd32cd-f16d-464d-ac4b-2b0408354d5c · inbound

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning cites this paper.

Learning Spatiotemporal Sensitivity in Video LLMs via Counterfactual Reinforcement Learning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:46:14.915462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T07:45:56.473188Z digest=sha256:aba402f65dbefa96b18ba675f434ed83afb4d05354066d5b6ce3bed8dea1fd43

Observation b9223b89-0eee-487b-a66d-6691a801b119 · inbound

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis cites this paper.

VGenST-Bench: A Benchmark for Spatio-Temporal Reasoning via Active Video Synthesis Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:14:42.891160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T07:12:02.612292Z digest=sha256:673c8440141c1d5ff05437412f80e95af32b3af1b92e221db10bff6e1b398653

Observation 54d8863a-6e9a-4d9b-94d0-2e2186023ae7 · inbound

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs cites this paper.

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:44:38.756827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-22T05:41:39.396469Z digest=sha256:0cac0db96d39a9637b25aae44f600291b4fd634ec86153a334047f5802a7cccf

Observation 41b3e6fe-42de-40a4-820c-17a67258d8c3 · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.474609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:898f4faf834fbe8c41478af6a9ff3b44d1546c984d2bd736c4d4b700292e8ee0

Observation 00b81ec2-478a-4f2e-81e9-cce4e09940af · inbound

Do Video-LLMs Actually Watch? Diagnosing Character-Tracking Failures in Long-Form Video cites this paper.

Do Video-LLMs Actually Watch? Diagnosing Character-Tracking Failures in Long-Form Video Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T07:12:52.433947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:12:52.433947Z digest=sha256:bc39c0be2be726095f3d725303fa3f0e27964db3f087fb466b68de3a6f957664

Observation 9bddb493-8b2a-4d5b-9e5e-2f5ace7445cc · inbound

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks cites this paper.

Accuracy Without Grounding: Diagnosing Visual Dependency Dissociation in Video LLM Benchmarks Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T05:39:58.511387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:39:58.511387Z digest=sha256:9c79fc95bb8444216c5dbbf55bd70df9b1a84e9ffff6c8825c737753e56f713a

Observation 8e3ba013-2d83-496e-9ee5-82869176d604 · inbound

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding cites this paper.

VideoChat3: Fully Open Video MLLM for Efficient and Generalist Video Understanding Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T00:44:46.152973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:44:46.152973Z digest=sha256:8903f29cb936b4465766cb8c3c2e5e744ceb89e8e32a3ffb69f095621ddd6e43

Observation 24a8a509-1e7d-4172-b845-151ab4876e45 · inbound

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning cites this paper.

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-04T17:24:16.126597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:24:16.126597Z digest=sha256:c069dca2e93ab2c45273eb6d3ed0da5b68761da0133f466e8772a67c15f2e7b6

Observation 6b53c5ab-bd8d-4503-b7a6-f4ac34dce6dc · inbound

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models cites this paper.

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models Lost in Time: A New Temporal Benchmark for VideoLLMs

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-14T04:35:47.905761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:35:47.905761Z digest=sha256:6f6d9d10fc4b5a1ca9454ea04653651874b71dcff56ba1b486ff282042b54391