Pith. sign in

Paper Citation Record · LEDGER

VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.06491.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.06491 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:15:07.890675Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T02:56:28.983694Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d74d232f-e4f9-40be-b7c5-f9d0342000ba · inbound

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning cites this paper.

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:56:07.815156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T20:56:07.247122Z digest=sha256:f184484e948a15636c657d61db7edea3aeae791f85e6eec801bbe332f66cbf66

Observation a8ace931-45ae-41a2-bcf2-826ac6c4ec44 · inbound

SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications cites this paper.

SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:07.890675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:07.890675Z digest=sha256:f1230b3d98fd71e0193db3dfc8e195c83642f15720279870cf05936382391324

Observation 9e4bb98b-e927-4bd8-927e-3c38eb987333 · inbound

"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth cites this paper.

"Harmless to You, Hurtful to Me!": Investigating the Detection of Toxic Languages Grounded in the Perspective of Youth VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T05:15:42.277048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:15:42.277048Z digest=sha256:5f5abf33e75fe2dd51a0db3c967952e3a49a77b9048f0be79060250ff36f56d4

Observation 018d2dbe-f6e4-4b6e-8b39-75981ee43f78 · inbound

VLM4D: Towards Spatiotemporal Awareness in Vision Language Models cites this paper.

VLM4D: Towards Spatiotemporal Awareness in Vision Language Models VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T05:12:36.307349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T05:12:36.307349Z digest=sha256:19b984bbffb3d34b1282bf7a0efe1f41881e925212ddd8b83fb96723b41db28a

Observation 1bf45568-0dff-4d0b-8fa8-f2a19810e54e · inbound

Where Do We (Not) Need Temporal Context in Low-Resource Video Task Adaptation? cites this paper.

Where Do We (Not) Need Temporal Context in Low-Resource Video Task Adaptation? VideoEval: Comprehensive Benchmark Suite for Low-Cost Evaluation of Video Foundation Model

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:56:28.985709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T10:31:10.409674Z digest=sha256:f556f8c7a69ee3ca9cb3094d6e06e6e5030ca50ad03c01c80f9431be78675f6f