Pith. sign in

Paper Citation Record · LEDGER

Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

As of 3 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2503.12496.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.12496 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T21:32:16.939563Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:39:04.701981Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eab857a9-1f34-4e9a-82cf-91757979e69b · inbound

Mitigating Object Hallucinations via Sentence-Level Early Intervention cites this paper.

Mitigating Object Hallucinations via Sentence-Level Early Intervention Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:35:32.671787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-25T08:31:24.173135Z digest=sha256:744603c0fcae2b82097d71a62a51135eefa032cf7ca578e8e512f640e76ba052

Observation 8be86950-d6de-4d84-bf9c-243eb57add5d · inbound

EgoSelf: From Memory to Personalized Egocentric Assistant cites this paper.

EgoSelf: From Memory to Personalized Egocentric Assistant Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:06:03.813089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-10T02:23:21.119521Z digest=sha256:059a5dff79de4634ace354e0b07761ef669c3d262cbe1b6681ce4b839581701d

Observation fb71b523-eaff-48d3-87a1-a6f35c8a1649 · inbound

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe cites this paper.

Uni-OPD: Unifying On-Policy Distillation with a Dual-Perspective Recipe Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:26:16.162141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T17:00:49.448352Z digest=sha256:c3a2790a073ec7b749ef626ced2e3c8fa01bc02ed68b0797479d06a354a5a1fd

Observation a0eb97c8-35c1-4a53-b269-e86c34f5582c · inbound

Video-Zero: Self-Evolution Video Understanding cites this paper.

Video-Zero: Self-Evolution Video Understanding Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:35:04.432104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T21:32:16.939563Z digest=sha256:f96e3b5297143c6d69deec576e45e0aa2f2b639fb46bded0d2e86289a55cd2b5

Observation d483c9cc-90a3-4602-9dbd-c336ec92fc5f · inbound

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding cites this paper.

How Well Can Your Video Model Remember? Measuring Memory-Budget Trade-offs in Long Video Understanding Does Your Vision-Language Model Get Lost in the Long Video Sampling Dilemma?

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:39:04.705018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=arxiv_source observed=2026-06-26T21:51:04.050833Z digest=sha256:146225285f9ffb44e99c94078ce309aad2bf5322767b60ff0791916bbab4e854