Pith. sign in

Paper Citation Record · LEDGER

Adaptive Keyframe Sampling for Long Video Understanding

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2502.21271.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.21271 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:20:58.751103Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:16:01.333138Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f24fb277-c4a9-41b8-a85c-ff10bdafcc62 · inbound

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval cites this paper.

LiveVLM: Efficient Online Video Understanding via Streaming-Oriented KV Cache and Retrieval Adaptive Keyframe Sampling for Long Video Understanding

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-22T14:31:40.790642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T14:26:59.015559Z digest=sha256:81d0c7a05bb854b2eb639beadb5de32c95ccf2c8a855b9c987b324c67b24d942

Observation a5fc3f40-2601-4cb8-a4ee-395a76b206cf · inbound

ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning cites this paper.

ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning Adaptive Keyframe Sampling for Long Video Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:20:58.751103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:20:58.751103Z digest=sha256:220bff64120a43e5a397938e2819eccb97b36e4ea8380b71e22d900b4a588e1a

Observation 3df37c4b-fbea-44e8-971a-bf77ec8063c4 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes Adaptive Keyframe Sampling for Long Video Understanding

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:07.977401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:07.977401Z digest=sha256:9d1fc229ac614027e9154e476498f4c19d30c34aa3ec3421c467fd994415e6e2

Observation cf1d15db-9e10-4c41-918f-dccbe1e2c010 · inbound

DATE: Dynamic Absolute Time Enhancement for Long Video Understanding cites this paper.

DATE: Dynamic Absolute Time Enhancement for Long Video Understanding Adaptive Keyframe Sampling for Long Video Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T19:28:31.595682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:28:31.595682Z digest=sha256:d9447b2176efc048aec84b0b512d00edb01ca91a6f8625cd0150507196423b56

Observation 854a22aa-db1d-4618-ad27-7850edfbf80e · inbound

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval cites this paper.

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval Adaptive Keyframe Sampling for Long Video Understanding

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:11:23.225426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:10:01.596422Z digest=sha256:79226c592938875e90ef23d02528636fae64f9ef12392f814c8d47f7166e6f54

Observation 296506d5-3962-4999-992c-24480b648df6 · inbound

LFS: Learnable Frame Selector for Event-Aware and Temporally Diverse Video Captioning cites this paper.

LFS: Learnable Frame Selector for Event-Aware and Temporally Diverse Video Captioning Adaptive Keyframe Sampling for Long Video Understanding

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:22:52.390755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:21:19.606631Z digest=sha256:549b17e7f212e36bdf807322e22625969288440e74d453003b717049d3adc826

Observation 3df1e43e-e7fd-49b4-8f1c-7de83fe1a2b6 · inbound

CREST: Curvature-Regulated Event-Centric Sampling for Efficient Long-Video Understanding cites this paper.

CREST: Curvature-Regulated Event-Centric Sampling for Efficient Long-Video Understanding Adaptive Keyframe Sampling for Long Video Understanding

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:06:18.749863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T03:06:09.753634Z digest=sha256:02aec7813cbf537a5eae51d67feb46e5a92fb442553a939ee307dc39bd1c6c14

Observation e2f49626-964a-4c93-ac9c-e8522015cd63 · inbound

TRACE: Evidence Grounding-Guided Multi-Video Event Understanding and Claim Generation cites this paper.

TRACE: Evidence Grounding-Guided Multi-Video Event Understanding and Claim Generation Adaptive Keyframe Sampling for Long Video Understanding

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:37:47.870546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T21:36:43.971666Z digest=sha256:9f76f3bd9a2ae3fd7284ec5568b297934de95054d0e43a172e968c830034afd9

Observation 2ec06c4b-e044-4144-88bf-174adad26ad6 · inbound

TRACE: Evidence Grounding-Guided Multi-Video Event Understanding and Claim Generation cites this paper.

TRACE: Evidence Grounding-Guided Multi-Video Event Understanding and Claim Generation Adaptive Keyframe Sampling for Long Video Understanding

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:35:01.054222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T19:33:24.558488Z digest=sha256:0bee2ecad8bb8c0ab73ca5e35eccb945ef973a30fab56e36cca2a57f103ca305

Observation 4dfe50cf-dc96-4921-b60b-be19ca401a1b · inbound

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning cites this paper.

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning Adaptive Keyframe Sampling for Long Video Understanding

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.956663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T18:15:28.261185Z digest=sha256:2f548bc7d6c1d218398e3384522d6351ef4cf215c9d2ffe1d013819a7f8b4d10

Observation 330df6f8-1c78-4022-b323-0fb6cdf2b1ed · inbound

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning cites this paper.

Q-GeoMem: Question-Guided Geometric Memory for Video Spatial Reasoning Adaptive Keyframe Sampling for Long Video Understanding

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T15:54:12.909974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T15:54:12.909974Z digest=sha256:e93a213715a16d3d8003681e1565973b22a922f93974a68ed7c24b574c4bf896

Observation ee821844-32b4-469b-9792-2844de2a1265 · inbound

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs cites this paper.

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs Adaptive Keyframe Sampling for Long Video Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T15:34:37.002001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:34:37.002001Z digest=sha256:b1b118152636135e866c979781d9e270ffdb3f315c98eb5d18330217a87b0c17

Observation fc0e7d17-ec2d-4cdd-9b5e-3d7b1b8e264e · inbound

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs cites this paper.

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs Adaptive Keyframe Sampling for Long Video Understanding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T18:39:24.915547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T18:39:24.915547Z digest=sha256:4c6fcfb9279418425de293b7029765c1247453a9ae1b70c4e6219e1aec908247

Observation b8cf0f28-bf73-4a21-aa0d-5d8c0131bc4c · inbound

PEEK: Picking Essential frames via Efficient Knowledge distillation cites this paper.

PEEK: Picking Essential frames via Efficient Knowledge distillation Adaptive Keyframe Sampling for Long Video Understanding

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:01.334887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T22:46:53.704892Z digest=sha256:4f6421cd0834319f720f9a17015e7d677f5f3d25590bc19c2e0576f684021d52

Observation c6a83721-15f9-4a6d-bcd5-146ea0611902 · inbound

Agent-Computer Observation Interfaces Enable Dynamic Computer Use cites this paper.

Agent-Computer Observation Interfaces Enable Dynamic Computer Use Adaptive Keyframe Sampling for Long Video Understanding

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:04:21.259945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T06:59:20.295818Z digest=sha256:f5f3c72f1a33bff67865c3ca649b688ecbdbaefba892f528132075279b30e319

Observation 6a7e20e5-6fe1-49cf-b008-2c2004a582a9 · inbound

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs cites this paper.

DynTrace: Tracking Dynamic Object Evidence for 4D Spatio-Temporal Reasoning in MLLMs Adaptive Keyframe Sampling for Long Video Understanding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T06:34:29.611393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:34:29.611393Z digest=sha256:da755c36af5708cb517437d100fc476e5aa440d07b2771ac454692ceeb49c570

Observation af7d562b-05a6-4f33-bbc3-475a9f2a52a7 · inbound

MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning cites this paper.

MarineEVT: Advancing Event-Centric Marine Video Understanding via Visual Tool Reasoning Adaptive Keyframe Sampling for Long Video Understanding

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-31T23:10:26.849086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T23:10:26.849086Z digest=sha256:6ed90dba902ecadace495c806b51fd5740af6d3f9453897cf289d83d87539e25