Pith. sign in

Paper Citation Record · LEDGER

KeyVideoLLM: Towards Large-scale Video Keyframe Selection

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2407.03104.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.03104 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T09:17:00.597095Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b2aa391c-367c-48d0-ac88-a8272f731738 · inbound

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs cites this paper.

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:07:15.442129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T11:03:59.222849Z digest=sha256:52153bc9142444058933c99e04d2709a1e60b3eef17477ac8bcefb8e754d33a3

Observation 7e66f06b-3700-47bf-8853-b5f39ea919ff · inbound

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models cites this paper.

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:45:48.672481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T19:36:42.100191Z digest=sha256:325ad55a8cca2e731a15925b715d5a6ebde241948783318cd834b505bbc311e0

Observation 9707cea1-f099-4dc8-96b8-b678058fa7b8 · inbound

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models cites this paper.

OpenWorldLib: A Unified Codebase and Definition of Advanced World Models KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 73

Resolution
unresolved
no resolver link, observed 2026-07-13T09:42:23.808691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:42:23.808691Z digest=sha256:8ee336f42049e06dda128007e98513f7b6ba014b161d479e2090ce4331cb130e

Observation c99e4a4a-b0d5-494b-bd4a-ba7a6dc0c52d · inbound

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects cites this paper.

Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T23:45:50.934752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:54:04.104227Z digest=sha256:dfeaa94e96ab63c41b6389beca62a4ef40a124d81435951729502f855e36bfef

Observation 284dbbde-cc6e-41b3-aa94-546ccfcbc54c · inbound

LDDR: Linear-DPP-Based Dynamic-Resolution Frame Sampling for Video MLLMs cites this paper.

LDDR: Linear-DPP-Based Dynamic-Resolution Frame Sampling for Video MLLMs KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:22:10.129873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T03:20:42.684286Z digest=sha256:f3657067fe2f318b5a3adbc9f8d8ae06e7b8603896a4d7c6a5b84c3ae365a0fa

Observation 02682ce4-8788-49cd-bd2e-0e84da35cd75 · inbound

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs cites this paper.

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T15:34:37.002001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:34:37.002001Z digest=sha256:fbbb426244b1aee8738660ef3fe08497bcac906a9a01f52caeb530915e5d14d5

Observation 457b445e-5618-4d49-83f0-6d365ce1734a · inbound

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs cites this paper.

Multi-Scale Separable Fourier Neural Networks for Solving High-Frequency PDEs KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-14T18:39:24.915547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T18:39:24.915547Z digest=sha256:336d6f92747fcdd4781ee703139ad295ff7e3d2000eec52c013990fceff93255

Observation 77a491e7-d551-4df2-bbed-52b973c054ee · inbound

PEEK: Picking Essential frames via Efficient Knowledge distillation cites this paper.

PEEK: Picking Essential frames via Efficient Knowledge distillation KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:16:01.296697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:46:53.704892Z digest=sha256:3b06e3c09c3bc42e478b771c8178710b983bd2a5a0f98003df2e88ac3dfdbda5

Observation 73a56052-63e4-4ab0-96e0-4439b22c03bb · inbound

AdaCodec: A Predictive Visual Code for Video MLLMs cites this paper.

AdaCodec: A Predictive Visual Code for Video MLLMs KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-28T15:22:19.548791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T15:20:48.248576Z digest=sha256:55807882da01c7b09de95f0d0d1d60b81f4a83cd7406f331dd6b3a6f26d2eb0d

Observation 3c27c65e-00ef-4d6b-a49f-1eadd7d241e0 · inbound

Q-Fold: Query-Aware Focus-Context Spatio-Temporal Folding for Long Video Understanding cites this paper.

Q-Fold: Query-Aware Focus-Context Spatio-Temporal Folding for Long Video Understanding KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:27:56.042192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T10:04:29.739632Z digest=sha256:d6f8af3fd1d47508ce7973c679f95428ac67f091a70b498aeb6959ddf02d4005

Observation e7ad3a05-4fa2-49c3-ab26-2558ecaec89c · inbound

QCA: Query- and Content-Aware Keyframe Selection for Long Video Understanding cites this paper.

QCA: Query- and Content-Aware Keyframe Selection for Long Video Understanding KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T14:17:02.660613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-02T14:09:54.549499Z digest=sha256:4a12503362cf18124efb06bf8fdb362642fc93bf9f97ba459ec95453f3c7ffbf

Observation 278fde52-fa8a-4fbb-9267-5625a466fbdc · inbound

Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding cites this paper.

Reasoning with Memory: A Temporal Granularity-Adaptive Framework for Training-Free Long Video Understanding KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T09:17:00.597095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T09:17:00.597095Z digest=sha256:97c6ee3f5fa6464154c6be0f594b3b781f55a5dd15c198c81637ce9b1eaa0c9b

Observation e8524394-313c-4732-8c3b-06a91ccf5ab6 · inbound

VisualRouter: Query-Grounded Visual Sampling for Long Video Understanding cites this paper.

VisualRouter: Query-Grounded Visual Sampling for Long Video Understanding KeyVideoLLM: Towards Large-scale Video Keyframe Selection

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T06:37:32.960181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:37:32.960181Z digest=sha256:183e0660186cad5934de7b6d098404e45bc4ff765523e86e5f688e06836d6d8e