Pith. sign in

Paper Citation Record · LEDGER

End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.15047.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.15047 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:56:37.569807Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T10:48:12.855049Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 11511695-2036-4371-bb90-7c06e4173f08 · inbound

ReasVQA: Advancing VideoQA with Imperfect Reasoning Process cites this paper.

ReasVQA: Advancing VideoQA with Imperfect Reasoning Process End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T15:56:37.569807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:56:37.569807Z digest=sha256:3f9f913590df435310b744345ac59aa49795ff6906ca6f2f7acf7cbade19932b

Observation c5b6ce89-4f84-4b2b-bdbd-d2ca24eee092 · inbound

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval cites this paper.

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T00:11:23.171904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T00:10:01.596422Z digest=sha256:26cbaff4dc13f11cb175182d61fcfe8f3214725b145ecfb1edfd601a3ddd9dd9

Observation ec4b263b-f997-4731-b04d-6a7916ed4723 · inbound

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling cites this paper.

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:58:25.539956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T00:56:47.841355Z digest=sha256:4a21baba994f1788b8a63ca3383189e44e98b2fe5d79afa85fbbfe52c3135a81

Observation 005c2245-a687-447f-985f-a480817be04c · inbound

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration cites this paper.

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:21:09.542544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-09T20:11:11.410051Z digest=sha256:e8fefa340d141d8e9e00c6ba942727d2603d2119ede5ca4c711867cc071bb7b4

Observation b08716df-c478-4c33-97e5-9bd74edc60af · inbound

CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering cites this paper.

CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering End-to-End Video Question Answering with Frame Scoring Mechanisms and Adaptive Sampling

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:48:12.856768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T10:45:46.597723Z digest=sha256:6118939eb2c3d3a0065862055486fef4be578cf14ceee6d26c3783b3bc880709