Pith. sign in

Paper Citation Record · LEDGER

Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2502.16427.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.16427 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:40.446631Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T19:50:11.312748Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f5a46dd9-92e6-472a-9871-bfdf327cf44c · inbound

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark cites this paper.

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:40.446631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:46:40.446631Z digest=sha256:ae236b5ea82fa5cbc47276e5d99c63c1eee213c1c50482116ec8efdc0d20baff

Observation e9841d58-d9f5-4f29-93ec-bd548fc09fd3 · inbound

Watch, Remember, Reason: Human-View Video Understanding with MLLMs cites this paper.

Watch, Remember, Reason: Human-View Video Understanding with MLLMs Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T17:27:15.765503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-27T22:00:28.350003Z digest=sha256:aec786062e6524232f7a2e54a6441f7a233fe42464e33a2cf760311f9d8a211b

Observation 4a3c1ceb-e825-480c-903f-222887dc8ae7 · inbound

CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales cites this paper.

CapRiCorn-1K: A Comprehensive Benchmark for Video Captioning and Subject Referential Consistency Across Temporal Scales Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:09:41.275122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-26T12:09:01.026544Z digest=sha256:016991a19a493a5ec700ab71f1417a029668fc45fa7b740dc3e3889ed0bbf4f1

Observation 0ed5d67b-c00e-460b-b136-a6212230d909 · inbound

Graph it first! Enabling Reasoning on Long-form Egocentric Videos through Scene Graphs cites this paper.

Graph it first! Enabling Reasoning on Long-form Egocentric Videos through Scene Graphs Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:50:11.314032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-06-25T20:58:01.925342Z digest=sha256:cfa2fa1ce72a09ba16f0ee95616e595ed24820a44a6981d2f66ebd464be0750b

Observation 85cc3e89-454d-48f7-9c21-45b5ea608467 · inbound

Graph it first! Enabling Reasoning on Long-form Egocentric Videos through Scene Graphs cites this paper.

Graph it first! Enabling Reasoning on Long-form Egocentric Videos through Scene Graphs Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:19:02.846566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-03T23:09:01.340759Z digest=sha256:049a97241d081cc5b67ff7cc9fec383ac6bfa6a8af07e6a886a679c1c604f050

Observation 869836ed-d635-412a-8e18-26e31c525685 · inbound

Learning to Evolve Scenes: Reasoning about Human Activities with Scene Graphs cites this paper.

Learning to Evolve Scenes: Reasoning about Human Activities with Scene Graphs Fine-Grained Captioning of Long Videos through Scene Graph Consolidation

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:08:32.383264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-07-03T15:06:06.430083Z digest=sha256:5b91eed3350d835ec22f530e1c5bf55d5910e0c362ef6029e7bfb0728840d1ff