Pith. sign in

Paper Citation Record · LEDGER

Localizing Moments in Video with Natural Language

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:1708.01641.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1708.01641 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:47:17.345702Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T20:46:13.589892Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d98d0760-e99f-4f4b-95cf-b737423006b9 · inbound

Do Language Models Understand Time? cites this paper.

Do Language Models Understand Time? Localizing Moments in Video with Natural Language

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T12:47:17.345702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:47:17.345702Z digest=sha256:c536731ccb930409dc7835e26a1a59488deb6658bdee32b1b16a013be1e20feb

Observation 82a166ee-239b-4fe5-afe9-e358b162ea9a · inbound

VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in Videos cites this paper.

VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in Videos Localizing Moments in Video with Natural Language

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:13.424298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:26:13.424298Z digest=sha256:6538ea37aeb5f22b86e41998079ddd190f1392d5f593c78b8f2fa01756b97519

Observation edf1b84f-16b8-4360-ad6d-c45d1a51f2be · inbound

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning cites this paper.

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning Localizing Moments in Video with Natural Language

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-04T22:06:58.751821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-08T08:04:15.238840Z digest=sha256:8d57c4a5f78db1ec40724a5f68bc386f17bed688623936df805b4ce3809ab65b

Observation 139f3aab-6821-4eeb-b534-d7e821f6df95 · inbound

TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs cites this paper.

TimeLens2: Generalist Video Temporal Grounding with Multimodal LLMs Localizing Moments in Video with Natural Language

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T18:04:12.732459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:04:12.732459Z digest=sha256:a784c068fdaeba295c713c0e9b04c0c8a9e1b0eabcfc7c5befa18c796a9fec2f