Pith. sign in

Paper Citation Record · LEDGER

Towards Understanding Camera Motions in Any Video

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2504.15376.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15376 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T07:14:42.446756Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:40:00.919006Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 580e7ed0-d1f7-4435-968c-c454c2334b5e · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Towards Understanding Camera Motions in Any Video

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.740824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:748a03b3c92ac42cf330fc471f09491a53533b1d26c2ddc06c7df249607cb964

Observation 25a997bb-2a43-4045-b099-9f931125154c · inbound

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding cites this paper.

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding Towards Understanding Camera Motions in Any Video

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-05-16T04:21:29.685627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T04:21:29.526008Z digest=sha256:5c0d703fc64453876ef7cde1792b9d20410b248f64bf84009ed73311170bb394

Observation 711ea883-a30c-4f23-9442-4316bc0f85f9 · inbound

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning cites this paper.

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning Towards Understanding Camera Motions in Any Video

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:02:42.471377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T10:02:20.477517Z digest=sha256:02c2221524dae276eb0b099c1b6f42c8f97e84079aa36b9dc3a11eea998dbcb8

Observation 220a4362-b9a8-4aaf-9cc7-53c8e9864681 · inbound

HumanScore: Benchmarking Human Motions in Generated Videos cites this paper.

HumanScore: Benchmarking Human Motions in Generated Videos Towards Understanding Camera Motions in Any Video

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:41:04.336350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T01:17:16.513512Z digest=sha256:a3b5d299d16903f78b7482b6694923667d7076c480cb5752ce422e15603f9277

Observation fc14d044-fdaf-4d48-9efb-2c47221a3f00 · inbound

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer cites this paper.

OmniShotCut: Holistic Relational Shot Boundary Detection with Shot-Query Transformer Towards Understanding Camera Motions in Any Video

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:46:49.149459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T04:15:20.045060Z digest=sha256:7b2aa52a40e6ab75fbfad6274d43ba8a16704c2730d8186dedefa9dc6c40f043

Observation 245941f5-f87c-4279-a090-4c45357dd2de · inbound

DMGD: Train-Free Dataset Distillation with Semantic-Distribution Matching in Diffusion Models cites this paper.

DMGD: Train-Free Dataset Distillation with Semantic-Distribution Matching in Diffusion Models Towards Understanding Camera Motions in Any Video

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:16:30.177543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T17:42:00.634333Z digest=sha256:66b1ec3ddec00e6f73bbed88e4e5a97e5586c5786ceab59683fab519ef9e97ee

Observation 1d7bdf41-e112-4a18-aecf-9514926b0aa0 · inbound

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs cites this paper.

Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs Towards Understanding Camera Motions in Any Video

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:11:15.446378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T02:10:27.595446Z digest=sha256:aafde7dc2769e7f4853a3b6c636ac89c3e4b3e92fbe56c4189f34436d2d56ba0

Observation f2102e95-3756-42a1-8188-7d646072807c · inbound

Probing into Camera Control of Video Models cites this paper.

Probing into Camera Control of Video Models Towards Understanding Camera Motions in Any Video

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.352310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T21:11:19.441408Z digest=sha256:b05c510252bea590ba68febb8d96d7429127e0b0b8fbe3ed82a6b0ff62fbc729

Observation b6df5c30-0758-4b1f-a8b9-e4a030b805ea · inbound

CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models cites this paper.

CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models Towards Understanding Camera Motions in Any Video

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:28:04.580227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T05:27:30.938311Z digest=sha256:f4be3142702d8fa42241c990f30dfc9f3d7570281c6a7cd7c5b38b28665016ac

Observation 309e9f1b-23d3-45b4-adf1-cce801da9589 · inbound

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning cites this paper.

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning Towards Understanding Camera Motions in Any Video

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.920422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-25T23:29:24.520537Z digest=sha256:b7494babfd2328af4ea71c8d76a6b20636e857aaa641683da5f3569b5cc85032

Observation 4f0f49a4-5bcf-435d-8d7f-30686c822220 · inbound

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds cites this paper.

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds Towards Understanding Camera Motions in Any Video

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:49:52.340383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T04:51:38.355857Z digest=sha256:9e8dd2a4cb9ae0eb434efc8ebb9f784cfe033e9830badd97ef27b5d90d967e04

Observation ed033391-79e8-45d1-a3d5-dfdef1130468 · inbound

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds cites this paper.

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds Towards Understanding Camera Motions in Any Video

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:13:52.839672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-29T04:55:32.982338Z digest=sha256:c6e5d11c14edfaba798a955b597e29c0613257a501e2327d4e91ccca13a8acb7

Observation 91c0c318-7fc0-4dc1-a37a-11ee04506787 · inbound

Natural Language Camera Movement Understanding cites this paper.

Natural Language Camera Movement Understanding Towards Understanding Camera Motions in Any Video

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T05:16:19.750976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:16:19.750976Z digest=sha256:afedb11b23127a376a8db06e3b1ea0b08568dd3c4e664722fdf743f7611ea749

Observation 18e482bd-c9b6-4f99-8506-30cf03e0c49c · inbound

GS-Agent: Creating 4D Physical Worlds With Generative Simulation cites this paper.

GS-Agent: Creating 4D Physical Worlds With Generative Simulation Towards Understanding Camera Motions in Any Video

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T07:14:42.446756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:14:42.446756Z digest=sha256:9096c054a3c1758e1a257e1a947fb9293857700885ba93161efbc2ba347d8d43

Observation 28d0b606-de80-41da-bdf2-c07efc56d5bc · inbound

Self-Supervised Learning of Structured Dynamics from Videos cites this paper.

Self-Supervised Learning of Structured Dynamics from Videos Towards Understanding Camera Motions in Any Video

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T07:07:43.948987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:07:43.948987Z digest=sha256:23b1d565dcf60994f72d5782517c2500d495d6890ea7898a9a962a2fd5102e82

Observation a29f6672-bca4-4dc4-97d7-1e24f41ee6f1 · inbound

CameraAnything: Refilming Videos with Arbitrary Camera Control cites this paper.

CameraAnything: Refilming Videos with Arbitrary Camera Control Towards Understanding Camera Motions in Any Video

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-31T11:10:41.548774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T11:10:41.548774Z digest=sha256:c67c0915605f13281b42a86545181e0bb8536e6508eee3e4045fb08ea9c654e4