Pith. sign in

Paper Citation Record · LEDGER

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 26 inbound Pith citation observations for arXiv:2501.12375.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.12375 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:58:41.807576Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:00:07.459591Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 56e8e8b7-e70e-44c4-b103-8c88dfbc9e82 · inbound

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation cites this paper.

DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:18:13.976065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T17:17:48.961672Z digest=sha256:717bb608faa9fb0e337247d3f264fd3d2f2ea4978b132a2b609ec51a11402b44

Observation 118f8165-5b8a-471a-a0ce-a0ed023223e8 · inbound

RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation cites this paper.

RaCalNet: Radar Calibration Network for Sparse-Supervised Metric Depth Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:58:41.807576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:58:41.807576Z digest=sha256:0ae00e44ced72fd8d034b5abff78f73de57959aee2cca2f19e69e09af21e8e5c

Observation ddf523f2-5879-4fe0-9116-f32619e7454c · inbound

RoboScape: Physics-informed Embodied World Model cites this paper.

RoboScape: Physics-informed Embodied World Model Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T21:56:09.428813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:56:09.428813Z digest=sha256:ce60ec278ff5108ce14327d24f8da4329ddb1de6481f95c01c8740f70a00dbc6

Observation ecf4f2db-7073-4213-beb6-daefed166cac · inbound

MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion cites this paper.

MM-Gesture: Towards Precise Micro-Gesture Recognition through Multimodal Fusion Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:27:06.939745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:27:06.939745Z digest=sha256:ff34a2b9bf77330df130cd6aa12864a1017f1030562611ce9a3d6f5360762fbf

Observation 98217f03-2f15-440f-9424-261dcd1fb3f8 · inbound

RCG: Safety-Critical Scenario Generation for Robust Autonomous Driving via Real-World Crash Grounding cites this paper.

RCG: Safety-Critical Scenario Generation for Robust Autonomous Driving via Real-World Crash Grounding Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:34:05.903479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:34:05.903479Z digest=sha256:b6a479b12148f30132e595f5a154d1a4cc93e5e4706dcf84a1c98a0bb31e783d

Observation c3c5ea58-a021-4c05-b163-8065c0f45e55 · inbound

SpatialTrackerV2: 3D Point Tracking Made Easy cites this paper.

SpatialTrackerV2: 3D Point Tracking Made Easy Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:49:48.565573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:49:48.565573Z digest=sha256:0607ea408698d4a3ffebbd789a9bc17279878d70a766818f1ee0e8ac82e11aa8

Observation 96348a5e-db49-4c7b-86cc-e3b688b5be64 · inbound

Reconstructing 4D Spatial Intelligence: A Survey cites this paper.

Reconstructing 4D Spatial Intelligence: A Survey Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 135

Resolution
unresolved
no resolver link, observed 2026-08-06T13:02:29.086195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:02:29.086195Z digest=sha256:7bcd7aac775af5bef8191269aa4794ad9dbfd127d0cb5f6134eb59386931f3c5

Observation 3a193b00-fc7b-44d1-a0ab-53127283cb3e · inbound

ViPE: Video Pose Engine for 3D Geometric Perception cites this paper.

ViPE: Video Pose Engine for 3D Geometric Perception Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:41:08.727344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T16:41:08.620285Z digest=sha256:7b2051429daad7f8dbdf773ddacc0fbb38a31059252712ab21aaf3229f0d999e

Observation c11040f0-350d-4e24-a23f-8f6aa3498ec0 · inbound

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation cites this paper.

Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-05T13:46:44.930591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:46:44.930591Z digest=sha256:fbd5f75e97ddc27f95e1a447b4a02257c014e07ec7a7d5c6b9e4acbebd7e195a

Observation 96c1b55d-adfa-44e5-9ef0-2d72c5c1f332 · inbound

Feedback Matters: Augmenting Autonomous Dissection with Visual and Topological Feedback cites this paper.

Feedback Matters: Augmenting Autonomous Dissection with Visual and Topological Feedback Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T11:35:58.419641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:35:58.419641Z digest=sha256:020c38ba05d4ed537604e9708c65dd16a18718a46a4ebc1b06c2b527ee348768

Observation 6be966d9-301c-4180-901d-c6d8cd0236ce · inbound

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction cites this paper.

CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T22:38:37.596051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T22:38:01.008280Z digest=sha256:226f003f1fbcfd5fd672446508b174cde4fb3321551b9cb45a6ce4a753d1ccea

Observation e0f3cdf7-c86a-49ef-a33e-0492f7d05222 · inbound

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation cites this paper.

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T22:13:53.383917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:13:53.383917Z digest=sha256:1878bb4d999c8c365e796b483cd40a1a2182b13f0e7f9f208c08044b65140d80

Observation f87dda88-df7c-41f9-ba6e-74ead34e61c1 · inbound

HOIGS: Human-Object Interaction Gaussian Splatting cites this paper.

HOIGS: Human-Object Interaction Gaussian Splatting Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T17:28:02.544902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T17:24:50.043908Z digest=sha256:01a819671556df53068b0eb056d7699bdedaa5e3a167ef1cf2f92e342f7757af

Observation 2bb3e8db-2aea-477e-b6f7-934d607c3940 · inbound

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data cites this paper.

From Video to Control: A Survey of Learning Manipulation Interfaces from Temporal Visual Data Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:03:01.206201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T17:02:18.358675Z digest=sha256:2ba094aef290fe606f9c8b58d869be6a82a8803a12b390814179bf089a5fdb55

Observation fbd445bf-0125-4f54-bbcb-db9b20b08ad8 · inbound

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation cites this paper.

LuMon: A Comprehensive Benchmark and Development Suite with Novel Datasets for Lunar Monocular Depth Estimation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:00.772158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T17:18:51.726719Z digest=sha256:f768745afb63a3f870a82a7f52c1f0a7d4d4c38f077637fcd7458deaefe32aea

Observation 2c3a4796-b1a6-4246-94e3-9497a21452a6 · inbound

Controllable Video Object Insertion via Multiview Priors cites this paper.

Controllable Video Object Insertion via Multiview Priors Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:45:20.659728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:3eb85b47287cce23fe6b70b5d77d2f8dd33b6504e59104c599a953a8fa0f1782

Observation 3d744e5f-9ae0-4b6c-936c-3826d068ad69 · inbound

GenMatter: Perceiving Physical Objects with Generative Matter Models cites this paper.

GenMatter: Perceiving Physical Objects with Generative Matter Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:01:18.256786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T12:53:18.504365Z digest=sha256:7b0a410589690ec8feda4db7125abd9de970e9a9ce54f03eac5decddc326defc

Observation a9a13e12-5f04-48e4-b699-7817f3f75bbd · inbound

GenMatter: Perceiving Physical Objects with Generative Matter Models cites this paper.

GenMatter: Perceiving Physical Objects with Generative Matter Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:00:07.462600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-04T19:55:05.116930Z digest=sha256:e4d4883b728fc9868b91595e820e95d4af6917015b2121f3659f4b0380fd0348

Observation 86fb3299-e315-4310-b09a-74670ef9662b · inbound

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring cites this paper.

WildLIFT: Lifting monocular drone video to 3D for species-agnostic wildlife monitoring Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:46:40.192466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T04:26:53.225501Z digest=sha256:da6f542891cad1f9afb77b954196f4af4ffea6b5e1973b12febe1db28011228a

Observation 05596cf5-ac2a-47a5-9472-5b6e82108ab7 · inbound

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors cites this paper.

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:26:07.902348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-09T20:05:21.723724Z digest=sha256:5630d2e8196bed303fbf0bce5f716d77a524090092cf27cf842d6a9a48e074b2

Observation dd2f7c12-0bcf-428f-afbb-cffd18c90535 · inbound

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians cites this paper.

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:16:39.146720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T05:16:28.339361Z digest=sha256:16c0cb6cbd9bbb31b3de0eb356656bf93eb44f01f7bc8452dee574a5f3324287

Observation faf378b5-508b-4f61-bdbc-ef01898e89df · inbound

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization cites this paper.

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:02.104002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T23:15:20.254402Z digest=sha256:5806885cb7adc96227404dd23d9cf91ae5a47701a219cb27bf31583598278321

Observation 0bbdd005-672f-4bdc-a482-a51255ba26fe · inbound

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection cites this paper.

Neural Voxel Dynamics: Learning Implicit 3D Physics via Volumetric Feature Advection Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:49:58.020111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T01:17:06.587807Z digest=sha256:49fa1e799449c66fc4c1b86915e15105f607ecf66e194b0b13c99158cc447d42

Observation 1c980236-bde6-4537-a8d9-05d5838c4609 · inbound

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation cites this paper.

PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T20:03:56.956882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T04:34:38.286863Z digest=sha256:44e2a95cabb137b9702f1d31298702901875b1ffbf2aea4914685f2da10ad90b

Observation 7b749165-ee2a-424f-81b7-5157dd18f758 · inbound

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras cites this paper.

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-02T06:16:27.239660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:16:27.239660Z digest=sha256:b87e124a14fe2da9c20e5a1ac13d46a465e314d6d9c44d2ca9041ac83aac8ee7

Observation a1ac38ed-ec1a-4c70-8835-a89107fa843d · inbound

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models cites this paper.

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models Video Depth Anything: Consistent Depth Estimation for Super-Long Videos

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T18:39:26.592889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:39:26.592889Z digest=sha256:0e8df2a347ea9ba61b0d6461fce84fda1f84102ed24cb6d3150f6521bb539321