Pith. sign in

Paper Citation Record · LEDGER

Learning Temporally Consistent Video Depth from Video Diffusion Priors

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2406.01493.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.01493 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 30 of 30 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:48:35.195224Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:59:53.009791Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6a61176e-3495-4abe-b6d3-c608b9851ddd · inbound

MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion cites this paper.

MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 140

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:41:13.196367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-15T14:41:13.120099Z digest=sha256:21ee1ac9660c5b26f970b474cf94a2ac2e0b12871d853cd15007e8f676ac6a07

Observation 6a7e7c6a-74ed-40ba-84bd-f80105681ccc · inbound

Guiding a diffusion model using sliding windows cites this paper.

Guiding a diffusion model using sliding windows Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T19:51:59.272881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T19:51:59.272881Z digest=sha256:0f8dd8323e15454e5b2f61eb73b9abe3b234faf0691b3165bc9882d962b92133

Observation 4a270ddf-dbfc-46e4-8b8c-ead701d2875b · inbound

Generative Omnimatte: Learning to Decompose Video into Layers cites this paper.

Generative Omnimatte: Learning to Decompose Video into Layers Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T12:56:32.905969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:56:32.905969Z digest=sha256:8b948e293de6ae24fa1821b83a486cc414635bed80383fdb6fff20b930b7430f

Observation 55069968-05a1-4281-9857-e5a3511f756e · inbound

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors cites this paper.

Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T12:26:49.531022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:26:49.531022Z digest=sha256:64f5e4048b745936a46392fba76613326cbaaa4cb1b0811c093baad359dcbb2a

Observation 3df7feef-382a-4e39-80d7-41a017370b5a · inbound

Video Depth without Video Models cites this paper.

Video Depth without Video Models Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T10:31:01.850501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:31:01.850501Z digest=sha256:94c2f5919a18b7011b2f486c0b93a1a7311f6f475c79ca696e5ae43b8209f787

Observation 3bfe619a-3acd-44e1-aaf1-a3c59b33b46f · inbound

Align3R: Aligned Monocular Depth Estimation for Dynamic Videos cites this paper.

Align3R: Aligned Monocular Depth Estimation for Dynamic Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T22:54:11.440353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:54:11.440353Z digest=sha256:53d6f273ff5ffa59e10953919a75071719b95ea2c54c9f069097b8a0eb693c74

Observation 54025329-24ce-48ce-bdbf-c6f2d6270f9e · inbound

MegaSaM: Accurate, Fast, and Robust Structure and Motion from Casual Dynamic Videos cites this paper.

MegaSaM: Accurate, Fast, and Robust Structure and Motion from Casual Dynamic Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T21:29:25.339643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:29:25.339643Z digest=sha256:e47c8c89b2e44b5078235d4588bde8db3b7a1140de4749eb966dd9ec6b785bf5

Observation aafbcb54-14f1-4cf8-873b-2916c8460791 · inbound

Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono Fail cites this paper.

Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono Fail Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T21:29:59.011587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:29:59.011587Z digest=sha256:10f3b31d42b722fc966888d0e70c0931bf1f491b58bb958bb6917a81fe1a4671

Observation c11b90d7-78c4-4ca4-9c0d-8ccfc271f13c · inbound

Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos cites this paper.

Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T16:56:07.288939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:56:07.288939Z digest=sha256:4f16f611d90bf5670a7caaa87c2179e099329e807df91487fe2b3579c50d450f

Observation ca8f03f1-0dfc-44e8-adc5-82c4104a9a24 · inbound

Gramian Multimodal Representation Learning and Alignment cites this paper.

Gramian Multimodal Representation Learning and Alignment Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T14:31:23.938986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:31:23.938986Z digest=sha256:6709141a0ee16c797d33393c9943aed660f5b29c792f27458cbdd69b8dbcea04

Observation fbc1eb73-e63e-4fe0-acf4-4bd6ccbff26a · inbound

Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation cites this paper.

Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-10T23:07:40.811860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:07:40.811860Z digest=sha256:19a0282f05ac4bff4178d4b297e1b41bbe70bc7e34f0b375ca83f69a9a35c98f

Observation 58f70e33-90ee-4572-960c-c6778e8f34df · inbound

GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking cites this paper.

GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T22:11:41.543906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:11:41.543906Z digest=sha256:5f12e75030a846dea50eabde359fc65b6cde19a1ad1099af2c6eb17e73149ac3

Observation a557e620-21f8-40ef-aefb-e9015d500253 · inbound

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos cites this paper.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.657944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.657944Z digest=sha256:23568b19551ec7da4c21988712821840949115b32f8f4b6bb442acfb8544ea52

Observation f1803207-3ae8-4a72-9cbc-5169f9ceab7f · inbound

Continuous 3D Perception Model with Persistent State cites this paper.

Continuous 3D Perception Model with Persistent State Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-10T17:20:07.407052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:20:07.407052Z digest=sha256:eb146cfd217c3d298c92c74b41722bf95165f78537f70c8614c9917fc95f258d

Observation 512d3bb8-d14a-417c-8533-99af4b6fbe22 · inbound

PhysAnimator: Physics-Guided Generative Cartoon Animation cites this paper.

PhysAnimator: Physics-Guided Generative Cartoon Animation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-10T12:27:48.132735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T12:27:48.132735Z digest=sha256:e3debc1d7701b8e64247d0dfe2cef4046609689373f656b12ae65709bb180691

Observation af0f1580-c2c8-4d09-82aa-d73e7ef778af · inbound

Seurat: From Moving Points to Depth cites this paper.

Seurat: From Moving Points to Depth Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:48:35.195224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:48:35.195224Z digest=sha256:bd0e4a3b63f6290c0fa5d775c0d6dfb6296e658f08b7e3616420e7b08d470f7a

Observation 90fceb19-691b-4571-b8ef-2c7cb7fdfd17 · inbound

FoodTrack: Estimating Handheld Food Portions with Egocentric Video cites this paper.

FoodTrack: Estimating Handheld Food Portions with Egocentric Video Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:41:42.183684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:41:42.183684Z digest=sha256:d38f3d56604e139906db913f1f1debbf36053b3de544942fbb0310eb1d91e102

Observation 6fdbdbc8-4cb1-4772-8df5-c29784bcd98f · inbound

Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis cites this paper.

Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T21:39:14.292371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:39:14.292371Z digest=sha256:16e9dc6d6adda15ee07a19aa85ca6680a8d45b316d107cc4908f9766ecb1876d

Observation 90fd036b-a57d-4460-88ec-e203654b88ea · inbound

UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation cites this paper.

UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:27:00.948363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:27:00.948363Z digest=sha256:a326bc6b1b2431f4dfb82be30a128cd685eb73fb890b30c121c64d922df6ce93

Observation f131cda4-f8f1-4561-b80d-4812df0f9db7 · inbound

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models cites this paper.

E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-07T11:36:14.176710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:36:14.176710Z digest=sha256:842ac577388c3a338424d2f9a69da5cf401ab26b326c3aac028c05d14494f910

Observation 67426323-b2ae-4a45-a482-0039aa1bae2b · inbound

Vid-CamEdit: Video Camera Trajectory Editing with Generative Rendering from Estimated Geometry cites this paper.

Vid-CamEdit: Video Camera Trajectory Editing with Generative Rendering from Estimated Geometry Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:13.445467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:13.445467Z digest=sha256:d6ae497c63b4b685c0a83a6610f7e1be3bb5e7b87e4704107c25455163f4ba1c

Observation 98fde7d7-a816-4244-9ba4-206e26a065fd · inbound

Geometry-aware 4D Video Generation for Robot Manipulation cites this paper.

Geometry-aware 4D Video Generation for Robot Manipulation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:14:27.927985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T00:12:39.787489Z digest=sha256:f3949dbd53119054f9383f14511d88f11c03336aae2051396f3e02fbb75e6492

Observation f796409e-3e49-47f6-be21-de314265e29b · inbound

LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion cites this paper.

LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T20:26:10.660772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:26:10.660772Z digest=sha256:81f257a86b136a4b782f8ab75a932986442cd3eebc3e8d544009e572ce66dcf4

Observation d7ae55bb-c431-40a9-a22d-8c389d5d5805 · inbound

Reconstructing 4D Spatial Intelligence: A Survey cites this paper.

Reconstructing 4D Spatial Intelligence: A Survey Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-06T13:02:29.073998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:02:29.073998Z digest=sha256:a9d060c39bd3c7cb647bc41d441f76387efd02fc9fcb5a23372414d8eb48dc7a

Observation f30f277a-3710-4c28-8e54-deda479ec830 · inbound

Towards Consistent Video Geometry Estimation cites this paper.

Towards Consistent Video Geometry Estimation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:16.090588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T08:03:13.579650Z digest=sha256:6465f1fb954568f73f60e7b82325c2e5756b58f7fed48a7dd1c3dc2a9ddf19ff

Observation 8f5898a9-8bd1-48b8-9522-3f3fe2e37313 · inbound

Towards Consistent Video Geometry Estimation cites this paper.

Towards Consistent Video Geometry Estimation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T12:54:18.419282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:54:18.419282Z digest=sha256:903792a5d43261d4758461d2099f6fe5ed09f5ea3412f7e5b2fbaabe6e30db9c

Observation 9f176d27-5a97-4dc9-8559-9d6c3f1a766c · inbound

Forget, Anticipate and Adapt: Test Time Training for Long Videos cites this paper.

Forget, Anticipate and Adapt: Test Time Training for Long Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:59:53.011409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T05:34:05.109347Z digest=sha256:0c56c610c941984f6d8eacc78a4a0dff8411720a52a08bf7b246fa235658e64b

Observation 5667fa5c-f124-420f-8ec6-fd91b86bd231 · inbound

Forget, Anticipate and Adapt: Test Time Training for Long Videos cites this paper.

Forget, Anticipate and Adapt: Test Time Training for Long Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:54:38.960634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T10:22:49.075651Z digest=sha256:f26a7281c718ffcc8fb4aa71384aebf1bc237fbbf72901ae7ce6527492a6bedc

Observation 650ede2f-1a03-4907-877f-2b298c111e92 · inbound

Forget, Anticipate and Adapt: Test Time Training for Long Videos cites this paper.

Forget, Anticipate and Adapt: Test Time Training for Long Videos Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-14T17:16:44.626166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:16:44.626166Z digest=sha256:fa045ff1a597533495e6900acc49f77401a7e888404956880c8876fa1679fe1d

Observation 6d9e5e56-ee7e-45e5-b2d3-50e5bfad8628 · inbound

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation cites this paper.

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation Learning Temporally Consistent Video Depth from Video Diffusion Priors

Reference 178

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:44.834806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:23:44.834806Z digest=sha256:df5419d607889801fc0626ce45fd418ca58f85915ec21b4c6f52bf7259729152