Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T00:59:22.654750Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 71 of 71 outbound references and 5 inbound Pith citation observations for arXiv:2412.01821.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T00:59:22.654750Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T19:13:40.201882Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T14:24:45.166579Z
71 of 71 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6a40f213-86c6-4610-9a16-fcefa4de87c1 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Lumiere: A Space-Time Diffusion Model for Video Generation
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec2399cf-bf68-446c-bea6-32c26bcd16e6 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbe03498-7435-489d-b3f5-1430a889cc3d · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Video generation models as world simulators
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fab3a95d-971e-4805-a868-738577e5d40e · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Video generation models as world simulators
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 91753383-d7f3-46ac-863f-be0700137d14 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Efficient Geometry-aware 3D Generative Adversarial Networks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 896ce24f-a253-4f15-ad1d-ff0351e98e21 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf8d4779-4dd3-4e05-a466-9b0d20963b5d · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Chang, Manolis Savva, Maciej Hal- ber, Thomas Funkhouser, and Matthias Nießner
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09f1c202-05df-4126-be9f-c8a7f318b5d8 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Depth-supervised NeRF: Fewer views and faster training for free
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48c0f20a-6148-402b-b669-3c37d268ed8f · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ce86b0c-b623-41a4-965a-34d5a9d90557 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography.Communications of the ACM, 24(6):381–395, 1981
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3a0d97d-87c2-4bf8-9b98-29e89f28876d · outbound
World-consistent Video Diffusion with Explicit 3D Modeling CAT3D: Create Anything in 3D with Multi-View Diffusion Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b9ddeb9-6603-4097-b682-afc21f3c449a · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33ba5c2d-344d-46af-87e2-3dfb71dac17e · outbound
World-consistent Video Diffusion with Explicit 3D Modeling NerfDiff: Single-image View Synthesis with NeRF-guided Distillation from 3D-aware Diffusion
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 127b668d-3215-4bf9-a42f-b31e3b6ba2ce · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Control3diff: Learning con- trollable 3d diffusion models from single-view images
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6129322b-e9b5-4db2-b582-edc1e173180e · outbound
World-consistent Video Diffusion with Explicit 3D Modeling AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9293ea54-ca52-4457-99c4-1baaa74ff268 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling CameraCtrl: Enabling Camera Control for Text-to-Video Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdc74aa5-cd08-4d1b-b18c-565979ef928b · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Gans trained by a two time-scale update rule converge to a local nash equilib- rium
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 213824c2-1b9d-41e4-bf1d-709406f286fe · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Denoising diffu- sion probabilistic models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fc87e58d-6030-4525-8c72-1f10833d7c8f · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Imagen Video: High Definition Video Generation with Diffusion Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e20f642-fd3a-413b-9de6-a6e9e6bd37b9 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Video diffu- sion models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1ce77cb0-62c5-4408-87cd-869a604a0670 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Viewdiff: 3d-consistent image generation with text-to-image models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 139530ad-7517-4da4-b89f-f6097094dd90 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling LRM: Large Reconstruction Model for Single Image to 3D
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 941151d8-815f-48da-b669-4bc0a882d662 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Vbench: Comprehensive bench- mark suite for video generative models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 83c05cd4-c6f9-4bdf-b616-f21ac42b4b0c · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Auto-Encoding Variational Bayes
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d47135f-b931-49f2-a3f3-a368d7ee94aa · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Ground- ing image matching in 3d with mast3r, 2024
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81ad3779-aee3-46c9-9837-aaa5a2389d71 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Zero-1-to-3: Zero-shot one image to 3d object, 2023
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d9e72d13-c0cf-4c82-8869-91ed41236709 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling SyncDreamer: Generating Multiview-consistent Images from a Single-view Image
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56ffc5a0-648a-4fe1-b7e6-1cad262bdabf · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Repaint: Inpainting using denoising diffusion probabilistic models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b8316ea5-4ac2-4a8a-a23d-fa48951d0e08 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 842ad03b-bcd3-42e1-af5c-759d19c896e2 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Gnerf: Gan-based neural radiance field without posed camera
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c34951db-2055-44b7-a6f4-c53e941666e8 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f2bbfd2-a178-4401-a5ae-c55148462ac9 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling GIRAFFE: Repre- senting Scenes as Compositional Generative Neural Feature Fields
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b6749ee4-a787-4887-8be1-d39feb144952 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Refusion: 3d reconstruction in dynamic environments for rgb-d cameras exploiting resid- uals
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f79a3bf5-b65a-48e1-912b-d0c8d69302c4 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Scalable diffusion models with transformers
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66ed9a93-edac-4d6d-a062-3f1df3dc3850 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5516aaa0-231c-459f-96b7-488c099b6de8 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0cb21094-00d2-4a5f-9dbd-aba91cea4c70 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Com- mon objects in 3d: Large-scale learning and evaluation of real-life 3d category reconstruction
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2324abcb-fa68-412a-8008-f1ff9fb805e7 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling High-resolution image synthesis with latent diffusion models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a5b5715-0424-4914-aab4-0e61e2eb4f27 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling U- net: Convolutional networks for biomedical image segmen- tation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation be327aac-0eed-4f3a-a5d1-66ba79f8d094 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Habitat: A Platform for Embodied AI Research
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 81961f21-0ae7-47d3-891b-d5d4930ab393 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Pixelwise view selection for un- structured multi-view stereo
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aecc5aaa-856b-4946-b8a9-2bb092dd7153 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling A benchmark and a baseline for robust multi- view depth estimation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6bb927f5-ece4-4729-ac5a-4e578824560a · outbound
World-consistent Video Diffusion with Explicit 3D Modeling MVDream: Multi-view Diffusion for 3D Generation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef8665a9-f7bd-49ed-a54d-fe81c311333f · outbound
World-consistent Video Diffusion with Explicit 3D Modeling 3D Photography using Context-aware Layered Depth Inpainting
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0528201c-c770-4638-9f68-57fac0fd99d8 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Indoor Segmentation and Support Inference from RGBD Images
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bdc6617a-d761-47d0-af65-b83021125512 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Light Field Networks: Neural Scene Representations with Single-Evaluation Rendering
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b5cc56a-e1d2-41d2-8faf-2c1fccb37afa · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Deep unsupervised learning using nonequilibrium thermodynamics
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f8384e1-d6ba-4514-97d3-32512b8c2c1e · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Score-Based Generative Modeling through Stochastic Differential Equations
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91e4a68c-ba22-4fa2-853a-4fcd5f795813 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Kick back & relax: Learning to reconstruct the 10 world by watching slowtv
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e16ca5f5-1ff3-41f5-a2eb-6735857dcfb7 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Roformer: Enhanced transformer with rotary position embedding
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4ae820-3139-46f9-8e1c-962b73805ebc · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Loftr: Detector-free local feature matching with transformers
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f84f4506-1b4a-4e58-bec9-4400675d9a8c · outbound
World-consistent Video Diffusion with Explicit 3D Modeling DeepV2D: Video to Depth with Differentiable Structure from Motion
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d6f419-23ee-47d1-8935-dc60abc0fecc · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Demon: Depth and motion network for learning monocular stereo
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f2f7b85-b8c7-4f33-a37f-8b99dade6290 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Vggsfm: Visual geometry grounded deep structure from motion
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4d7b6a2-affa-46d6-a743-d41ea0286182 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling ImageDream: Image-Prompt Multi-view Diffusion for 3D Generation
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53697f09-d39e-4936-a04c-3a56b2206e0e · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Dust3r: Geometric 3d vi- sion made easy
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 95373792-f642-45a2-b23e-34fa89119db4 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23ddbe7d-ff02-4306-af1f-11bdc1c80743 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Driving into the future: Multiview visual forecasting and planning with world model for au- tonomous driving
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ead9686-bdc8-4f0f-a095-51fe27d2ec71 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Motionctrl: A unified and flexible motion controller for video generation
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6b6e4247-3dc0-4900-a3e0-b44f7692a8ee · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Controlling Space and Time with Diffusion Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c908c1a-4947-4128-92f2-bd648979eaf3 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65d8e9a2-2035-4030-8087-9a677ee8cabb · outbound
World-consistent Video Diffusion with Explicit 3D Modeling 3d-aware image synthesis via learning struc- tural and textural representations
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b2ca9657-cd1f-49fd-98f7-9271d716fd7a · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Consistnet: Enforcing 3d consistency for multi- view images diffusion
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bf1f76fb-be3f-45b2-af73-72c608bb8ff6 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Mvs2d: Efficient multi-view stereo via attention-driven 2d convolutions
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 7c339a15-5a63-4155-94c9-06ed48a4f344 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Mvsnet: Depth inference for unstructured multi-view stereo
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6a942ca6-8fe8-4d4c-a3c9-b796c443d308 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Scannet++: A high-fidelity dataset of 3d in- door scenes
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fa8ad4e5-f085-49ee-bf5c-99c6bdd09f9a · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Mvimgnet: A large-scale dataset of multi-view images
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9d5b0382-f830-46bb-aa85-ccbd429ea3f8 · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Root mean square layer nor- malization
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe658124-dec4-432a-a465-0d6148dbd5cc · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Vis-mvsnet: Visibility-aware multi-view stereo net- work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 696b4bcc-096e-48c7-a4b4-873a5b789aea · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Stereo magnification: Learning view synthesis using multiplane images
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 45a56b71-fdc1-4570-b2f9-d437619c2f6b · outbound
World-consistent Video Diffusion with Explicit 3D Modeling Stereo magnification: Learning view syn- thesis using multiplane images
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation bab06ec1-95da-4441-aae9-cdc3817c1259 · inbound
Emergent Temporal Correspondences from Video Diffusion Transformers World-consistent Video Diffusion with Explicit 3D Modeling
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff09493-8686-47ca-9a6d-86ba1e844feb · inbound
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations World-consistent Video Diffusion with Explicit 3D Modeling
Reference 132
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b0781bd0-2d4a-41bb-aa5e-314da47346e8 · inbound
SeqTex: Generate Mesh Textures in Video Sequence World-consistent Video Diffusion with Explicit 3D Modeling
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4501095a-4b6d-4bae-b511-54d43a1e959f · inbound
Epipolar Geometry Improves Video Generation Models World-consistent Video Diffusion with Explicit 3D Modeling
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 402390c6-7c05-4c0b-a9cf-af707026ac1a · inbound
Unified 3D Scene Understanding Through Physical World Modeling World-consistent Video Diffusion with Explicit 3D Modeling
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.