Pith. sign in

Paper Citation Record · LEDGER

AR4D: Autoregressive 4D Generation from Monocular Videos

As of 13 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 4 inbound Pith citation observations for arXiv:2501.01722.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.01722 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:25:55.528521Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T15:09:46.865981Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

58 of 58 outbound references displayed

  • verified exact3
  • verified fuzzy15
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ac187ac8-0c59-437b-bb4d-56bba35e2ab7 · outbound

This paper cites Hyperreel: High-fidelity 6-dof video with ray- conditioned sampling.

AR4D: Autoregressive 4D Generation from Monocular Videos Hyperreel: High-fidelity 6-dof video with ray- conditioned sampling

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.402048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.221308Z digest=sha256:8f25fd0ae5c14d59a4bda7ed23c73d1614834a20c781cff667b222577c695f9b

Observation 5668ee21-4407-4e2b-afdf-98b12c1a5987 · outbound

This paper cites 4d-fy: Text-to-4d generation using hybrid score dis- tillation sampling.

AR4D: Autoregressive 4D Generation from Monocular Videos 4d-fy: Text-to-4d generation using hybrid score dis- tillation sampling

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.386121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.226788Z digest=sha256:f6072e7a43fbdc613be141122d540e4485c5122ec556889f49cc455753fa9d3e

Observation 8272e4d1-ad90-4f20-bbb6-59da5e7a5505 · outbound

This paper cites VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control.

AR4D: Autoregressive 4D Generation from Monocular Videos VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.232404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.232404Z digest=sha256:196fb642b9b563d3956434b47e7f2a6d7fbf6235da5ebe4069e12fbdbbe49a83

Observation 9622abdd-3cde-4c9b-b1b7-3c26f9ceab78 · outbound

This paper cites Tc4d: Trajectory-conditioned text-to-4d generation.

AR4D: Autoregressive 4D Generation from Monocular Videos Tc4d: Trajectory-conditioned text-to-4d generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.366161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.237986Z digest=sha256:cf6a0aa0cd7056ee97926d36c00e5b31bf2c3e39904a2bde350f08e70a3e5f9e

Observation 2787f60a-a828-4dae-bfee-a372a42d8947 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

AR4D: Autoregressive 4D Generation from Monocular Videos Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.245383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.245383Z digest=sha256:e3448ad53b24625f53bf35a95c19f0432222b29b477826d6f4983bd07e32d5f4

Observation 3af9f115-3871-4f1d-9799-800dede488f3 · outbound

This paper cites Objaverse: A universe of annotated 3d objects.

AR4D: Autoregressive 4D Generation from Monocular Videos Objaverse: A universe of annotated 3d objects

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.251368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.251368Z digest=sha256:246b131d1066f0863681aa588aa8a5b7a3384bf0110223b1c2ad8e6e7b149aeb

Observation da3e79aa-4d5d-40c5-b017-47c3b42e5456 · outbound

This paper cites GaussianFlow: Splatting Gaussian Dynamics for 4D Content Creation.

AR4D: Autoregressive 4D Generation from Monocular Videos GaussianFlow: Splatting Gaussian Dynamics for 4D Content Creation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.257309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.257309Z digest=sha256:7524e5116f4c9e1fe09b2601c0bd212e033cb29eac821b695d6639c5eac2d251

Observation 55c80ddf-9abb-4c32-a772-a8da4be888ea · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

AR4D: Autoregressive 4D Generation from Monocular Videos CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.262370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.262370Z digest=sha256:1f5d71511c1be135092320ac385299ad9c11d0a0a1c4dbca76f38d22c8540787

Observation f9e11057-ca17-412e-917f-f969b38d1782 · outbound

This paper cites Training-free Camera Control for Video Generation.

AR4D: Autoregressive 4D Generation from Monocular Videos Training-free Camera Control for Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.267718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.267718Z digest=sha256:12b22006d6ce509d7a973b991f0f7b035537018527d2beca718d70c929f4ba62

Observation 827a5666-6b0f-460a-8539-60991774bb10 · outbound

This paper cites Animate3D: Animating Any 3D Model with Multi-view Video Diffusion.

AR4D: Autoregressive 4D Generation from Monocular Videos Animate3D: Animating Any 3D Model with Multi-view Video Diffusion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.273568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.273568Z digest=sha256:1cfece90084c848e0f6864bc77c4511082ecea31fdcb87524aafd09cfb5d4278

Observation 31667c08-6544-418a-8d8e-d53bf588752b · outbound

This paper cites Consistent4d: Consistent 360 {\deg} dynamic object gener- ation from monocular video.

AR4D: Autoregressive 4D Generation from Monocular Videos Consistent4d: Consistent 360 {\deg} dynamic object gener- ation from monocular video

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.337974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.279430Z digest=sha256:a1491f07dca5e7550d630755d2723ac0f6ace69ad6a697756b11ae7e3902d5e9

Observation d9556d85-c66e-49b1-b559-8d3e0f06e65a · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

AR4D: Autoregressive 4D Generation from Monocular Videos 3d gaussian splatting for real-time radiance field rendering

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.283775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.283775Z digest=sha256:7845c77a301922792ff5f984fd11901116c83cd14cc63f931bafcdd7884f06e8

Observation d996f46c-c7a6-4a32-9055-3104d9ce1cc5 · outbound

This paper cites Vivid-ZOO: Multi-View Video Generation with Diffusion Model.

AR4D: Autoregressive 4D Generation from Monocular Videos Vivid-ZOO: Multi-View Video Generation with Diffusion Model

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:25:55.959273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.288588Z digest=sha256:cbca7a6ae6db31c83386f5b5a3074a3f690fa5a2324a233ac83c10f7e4f70aed

Observation 154da7ec-1903-42d8-a3ef-b2fe2cdad257 · outbound

This paper cites DreamMesh4D: Video-to-4D Generation with Sparse-Controlled Gaussian-Mesh Hybrid Representation.

AR4D: Autoregressive 4D Generation from Monocular Videos DreamMesh4D: Video-to-4D Generation with Sparse-Controlled Gaussian-Mesh Hybrid Representation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.293116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.293116Z digest=sha256:85ac32cf087154c94aaacf71fa3f1284e3a5ed865742c3161343af7eac0b84df

Observation ab9267d5-09a9-4e56-a21c-8b140cfbd8c1 · outbound

This paper cites Spacetime gaus- sian feature splatting for real-time dynamic view synthesis.

AR4D: Autoregressive 4D Generation from Monocular Videos Spacetime gaus- sian feature splatting for real-time dynamic view synthesis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.298927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.298927Z digest=sha256:e7dd89ef2233a07fb2c2ef1a7f59d736250b849fb49b11723757b8e94e399134

Observation 81ba0f20-589f-4120-840c-fb4394ce645d · outbound

This paper cites Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models.

AR4D: Autoregressive 4D Generation from Monocular Videos Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.303806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.303806Z digest=sha256:7b6511e9703ada4408c757a5f728d7eb0f948b0a9d1e46af84a4666e60572045

Observation 8708bc44-a8a7-43ee-89e8-4f78f30885d4 · outbound

This paper cites Luciddreamer: Towards high- fidelity text-to-3d generation via interval score matching.

AR4D: Autoregressive 4D Generation from Monocular Videos Luciddreamer: Towards high- fidelity text-to-3d generation via interval score matching

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.299890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.308554Z digest=sha256:7878986e68f4b14ced8ea528004620e37fb35e261b2fb243a7421ece5d8e44c8

Observation 8ad5c261-7b40-400c-bc30-20c829f95cb6 · outbound

This paper cites Align your gaussians: Text-to-4d with dynamic 3d gaussians and composed diffusion models.

AR4D: Autoregressive 4D Generation from Monocular Videos Align your gaussians: Text-to-4d with dynamic 3d gaussians and composed diffusion models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.284937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.313307Z digest=sha256:9b98abd3f7a1e76e2a95bac757b45dfb37ff12924b4cc0b2810b4c5ead793a8a

Observation 8a09c225-3808-4951-b8b1-a1f30bbf60f9 · outbound

This paper cites One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimiza- tion.

AR4D: Autoregressive 4D Generation from Monocular Videos One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimiza- tion

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.270429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.319223Z digest=sha256:c525185b71aafddf1651b39ef8cb8973da201600bd4a0e050ef08a9e115eef28

Observation 607ed7cb-cd97-4fc2-b02f-be95dcc8719c · outbound

This paper cites Zero-1-to- 3: Zero-shot one image to 3d object.

AR4D: Autoregressive 4D Generation from Monocular Videos Zero-1-to- 3: Zero-shot one image to 3d object

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.255827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.325133Z digest=sha256:4ed91c4bfbfa932c3bc3648e122c791149f615e1d0d56cd038b5d3a291ebd900

Observation 2eb9b73c-69b7-4522-8072-94dc79fbb68c · outbound

This paper cites SyncDreamer: Generating Multiview-consistent Images from a Single-view Image.

AR4D: Autoregressive 4D Generation from Monocular Videos SyncDreamer: Generating Multiview-consistent Images from a Single-view Image

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.330274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.330274Z digest=sha256:4a87fd4d8de51861597c42d0432a19ca0c85533013b202069f728e68d8ed00b6

Observation 01a9e790-e709-4424-bea7-d3a5524f42b3 · outbound

This paper cites Wonder3d: Sin- gle image to 3d using cross-domain diffusion.

AR4D: Autoregressive 4D Generation from Monocular Videos Wonder3d: Sin- gle image to 3d using cross-domain diffusion

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.335153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.335153Z digest=sha256:824ddebc4430ded04760f69512c6f434dc09acb7033dde5d4718ddf0709215a6

Observation a092987f-c37f-4eda-bd13-4afa7011ec56 · outbound

This paper cites PLA4D: Pixel-Level Alignments for Text-to-4D Gaussian Splatting.

AR4D: Autoregressive 4D Generation from Monocular Videos PLA4D: Pixel-Level Alignments for Text-to-4D Gaussian Splatting

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:25:55.897202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.339931Z digest=sha256:18e5b971fc1881fb30ca680c4604d059412bf42128f1cc06853efb43676b0449

Observation b29c3173-dab6-4ac9-ba7b-98086e5f65fe · outbound

This paper cites Nerf: Representing scenes as neural radiance fields for view syn- thesis.

AR4D: Autoregressive 4D Generation from Monocular Videos Nerf: Representing scenes as neural radiance fields for view syn- thesis

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.344632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.344632Z digest=sha256:f1bdacc3b402435c529d8d889432e478fd922dcdaef0bb76b4e0aca09ee42a36

Observation a027ce41-23eb-4d84-a093-58bf34af78e0 · outbound

This paper cites T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models.

AR4D: Autoregressive 4D Generation from Monocular Videos T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.350029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.350029Z digest=sha256:6b3b07528b20f22983ed503b894fb3f50b319adf5fa62ac63138e1431e2ef055

Observation 06542d68-efd5-4852-be6c-7621a1adcf6e · outbound

This paper cites Reg- nerf: Regularizing neural radiance fields for view synthesis from sparse inputs.

AR4D: Autoregressive 4D Generation from Monocular Videos Reg- nerf: Regularizing neural radiance fields for view synthesis from sparse inputs

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.210867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.358751Z digest=sha256:dd416d0bdc9e9c10ab3312c379505d138a0d5422b9918834aa8ed748acb6762e

Observation e71f31b3-e8c7-4c7a-b8b8-eb24b4f1aa5e · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

AR4D: Autoregressive 4D Generation from Monocular Videos SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.365367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.365367Z digest=sha256:ea8c5fd4da78374626b48fc9b18ffe926b6aa9182610cb0636be110be4c9c295

Observation 32acf034-f4b6-4149-84fe-8cf4e662266c · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

AR4D: Autoregressive 4D Generation from Monocular Videos DreamFusion: Text-to-3D using 2D Diffusion

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.370427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.370427Z digest=sha256:b4dbc1639c95dad25f3532c6e03a2d4e5bdeebec14db7b4cffe13ce4749d0d66

Observation 9bac6172-8bfe-4310-95ba-2e5514d9b2f2 · outbound

This paper cites D-nerf: Neural radiance fields for dynamic scenes.

AR4D: Autoregressive 4D Generation from Monocular Videos D-nerf: Neural radiance fields for dynamic scenes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.378312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.378312Z digest=sha256:d9867baf443c4cbe07306edbe610db997a79837b76882b2b4e68c6880e5d0d5b

Observation e4e3e470-db13-4f03-8b02-211f21f57c03 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

AR4D: Autoregressive 4D Generation from Monocular Videos Learning transferable visual models from natural language supervi- sion

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.383235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.383235Z digest=sha256:33c9d25a99811fb7c02f86f435d9f2e82ebee4200391221ce29d884fca36c2cc

Observation 24f68d4d-3e9d-4adf-98e3-80e016609e89 · outbound

This paper cites DreamGaussian4D: Generative 4D Gaussian Splatting.

AR4D: Autoregressive 4D Generation from Monocular Videos DreamGaussian4D: Generative 4D Gaussian Splatting

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.387765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.387765Z digest=sha256:cff163aee02bbd75f990795d55734ff8766e849ed6f487ab3ccd4d6181b66b51

Observation 2192cef0-58a5-4c88-b22b-180ea6d7445f · outbound

This paper cites L4GM: Large 4D Gaussian Reconstruction Model.

AR4D: Autoregressive 4D Generation from Monocular Videos L4GM: Large 4D Gaussian Reconstruction Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.392502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.392502Z digest=sha256:1353a1be21125a72dee169e337bd5bf47a5566c50414fe93daa73e00131601b5

Observation c935efe5-5518-49a7-bb9b-ef5d4d9c9963 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

AR4D: Autoregressive 4D Generation from Monocular Videos High-resolution image synthesis with latent diffusion models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.397401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.397401Z digest=sha256:3d37bec8b7b03c3c318cd7d46a2419918483bf42b65caa27599a5f663de48f50

Observation 9de1dca4-743b-413f-86cd-fe95a9b09178 · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

AR4D: Autoregressive 4D Generation from Monocular Videos MVDream: Multi-view Diffusion for 3D Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.401702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.401702Z digest=sha256:e55efa3de52bc68656b2b44d5ff9b6579fb599535f51511a2d2aa0c824fdd32b

Observation 7281cb85-7659-401e-8391-ac8be9d0c6c2 · outbound

This paper cites Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs.

AR4D: Autoregressive 4D Generation from Monocular Videos Splatt3R: Zero-shot Gaussian Splatting from Uncalibrated Image Pairs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.406820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.406820Z digest=sha256:781beccb08d67db4ae3a727d1a3c48d3ed93ece111e719c60a025008234a10fe

Observation 5a1971f8-bf58-4842-bbe4-af3f0fa6b885 · outbound

This paper cites EG4D: Explicit Generation of 4D Object without Score Distillation.

AR4D: Autoregressive 4D Generation from Monocular Videos EG4D: Explicit Generation of 4D Object without Score Distillation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.411919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.411919Z digest=sha256:d207f3b4b83de2b8e607f43c8fca5497d9db1777b2f417f085f909f2998f954d

Observation c1d1395b-a8e8-4249-aa95-cedf4e89bcd4 · outbound

This paper cites DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation.

AR4D: Autoregressive 4D Generation from Monocular Videos DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.417719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.417719Z digest=sha256:e22a46f4af6d97072abcb5fa8b7fc31f2627e936e226674c44bb628570a01463

Observation c8ad3180-a0eb-4c7e-bc1f-86afecabe852 · outbound

This paper cites Lgm: Large multi-view gaussian model for high-resolution 3d content creation.

AR4D: Autoregressive 4D Generation from Monocular Videos Lgm: Large multi-view gaussian model for high-resolution 3d content creation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.171089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.423990Z digest=sha256:42e70ab0f826f46b8da47c508c8205011097cfed652450c76e91b3ac8b7b093d

Observation a9d45ee2-fd62-45a9-aa07-1542e349dfd0 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

AR4D: Autoregressive 4D Generation from Monocular Videos Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.428902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.428902Z digest=sha256:287e484bd75446d40a717052862106c8c3c3ced330c0f6ff3988c844ac111398

Observation 23314bc1-9bbb-4e2d-8f68-132c9ad853e9 · outbound

This paper cites Phenaki: Variable length video generation from open domain textual descriptions.

AR4D: Autoregressive 4D Generation from Monocular Videos Phenaki: Variable length video generation from open domain textual descriptions

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.156098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.434719Z digest=sha256:6f69baa7905dc2b0d482f5d4ec28c1af8630b8ab9113e1afb40eace7fe6ec53d

Observation 54b479a3-ed4b-4400-bddb-7d6fe4f4709f · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.

AR4D: Autoregressive 4D Generation from Monocular Videos Image quality assessment: from error visibility to structural similarity

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.440063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.440063Z digest=sha256:9f8b2bf425f521b87ac529481762007ca70c3bf8dbff3475eb9417863850333c

Observation 1187e5f3-9da7-49a8-b0c4-483f1e089c1e · outbound

This paper cites Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distilla- tion.

AR4D: Autoregressive 4D Generation from Monocular Videos Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distilla- tion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.131782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.445048Z digest=sha256:13c239edad6cc0fc3cb8d019f63ec241ff03791ac291153138dd818f655c0a58

Observation 67b6152b-fbec-473a-8c74-289af210672b · outbound

This paper cites 4d gaussian splatting for real-time dynamic scene rendering.

AR4D: Autoregressive 4D Generation from Monocular Videos 4d gaussian splatting for real-time dynamic scene rendering

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.451389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.451389Z digest=sha256:add60cbfd89558c97ce6dfcb7d07fe07aa4f6dc605d17b07a58ed4369be284cf

Observation d134d139-da09-44ac-a72b-5cceb0a0cdf2 · outbound

This paper cites Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation.

AR4D: Autoregressive 4D Generation from Monocular Videos Tune-a-video: One-shot tuning of image diffusion models for text-to-video generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.456918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.456918Z digest=sha256:5f7685f698318a383729065da42a6602ac1856446c72d6176dae57f5f4b31720

Observation a269dc50-4099-4822-a701-bf4b8dad37ca · outbound

This paper cites SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency.

AR4D: Autoregressive 4D Generation from Monocular Videos SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.461513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.461513Z digest=sha256:8d1267879a0b8c3b4a9fc7f4cfe35b49726ad31f6463e51b807d77c5d28bf9c8

Observation f85db101-c269-4623-9ca1-393c4d9f759e · outbound

This paper cites CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation.

AR4D: Autoregressive 4D Generation from Monocular Videos CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.466525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.466525Z digest=sha256:79c8f29547ccde3ef2296c96262782f92bee5c35be9af5193c082998d6ed633d

Observation 13cb0d96-84bc-49f3-bd74-1524ef211d13 · outbound

This paper cites Deformable 3d gaussians for high- fidelity monocular dynamic scene reconstruction.

AR4D: Autoregressive 4D Generation from Monocular Videos Deformable 3d gaussians for high- fidelity monocular dynamic scene reconstruction

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.097702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.471069Z digest=sha256:94650c13030cc4d05b2e0069efaf27899b42321148ff38ff53de739b4624fe5f

Observation 2dd2a1b0-7d35-4cea-ae55-d6c32821eba2 · outbound

This paper cites Diffusion$^2$: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models.

AR4D: Autoregressive 4D Generation from Monocular Videos Diffusion$^2$: Dynamic 3D Content Generation via Score Composition of Video and Multi-view Diffusion Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.475396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.475396Z digest=sha256:a8d40e417be2220afe3f81957b051bd8db35c1ca6994a6bf1421f00692d108e6

Observation eaaccaba-424d-4c16-9fd8-74cc01ed5c8f · outbound

This paper cites GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models.

AR4D: Autoregressive 4D Generation from Monocular Videos GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.479899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.479899Z digest=sha256:4501aae6256643a699d601d39548351dd742a024c707b7db31a5f3876cbc0ee5

Observation 04485df3-271c-43aa-a162-155642fba33a · outbound

This paper cites ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis.

AR4D: Autoregressive 4D Generation from Monocular Videos ViewCrafter: Taming Video Diffusion Models for High-fidelity Novel View Synthesis

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.486085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.486085Z digest=sha256:1cd04ddccaa7a27ebefe711fa8c6019a817dcd0ed12cd1cf83ee56fc659b8932

Observation 63ef4725-28c2-42a9-9da5-84a8d58a47ae · outbound

This paper cites 4Dynamic: Text-to-4D Generation with Hybrid Priors.

AR4D: Autoregressive 4D Generation from Monocular Videos 4Dynamic: Text-to-4D Generation with Hybrid Priors

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-10T22:25:55.643470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.491407Z digest=sha256:1cdc53fab168ff6de48bc40d88ee805fd4cd77b31bc63a98d1b998543931cde6

Observation d0df6088-3290-48cc-b31b-85f4605544e9 · outbound

This paper cites Stag4d: Spatial-temporal anchored generative 4d gaussians.

AR4D: Autoregressive 4D Generation from Monocular Videos Stag4d: Spatial-temporal anchored generative 4d gaussians

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.083486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.497269Z digest=sha256:ea669976fcd2c803a3501ddcec53477eadd3f1e39c2d197df7b1a39992bec9ae

Observation 9490a453-7b97-49d0-90ff-7c63f3374e69 · outbound

This paper cites Show-1: Marrying pixel and latent diffusion models for text-to-video generation.

AR4D: Autoregressive 4D Generation from Monocular Videos Show-1: Marrying pixel and latent diffusion models for text-to-video generation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T22:25:56.068472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T22:25:55.501958Z digest=sha256:c4d0ada7900b293b2c3e2fdd1caf41fee3d344bb73618ce79395de9b478db837

Observation a531618b-919b-4f8d-8aba-73e16874d421 · outbound

This paper cites 4Diffusion: Multi-view Video Diffusion Model for 4D Generation.

AR4D: Autoregressive 4D Generation from Monocular Videos 4Diffusion: Multi-view Video Diffusion Model for 4D Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.508339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.508339Z digest=sha256:b5fb2f798bf9cb00c83068ba972f24577e98916f453d7a086ea1f05c7d752104

Observation bc072b74-685f-4d30-98e7-f48cf82ac925 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

AR4D: Autoregressive 4D Generation from Monocular Videos Adding conditional control to text-to-image diffusion models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.513275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.513275Z digest=sha256:2274867fb59c27367fe52119880775396d69a61c334112e594d34cac5f43a60b

Observation 7aca5b56-31ac-4d70-a825-080940d5ee54 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

AR4D: Autoregressive 4D Generation from Monocular Videos The unreasonable effectiveness of deep features as a perceptual metric

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.518142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.518142Z digest=sha256:0cea754ba883b8b8d16e42b87481e6eef8eb5ded9f3a2b983e35279002d7a990

Observation 6c952e3c-1df4-426e-9ceb-1d83293a8fba · outbound

This paper cites Animate124: Animating One Image to 4D Dynamic Scene.

AR4D: Autoregressive 4D Generation from Monocular Videos Animate124: Animating One Image to 4D Dynamic Scene

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.523118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.523118Z digest=sha256:647584cb9682317965d45a2d97cc21383e6a71dc921db1563a0514118bf70ee1

Observation 93ef31c7-24a3-41c7-af81-2ee5bbba9239 · outbound

This paper cites Compositional 3D-aware Video Generation with LLM Director.

AR4D: Autoregressive 4D Generation from Monocular Videos Compositional 3D-aware Video Generation with LLM Director

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-10T22:25:55.528521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:25:55.528521Z digest=sha256:b7711fe0b32856c19a0d15950c54047f0afb8d67c703b71bbc73d78d1527f24d

Pith citing papers

Observation ce779607-71f2-4381-bc57-11d7b348aa6e · inbound

CP4D: Compositional Physics-aware 4D Scene Generation cites this paper.

CP4D: Compositional Physics-aware 4D Scene Generation AR4D: Autoregressive 4D Generation from Monocular Videos

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:57:30.477209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-27T16:52:36.288955Z digest=sha256:745d5181cabc7b9f70bd184014fb69b34562e3801cd61dc6419f41d5e90ff096

Observation e4ae32d7-b2aa-495f-a6c0-f8768e9309c8 · inbound

Feed-forward Motion In-betweening for Any 4D cites this paper.

Feed-forward Motion In-betweening for Any 4D AR4D: Autoregressive 4D Generation from Monocular Videos

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-26T12:29:28.059344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-26T12:21:01.979010Z digest=sha256:a948fc854efb1969917d2b59f51efb1203d57b6e4518615f04e233fb6a01f5b7

Observation fb3e193e-4a57-4d2d-bc6f-8c0a73b26a8e · inbound

Follow Your Track: Precise Skeleton Animation Controlled by 3D Trajectories cites this paper.

Follow Your Track: Precise Skeleton Animation Controlled by 3D Trajectories AR4D: Autoregressive 4D Generation from Monocular Videos

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-04T19:40:06.062435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-06-25T21:12:03.751682Z digest=sha256:63b5baddfc63f5ee3a0886982f26b85b597268ec95f490aa7163c172200e7a2a

Observation 0aba1c3d-f863-4c26-bdbf-b691107be888 · inbound

AniGS: Bridging Rendering and Diffusion Prior for 3D Scene Animation cites this paper.

AniGS: Bridging Rendering and Diffusion Prior for 3D Scene Animation AR4D: Autoregressive 4D Generation from Monocular Videos

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T15:09:46.865981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:09:46.865981Z digest=sha256:4c7175bdfef1f8f877e181b1f618b95225f73980c92c6a4f516837f755f86ba7