Pith. sign in

Paper Citation Record · LEDGER

DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 58 inbound Pith citation observations for arXiv:2411.04928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.04928 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 58 of 58 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:28:23.752238Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:33:54.345685Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 47a150ca-52be-4477-a579-866171bc86c2 · inbound

CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models cites this paper.

CAT4D: Create Anything in 4D with Multi-View Video Diffusion Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-12T11:05:32.779919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:05:32.779919Z digest=sha256:92718300048e431b4c28c09e152ebacf9211b696ba250300bedeea9768331fde

Observation a9eafcef-b67e-4e3e-b2d9-c4ab65a9e864 · inbound

AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers cites this paper.

AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 129

Resolution
unresolved
no resolver link, observed 2026-08-12T11:07:39.864035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:07:39.864035Z digest=sha256:f85858dcec2aa70dabb871be402913a3bfa138a1a5344b53784c5c6854aad472

Observation dfae9cf9-d5e9-4399-ab64-fa3235f7e20e · inbound

You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale cites this paper.

You See it, You Got it: Learning 3D Creation on Pose-Free Videos at Scale DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T19:26:14.261266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:26:14.261266Z digest=sha256:00e179d2d776048121571ccbc7027585e614473f0d1c57e750fc8d9db84abc8e

Observation 403b6900-5b9e-4159-99e5-146bd1de6af6 · inbound

Grid: Omni Visual Generation cites this paper.

Grid: Omni Visual Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T15:44:33.624817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:44:33.624817Z digest=sha256:8bb1d7ef88d23da27073d3db6a18ffb78e25e3d680e03b258cea26d7e9d20e76

Observation b8b51882-b91b-42b5-b81a-22f5cf6057c5 · inbound

Wonderland: Navigating 3D Scenes from a Single Image cites this paper.

Wonderland: Navigating 3D Scenes from a Single Image DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T14:21:06.406092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:21:06.406092Z digest=sha256:e029d05a0ce2c4e1bf3f60654119d62f677940617df07095ac268a9b3ac2290e

Observation 7a2e3620-c47a-400b-ad92-88cef75d2d44 · inbound

Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation cites this paper.

Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T23:07:40.841180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:07:40.841180Z digest=sha256:1279e2f391afa7c49d91262571cd33b1fac4efc2920ba5cd4a2ad0f0e402b252

Observation 1ef1effa-716c-44e0-ab8a-30878f0c4c84 · inbound

Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation cites this paper.

Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T21:25:45.871868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:25:45.871868Z digest=sha256:493256f34edee9d87e2e38780816cd4771f4aca26658fa99f5ff5fbc2ca55cfb

Observation c9437dd1-38f6-49c4-959b-eb5ca6e7c7da · inbound

VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation cites this paper.

VidCRAFT3: Camera, Object, and Lighting Control for Image-to-Video Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T12:31:41.884338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:31:41.884338Z digest=sha256:293ea7a580b6e3c4291f121b02c6717d1642bbc06529167dc366feb3ccea30fb

Observation c4413c67-57f3-4af6-bc11-06ae244906c9 · inbound

SOPHY: Learning to Generate Simulation-Ready Objects with Physical Materials cites this paper.

SOPHY: Learning to Generate Simulation-Ready Objects with Physical Materials DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-16T12:28:23.752238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:28:23.752238Z digest=sha256:5a3d0669faa8f5bb23aaf1802d4682df5873c318ca7bd2901ae9b7af122f5701

Observation 708b1fa3-35e0-4d2d-b3b2-b12507299b67 · inbound

TwoSquared: 4D Generation from 2D Image Pairs cites this paper.

TwoSquared: 4D Generation from 2D Image Pairs DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T12:26:06.293702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:26:06.293702Z digest=sha256:c7e029b1e1b2c1745442f954c5c9b4ee332a71e6d07c2b97640f859e6c187de5

Observation 94f3efe0-da26-48e4-bc42-a491e4ffa6af · inbound

3D Scene Generation: A Survey cites this paper.

3D Scene Generation: A Survey DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:06:02.109992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:06:02.109992Z digest=sha256:d30c30efa4e6f1a809b1e639c4bbf8a9953f935bb8623ab7a8c5060749a52db9

Observation 7d4260a0-4068-4be0-8333-9923f955c174 · inbound

SpatialCrafter: Unleashing the Imagination of Video Diffusion Models for Scene Reconstruction from Limited Observations cites this paper.

SpatialCrafter: Unleashing the Imagination of Video Diffusion Models for Scene Reconstruction from Limited Observations DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T20:47:21.180298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:47:21.180298Z digest=sha256:6cc2746b0917240bbeadb7784cc2107c6682658fde8acae17aa6021527258703

Observation 134783af-078d-4b96-8ab5-119e72e61b7f · inbound

EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance cites this paper.

EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:28:13.250744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:28:13.250744Z digest=sha256:facba6f8bf37664d25fa9a9db773ff49ae8a19dcacaec25d8ab836a028eb6f1e

Observation 9c10fc28-52dc-4b6b-beb7-1d73681d3754 · inbound

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting cites this paper.

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T10:43:24.320728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:43:24.320728Z digest=sha256:7a93937aebac953000a1c2e5684149c11fd3329df52ef0a2dfb8ed650706ce1b

Observation aa8f1fe9-27d9-4348-a6a6-401f47abb153 · inbound

Emergent Temporal Correspondences from Video Diffusion Transformers cites this paper.

Emergent Temporal Correspondences from Video Diffusion Transformers DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T19:13:39.939649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:13:39.939649Z digest=sha256:e29a3546ceacf5067809ab28dd1be74d890a71fc32e84a35d138ab827ac89926

Observation c8baa711-dca9-4cf2-9990-043c2653c678 · inbound

BulletGen: Improving 4D Reconstruction with Bullet-Time Generation cites this paper.

BulletGen: Improving 4D Reconstruction with Bullet-Time Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:03:01.137403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-19T08:02:14.123151Z digest=sha256:374e662e1e1d7c966313fffd84806452c93a5f5db723e329ba953bf572a5886d

Observation 24d0e384-c26e-4068-b69c-73d154ab8a68 · inbound

From Virtual Games to Real-World Play cites this paper.

From Virtual Games to Real-World Play DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T18:47:22.962575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:47:22.962575Z digest=sha256:fbe2ca6b8ebc681c733356616e8acaa586b33aaded33dd21fd4d717573d4c378

Observation 2391828b-e872-4545-a342-40bad334e217 · inbound

LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion cites this paper.

LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:26:10.665816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:26:10.665816Z digest=sha256:e3d890d1c792d633c471283ddf9b03c692586539045510fed540eadebe5a1d8d

Observation c61477fe-6caf-4f65-ac18-c40b470665cd · inbound

Voyaging into Perpetual Dynamic Scenes from a Single View cites this paper.

Voyaging into Perpetual Dynamic Scenes from a Single View DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:12.284812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:12.284812Z digest=sha256:79c28e896451aed7b6f7dd58afa6f6a59114edad784d0ba6d76d29b180b3ce93

Observation aa499822-13ce-48d2-bfee-33fa71b14ed0 · inbound

LiON-LoRA: Rethinking LoRA Fusion to Unify Controllable Spatial and Temporal Generation for Video Diffusion cites this paper.

LiON-LoRA: Rethinking LoRA Fusion to Unify Controllable Spatial and Temporal Generation for Video Diffusion DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T19:25:28.485901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:25:28.485901Z digest=sha256:8737d5c2a3336d2f36941e1b930714c06c45efe1b7733593a605138447ccc6be

Observation 67a0346e-d221-4e78-bf57-1aae38abb2cd · inbound

HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels cites this paper.

HunyuanWorld 1.0: Generating Immersive, Explorable, and Interactive 3D Worlds from Words or Pixels DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T12:26:02.174516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:26:02.174516Z digest=sha256:be7a1c694b67850030eb6ba907781f3b358f5faa6d617343b6da9f11b43c2180

Observation 3fab5e38-1866-4cc7-bf1b-db2e3e1dc520 · inbound

4DVD: Cascaded Dense-view Video Diffusion Model for High-quality 4D Content Generation cites this paper.

4DVD: Cascaded Dense-view Video Diffusion Model for High-quality 4D Content Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T00:02:34.713208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:02:34.713208Z digest=sha256:318f7e60f8c8cde34da6e84c1ed103ea0b0ea2b65d79d63b38678653befc281a

Observation 16a285f7-1119-435d-8802-02402e75275d · inbound

Impact-driven Context Filtering For Cross-file Code Completion cites this paper.

Impact-driven Context Filtering For Cross-file Code Completion DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T23:03:18.066450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:03:18.066450Z digest=sha256:5d98a1928d552bbb6167615f7c805aef41d3c80124a55b867a0813210126be19

Observation 1043b0a1-dca6-4aa3-90a7-878a69bd36f9 · inbound

DIP-GS: Deep Image Prior For Gaussian Splatting Sparse View Recovery cites this paper.

DIP-GS: Deep Image Prior For Gaussian Splatting Sparse View Recovery DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T22:11:44.582968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:11:44.582968Z digest=sha256:7683ca58d5594f57e0f450ab4c0a0b3a3b64f33ddca40db8ef699623bcbccaed

Observation 67047dd6-3a46-4c94-bfb9-6ef2a4899287 · inbound

CharacterShot: Controllable and Consistent 4D Character Animation cites this paper.

CharacterShot: Controllable and Consistent 4D Character Animation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T22:13:13.715526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:13:13.715526Z digest=sha256:4c9b58dd71647a18efc3fdf5582bd324765de57a3ce1382f4c0c454aa40348c8

Observation 12b4b5c7-a6af-4308-991f-e81603bab018 · inbound

4DNeX: Feed-Forward 4D Generative Modeling Made Easy cites this paper.

4DNeX: Feed-Forward 4D Generative Modeling Made Easy DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:19:01.441905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:19:01.441905Z digest=sha256:39918d9564e5a2e5f35bdf8b68a826143ee5cc0e7f90544f3e3811ec8afcc986

Observation 5451d491-ed01-4d82-ad05-b2dd61145ff9 · inbound

Non-invasive Assessment of Pancreatic Duct Hypertension Using Computational Flow Modeling cites this paper.

Non-invasive Assessment of Pancreatic Duct Hypertension Using Computational Flow Modeling DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T18:05:48.158332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:05:48.158332Z digest=sha256:ad63e4515feb6d90925be831607660265203953179c06fd2b62450d1553b6a47

Observation dd7532bf-bf5c-44e9-9d5f-6d6a28aa03cc · inbound

CausNVS: Autoregressive Multi-view Diffusion for Flexible 3D Novel View Synthesis cites this paper.

CausNVS: Autoregressive Multi-view Diffusion for Flexible 3D Novel View Synthesis DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T16:20:44.953957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:20:44.953957Z digest=sha256:6b8eee39bc0862aa39408d22cd8eefe81a1e4ef26b4d6b7c85711363b075a975

Observation 5a4c2ffa-782e-46ef-bbcb-46291589b789 · inbound

PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation cites this paper.

PanoLora: Bridging Perspective and Panoramic Video Generation with LoRA Adaptation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T17:10:00.288556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:10:00.288556Z digest=sha256:ec7df361b06fbc7979606e886eb5b5d17957cc44948c96ecdbd2068639cfe215

Observation 734942c1-68a5-4079-8643-9ed039db3bb6 · inbound

Diff4Splat: Controllable 4D Scene Generation with Latent Dynamic Reconstruction Models cites this paper.

Diff4Splat: Controllable 4D Scene Generation with Latent Dynamic Reconstruction Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-18T02:00:39.437290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-18T01:59:23.928725Z digest=sha256:3fc665591d24247ff3d50411c85e103234923323b9b4e96e3031e3da92f6dcb3

Observation 3aec3c17-26ea-41b5-93d3-47aa3b82acf3 · inbound

GeoWorld: Providing Full-frame Geometry Features to Facilitate 3D Scene Generation cites this paper.

GeoWorld: Providing Full-frame Geometry Features to Facilitate 3D Scene Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T19:39:19.997502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:39:19.997502Z digest=sha256:95c7641858f553d4f23b3980962880eb0119edfa3597d810a1b5b0421f4755eb

Observation 31b16ef8-8470-431d-bc7b-7093fc574e54 · inbound

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation cites this paper.

Geometry-Aware Single-Image 4D Synthesis via Dense Trajectory Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T18:30:06.159144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:30:06.159144Z digest=sha256:f212ba32310edf00a8389c6cdc61cbd9cd5de4942d95f458026d267318b1a064

Observation f201ad96-053d-4e6f-96f0-21758256133a · inbound

InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem cites this paper.

InverseCrafter: Efficient Video ReCapture as a Latent Domain Inverse Problem DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T18:26:35.869199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:26:35.869199Z digest=sha256:7d65f9349cba9bc185c1bd57a50cda4f4ff13eac3ce28868bb9c0364fc590d76

Observation 6b05fe1f-b25b-4861-965e-9af52ed6bffa · inbound

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling cites this paper.

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:29:56.432941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-15T14:29:56.348733Z digest=sha256:f38d2c1cb2d871940417be4c7b6b310df0acdb710584acd68e47947d3dbbf8bf

Observation fedf39fb-5fc6-409d-b7ec-e292438c4c6a · inbound

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling cites this paper.

WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T16:11:16.996902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:11:16.996902Z digest=sha256:29f2b69458954eef6da2da045fcd0d679e6f318d8fe5cb4d1c5f751fa1b9e900

Observation 67243272-4049-4d30-b1bd-7cf1da5a9027 · inbound

CustomX: Unified Character, Action, and Scene Customization in Video World Models cites this paper.

CustomX: Unified Character, Action, and Scene Customization in Video World Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T15:28:58.145189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:28:58.145189Z digest=sha256:3e51bcee3fd955b23491119b42cc3b7d6a7039c1d77d499479dbccca476930d4

Observation 2359316e-8224-46bd-897d-6327e6c6ac0a · inbound

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation cites this paper.

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:17:40.229817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-16T09:16:57.528410Z digest=sha256:346dab77fdbe681052b282286af96265ffd1d1fc1026b7819433df201299354e

Observation 7dc91798-f8c0-4a5d-a1ee-6da955f707a4 · inbound

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation cites this paper.

VideoGPA: Distilling Geometry Priors for 3D-Consistent Video Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T06:13:20.046283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:13:20.046283Z digest=sha256:2bccab1c8fcea21e9f4c0ec76b1976934c3ce28a2e4e0f8cfb944e8ac28ffb0a

Observation 27b5c168-b5f7-401f-a108-e3ec61d1256b · inbound

Reevaluating the Intra-Modal Misalignment Hypothesis in CLIP cites this paper.

Reevaluating the Intra-Modal Misalignment Hypothesis in CLIP DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-13T23:59:10.664837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:59:10.664837Z digest=sha256:c6d7b92c9ad063c6fb94c5223fbb2f5687c8175cfe25970463553d8cac82c452

Observation 2ab9f36a-c574-41b6-8c62-d0d02bab3bd3 · inbound

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras cites this paper.

SparseCam4D: Spatio-Temporally Consistent 4D Reconstruction from Sparse Cameras DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:43:18.016795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-14T23:41:24.423859Z digest=sha256:68f733bd204319416376fdd692d5cffbffa57fe462dc5d1c2a23bdf62389cd18

Observation 143119c4-b953-4610-a46a-3ad3c891ea74 · inbound

Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models cites this paper.

Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T10:16:05.861778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T15:34:15.916685Z digest=sha256:54ee1aa09e0ba7f302dfac9261a6596a643d3a5a6370ce77450e24988c7a23a1

Observation 417ca2ce-e893-4d3a-88d2-6d38cbcda58a · inbound

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models cites this paper.

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:06:19.019208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T06:03:06.920592Z digest=sha256:f0415f1d55baf458df77095d23a99af2831c88f03aa55bd680188264ed01c291

Observation cbd3f361-c68c-425d-baaf-7d9dbbbab1f7 · inbound

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models cites this paper.

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:26:29.595046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-12T02:52:58.232585Z digest=sha256:057693d351e1754d257e8fba131920c78cf68583067b8cd0325fe444f89d91e9

Observation 2a066ccb-78c2-4234-bfe9-3f186f653cc9 · inbound

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models cites this paper.

UniGeo: Unifying Geometric Guidance for Camera-Controllable Image Editing via Video Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T16:51:14.237087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-05T16:47:32.853010Z digest=sha256:7e169af20d7129cc9a9a4b8c9e4d451100a288f0c6bef123fd6f74e66c0c3d5e

Observation fee85766-e775-47fe-88de-b45611d7cd40 · inbound

Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation cites this paper.

Memorize When Needed: Decoupled Memory Control for Spatially Consistent Long-Horizon Video Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:35:19.147994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T04:49:44.525757Z digest=sha256:1ca36acdcd20d2a3c88c2a24de48f6ca67f4bf13f0b8c82f7638019a36cdbbcf

Observation 2f834d82-383e-451b-bdc8-b69a3a4d597a · inbound

Sculpt4D: Generating 4D Shapes via Sparse-Attention Diffusion Transformers cites this paper.

Sculpt4D: Generating 4D Shapes via Sparse-Attention Diffusion Transformers DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:29:06.506297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-09T22:28:54.031663Z digest=sha256:2e20594dd7c95a08eef0e0218f2f572c0b90a8cabc525efda5e2260010afeb7e

Observation 43a9c4b4-b5c1-46b5-ab01-d339e53ff99a · inbound

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling cites this paper.

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:55:32.084388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-08T19:16:50.004359Z digest=sha256:db03d22ba608038612ef594798e1f7b5cbc8219be0a339390f64c525e8b8067a

Observation 97742f30-e10f-4e1b-a4ec-0b5012d6b087 · inbound

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling cites this paper.

Embody4D: A Generalist Data Engine for Embodied 4D World Modeling DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T00:25:09.858911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-01T00:17:39.409452Z digest=sha256:fa3745de9db22eb24d07387150c9b0e05d7b8a9ce545ec991be9235f39440a70

Observation 3c00313e-e4e8-4d38-b5d0-5779b4b8b67b · inbound

ST-Gen4D: Embedding 4D Spatiotemporal Cognition into World Model for 4D Generation cites this paper.

ST-Gen4D: Embedding 4D Spatiotemporal Cognition into World Model for 4D Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:59.306970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-11T01:56:13.909329Z digest=sha256:787d4e9de8060270b4573e20f59e837fa456797677f8500b5739d8d78223de7a

Observation f1cd2688-4729-4590-987f-92b7fa53a40d · inbound

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion cites this paper.

GTA: Advancing Image-to-3D World Generation via Geometry Then Appearance Video Diffusion DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.538276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-14T19:39:16.410139Z digest=sha256:bf5c407a3f14f7dfab7a593582a08780d415ea2bd11034e4328c00b21b6dbd5e

Observation 0de04546-e9ba-4bb7-89f6-4f3112b5627c · inbound

Probing into Camera Control of Video Models cites this paper.

Probing into Camera Control of Video Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:15:04.362015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-06-30T21:11:19.441408Z digest=sha256:78f02a81fd5d2cdd37877083543876660e45a3bb7a394094e178c235b8f45662

Observation 688c0fc0-c6b2-4958-b029-533559ceb455 · inbound

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space cites this paper.

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-07T14:33:54.347002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-07T14:29:23.795094Z digest=sha256:11d291da06e44bbe095f46d710cf2d90a12c18493d4e404e8e879b1c07c86a07

Observation c98c6b60-fc9a-4cc9-805b-840d0643a88a · inbound

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation cites this paper.

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 187

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:46.018532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:23:46.018532Z digest=sha256:7e1e9a9fb671d49c9539d3acc0d7e9a3cbeca85b10c575956c208fd79a40e2db

Observation aff8c725-aae5-4f7d-abea-a5cb52a02b2f · inbound

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans cites this paper.

4DHumanDiff: Direct Text-to-4DGS Generation for Consistent 360-Degree Dynamic Humans DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-01T04:24:27.214945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T04:24:27.214945Z digest=sha256:9ac9499a22f4b15d9f5da9b05200b5dbf129632b1c88eee134de1b9b58a07f6a

Observation ed5225bd-e263-41d2-b000-27ef5070a733 · inbound

Quo Vadis, World Modeling? cites this paper.

Quo Vadis, World Modeling? DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-07T00:14:11.683149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:14:11.683149Z digest=sha256:43cace3da1694f7b9054a830afd6cabfdb5591f2d3f3891e112be6aa59f805ae

Observation 6da1b104-eadc-4a4e-b66f-74ad57cc3554 · inbound

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models cites this paper.

UniWorld-View: Large-Baseline View Synthesis via Video Diffusion Models DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-06T18:39:31.932292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:39:31.932292Z digest=sha256:d75f52ac92b67445755a079370c44a9c5f04467eb7b54291aec59eb2d49af72f

Observation 4c05750e-30d8-47f5-9f99-42fb733b2b68 · inbound

High-Quality Exposure Correction with Diffusion-Based Image Generation Priors cites this paper.

High-Quality Exposure Correction with Diffusion-Based Image Generation Priors DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-14T04:29:23.284867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:29:23.284867Z digest=sha256:4e7c889142d885cd02476e892ba9be8658ac9e733a13cc33d5eaeaab4bb357c8

Observation f36caa24-139e-4522-b242-9bc1cbba093c · inbound

StateFlow: Building, Evolving, and Accessing 3D World States for Previsualization cites this paper.

StateFlow: Building, Evolving, and Accessing 3D World States for Previsualization DimensionX: Create Any 3D and 4D Scenes from a Single Image with Controllable Video Diffusion

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T00:12:27.478355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:12:27.478355Z digest=sha256:0b22ade301276405a1127d55249b573f61ead1b13c14473b7b09a20b379a090e