OVOW reconstructs instance-level, simulation-ready 4D mesh scenes from monocular video via a four-stage training-free pipeline and introduces a new benchmark for structured Video-to-4D evaluation.
arXiv preprint arXiv:2601.11514 (2026)
5 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
years
2026 5verdicts
UNVERDICTED 5roles
baseline 1polarities
baseline 1representative citing papers
A reproducible VLM-judge protocol with position-bias correction is validated as superior to CLIP similarity and geometry-validity proxies for assessing single-image 3D mesh quality.
PAD synthesizes 3D geometry in observation space via depth unprojection as anchor to eliminate pose ambiguity in image-to-3D generation.
Single-view mesh reconstruction generalizes poorly to robot camera rotations, inducing MDE distortion and layout drift, while a gravity-aware refinement cuts one-stage layout-orientation error by 47.1%.
FlowObject reformulates sparse-view 3D reconstruction as a training-free guided inverse problem in flow-matching models, augmented by 3DGS refinement to improve geometric completeness and fidelity.
citing papers explorer
-
One Video, One World: Turning Monocular Video into Physical 4D Scenes
OVOW reconstructs instance-level, simulation-ready 4D mesh scenes from monocular video via a four-stage training-free pipeline and introduces a new benchmark for structured Video-to-4D evaluation.
-
A Cross-Model VLM-Judge Protocol for Single-Image 3D Mesh Quality (and Why Cheap Proxies Fall Short)
A reproducible VLM-judge protocol with position-bias correction is validated as superior to CLIP similarity and geometry-validity proxies for assessing single-image 3D mesh quality.
-
Pose-Aware Diffusion for 3D Generation
PAD synthesizes 3D geometry in observation space via depth unprojection as anchor to eliminate pose ambiguity in image-to-3D generation.
-
Can Single-View Mesh Reconstruction Generalize to Robot Camera Rotation?
Single-view mesh reconstruction generalizes poorly to robot camera rotations, inducing MDE distortion and layout drift, while a gravity-aware refinement cuts one-stage layout-orientation error by 47.1%.
-
FlowObject: Flow Steering for Bridging Generative Priors and Reconstruction Fidelity
FlowObject reformulates sparse-view 3D reconstruction as a training-free guided inverse problem in flow-matching models, augmented by 3DGS refinement to improve geometric completeness and fidelity.