An end-to-end 3D editing framework achieves high-fidelity local edits from coarse bounding boxes and 2D image prompts using region-aware loss reweighting and a large-scale parts-derived training dataset.
Efficient part-level 3d object generation via dual volume packing.arXiv preprint arXiv:2506.09980
9 Pith papers cite this work. Polarity classification is still indexing.
abstract
Recent progress in 3D object generation has greatly improved both the quality and efficiency. However, most existing methods generate a single mesh with all parts fused together, which limits the ability to edit or manipulate individual parts. A key challenge is that different objects may have a varying number of parts. To address this, we propose a new end-to-end framework for part-level 3D object generation. Given a single input image, our method generates high-quality 3D objects with an arbitrary number of complete and semantically meaningful parts. We introduce a dual volume packing strategy that organizes all parts into two complementary volumes, allowing for the creation of complete and interleaved parts that assemble into the final object. Experiments show that our model achieves better quality, diversity, and generalization than previous image-based part-level generation methods.
citation-role summary
citation-polarity summary
representative citing papers
PhysForge generates physics-grounded 3D assets via a VLM-planned Hierarchical Physical Blueprint and a KineVoxel Injection diffusion model, backed by the new PhysDB dataset of 150,000 annotated assets.
AssemLM fuses SO(3)-equivariant point-cloud features into a VLM to predict discrete 6D assembly poses, reaching ~89% success on AssemBench and improved real-robot multi-step assembly.
UniRecGen unifies reconstruction and generation via shared canonical space and disentangled cooperative learning to produce complete, consistent 3D models from sparse views.
SegviGen shows pretrained 3D generative models can be repurposed for part segmentation via voxel colorization, beating prior methods by 40% interactively and 15% on full segmentation using only 0.32% of labeled data.
ISAP-3D proposes identity-slot aligned modeling with semantic identity tokens and one-to-one layout prediction to achieve stable part-aware 3D generation.
SceneConductor decomposes single-image 3D scene generation into initialization, environment construction, and multi-agent refinement stages with a geometry-aware layout predictor trained on sparse geometric priors from point maps.
A part-wise semi-autoregressive discrete diffusion model for point-cloud-to-mesh generation that separates global structure from local detail, beating prior SOTA on Objaverse.
Home3D 1.0 describes a four-module image-to-3D system using latent SDF, flow-matching, texture fields, material retrieval, and part-specific VAEs to produce meshes with PBR materials and decomposable components.
citing papers explorer
-
EditVerse3D: High-Quality 3D Object Editing with Region-Aware Learning
An end-to-end 3D editing framework achieves high-fidelity local edits from coarse bounding boxes and 2D image prompts using region-aware loss reweighting and a large-scale parts-derived training dataset.
-
PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World
PhysForge generates physics-grounded 3D assets via a VLM-planned Hierarchical Physical Blueprint and a KineVoxel Injection diffusion model, backed by the new PhysDB dataset of 150,000 annotated assets.
-
AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly
AssemLM fuses SO(3)-equivariant point-cloud features into a VLM to predict discrete 6D assembly poses, reaching ~89% success on AssemBench and improved real-robot multi-step assembly.
-
UniRecGen: Unifying Multi-View 3D Reconstruction and Generation
UniRecGen unifies reconstruction and generation via shared canonical space and disentangled cooperative learning to produce complete, consistent 3D models from sparse views.
-
SegviGen: Repurposing 3D Generative Model for Part Segmentation
SegviGen shows pretrained 3D generative models can be repurposed for part segmentation via voxel colorization, beating prior methods by 40% interactively and 15% on full segmentation using only 0.32% of labeled data.
-
ISAP-3D: Identity-Slot Aligned Part-Aware 3D Generation
ISAP-3D proposes identity-slot aligned modeling with semantic identity tokens and one-to-one layout prediction to achieve stable part-aware 3D generation.
-
SceneConductor: 3D Scene Generation from a Single Image with Multi-Agent Orchestration
SceneConductor decomposes single-image 3D scene generation into initialization, environment construction, and multi-agent refinement stages with a geometry-aware layout predictor trained on sparse geometric priors from point maps.
-
PartDiffuser: Part-wise 3D Mesh Generation via Discrete Diffusion
A part-wise semi-autoregressive discrete diffusion model for point-cloud-to-mesh generation that separates global structure from local detail, beating prior SOTA on Objaverse.
-
Home3D 1.0: A High-Fidelity Image-to-3D Asset Generation System for Interior Design
Home3D 1.0 describes a four-module image-to-3D system using latent SDF, flow-matching, texture fields, material retrieval, and part-specific VAEs to produce meshes with PBR materials and decomposable components.