Pith. sign in

REVIEW 6 cited by

An Object is Worth 64x64 Pixels: Generating 3D Object via Image Diffusion

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.03178 v1 pith:JZTA4UY3 submitted 2024-08-06 cs.CV cs.GRcs.LG

An Object is Worth 64x64 Pixels: Generating 3D Object via Image Diffusion

classification cs.CV cs.GRcs.LG
keywords generationimagemodelsobjectapproachdiffusiongeneratingpatch
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

We introduce a new approach for generating realistic 3D models with UV maps through a representation termed "Object Images." This approach encapsulates surface geometry, appearance, and patch structures within a 64x64 pixel image, effectively converting complex 3D shapes into a more manageable 2D format. By doing so, we address the challenges of both geometric and semantic irregularity inherent in polygonal meshes. This method allows us to use image generation models, such as Diffusion Transformers, directly for 3D shape generation. Evaluated on the ABO dataset, our generated shapes with patch structures achieve point cloud FID comparable to recent 3D generative models, while naturally supporting PBR material generation.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 6 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. MeshFlow: Mesh Generation with Equivariant Flow Matching

    cs.GR 2026-06 unverdicted novelty 7.0

    MeshFlow applies equivariant optimal-transport flow matching to generate triangle meshes as soups, matching autoregressive quality with an 18x inference speedup.

  2. Garment Particles: A 2D--3D Symmetric Garment Representation for Generation and Editing

    cs.GR 2026-05 unverdicted novelty 7.0

    Garment Particles is a 5D point cloud representation jointly encoding 2D sewing patterns and 3D geometry, supporting rectified flow generation from high-level inputs and diffusion-based editing of patterns or shapes.

  3. VesselRW: Weakly Supervised Subcutaneous Vessel Segmentation via Learned Random Walk Propagation

    cs.CV 2025-08 unverdicted novelty 5.0

    VesselRW expands sparse vessel annotations into dense probabilistic supervision via a jointly trained differentiable random walk model with uncertainty weighting and topology regularization for CNN-based subcutaneous ...

  4. DualResolution Residual Architecture with Artifact Suppression for Melanocytic Lesion Segmentation

    cs.CV 2025-08 unverdicted novelty 5.0

    Dual-resolution residual architecture with boundary-aware connections, channel attention, artifact suppression, and combined Dice-Tversky plus boundary and contrastive losses improves lesion boundary precision over st...

  5. Edge Detection for Organ Boundaries via Top Down Refinement and SubPixel Upsampling

    cs.CV 2025-08 unverdicted novelty 5.0

    A top-down backward refinement network with subpixel upsampling generates crisp high-resolution organ boundaries in medical images and improves downstream segmentation and registration performance.

  6. Deeply Dual Supervised learning for melanoma recognition

    cs.CV 2025-08 unverdicted novelty 4.0

    A dual-pathway deep learning model with attention mechanisms and multi-scale feature aggregation claims superior accuracy and fewer false positives for melanoma detection on benchmark datasets.