Pith. sign in

REVIEW 5 cited by

Generative AI meets 3D: A Survey on Text-to-3D in AIGC Era

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.06131 v4 pith:ORZSIP7W submitted 2023-05-10 cs.CV

classification cs.CV
keywords text-to-3dgenerationfieldincludingresearchaigccontentdata
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Generative AI has made significant progress in recent years, with text-guided content generation being the most practical as it facilitates interaction between human instructions and AI-generated content (AIGC). Thanks to advancements in text-to-image and 3D modeling technologies, like neural radiance field (NeRF), text-to-3D has emerged as a nascent yet highly active research field. Our work conducts a comprehensive survey on this topic and follows up on subsequent research progress in the overall field, aiming to help readers interested in this direction quickly catch up with its rapid development. First, we introduce 3D data representations, including both Structured and non-Structured data. Building on this pre-requisite, we introduce various core technologies to achieve satisfactory text-to-3D results. Additionally, we present mainstream baselines and research directions in recent text-to-3D technology, including fidelity, efficiency, consistency, controllability, diversity, and applicability. Furthermore, we summarize the usage of text-to-3D technology in various applications, including avatar generation, texture generation, scene generation and 3D editing. Finally, we discuss the agenda for the future development of text-to-3D.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Hallo4D uses vision-language models to detect and correct spatial and temporal mistakes in AI-generated 3D and 4D content, improving consistency without retraining the base generators.

  2. A11yShape: AI-Assisted 3-D Modeling for Blind and Low-Vision Programmers

    cs.HC 2025-08 conditional novelty 6.0 of 10

    A11yShape enables blind and low-vision programmers to create and modify 3-D models through an AI-assisted, code-based system, as demonstrated in a four-participant study.

  3. "You'll Be Alice Adventuring in Wonderland!" Processes, Challenges, and Opportunities of Creating Animated Virtual Reality Stories

    cs.HC 2025-02 conditional novelty 6.0 of 10

    A 21-creator interview study identifies ten common stages and nine unique challenges in creating animated VR stories, with story-driven and visual-driven workflow types.

  4. PlantDreamer: Achieving Realistic 3D Plant Models with Diffusion-Guided Gaussian Splatting

    cs.CV 2025-05 conditional novelty 5.0 of 10

    A diffusion-guided Gaussian splatting pipeline generates realistic 3D plants from L-system meshes or point clouds and beats GaussianDreamer on masked PSNR for bean, kale and mint.

  5. Reconstructing 4D Spatial Intelligence: A Survey

    cs.CV 2025-07 accept novelty 4.0 of 10

    A review that classifies 4D scene reconstruction methods into five progressive levels: low-level cues, scene components, dynamic scenes, interactions, and physics.

Pith tools