Pith. sign in

REVIEW 1 cited by

VRCopilot: Authoring 3D Layouts with Generative AI Models in VR

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2408.09382 v1 pith:UFD5PQPX submitted 2024-08-18 cs.HC cs.AIcs.ET

classification cs.HCcs.AIcs.ET
keywords authoringcreationgenerativeimmersiveuseragencyautomaticvrcopilot
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Immersive authoring provides an intuitive medium for users to create 3D scenes via direct manipulation in Virtual Reality (VR). Recent advances in generative AI have enabled the automatic creation of realistic 3D layouts. However, it is unclear how capabilities of generative AI can be used in immersive authoring to support fluid interactions, user agency, and creativity. We introduce VRCopilot, a mixed-initiative system that integrates pre-trained generative AI models into immersive authoring to facilitate human-AI co-creation in VR. VRCopilot presents multimodal interactions to support rapid prototyping and iterations with AI, and intermediate representations such as wireframes to augment user controllability over the created content. Through a series of user studies, we evaluated the potential and challenges in manual, scaffolded, and automatic creation in immersive authoring. We found that scaffolded creation using wireframes enhanced the user agency compared to automatic creation. We also found that manual creation via multimodal specification offers the highest sense of creativity and agency.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MS2Mesh-XR: Multi-modal Sketch-to-Mesh Generation in XR Environments

    cs.CV 2024-12 conditional novelty 3.0 of 10

    A system that turns mid-air sketches plus voice into textured 3D meshes in XR by chaining ControlNet image generation with convolutional mesh reconstruction.

Pith tools