REVIEW 5 cited by
Investigating the Design Considerations for Integrating Text-to-Image Generative AI within Augmented Reality Environments
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Generative Artificial Intelligence (GenAI) has emerged as a fundamental component of intelligent interactive systems, enabling the automatic generation of multimodal media content. The continuous enhancement in the quality of Artificial Intelligence-Generated Content (AIGC), including but not limited to images and text, is forging new paradigms for its application, particularly within the domain of Augmented Reality (AR). Nevertheless, the application of GenAI within the AR design process remains opaque. This paper aims to articulate a design space encapsulating a series of criteria and a prototypical process to aid practitioners in assessing the aptness of adopting pertinent technologies. The proposed model has been formulated based on a synthesis of design insights garnered from ten experts, obtained through focus group interviews. Leveraging these initial insights, we delineate potential applications of GenAI in AR.
Forward citations
Cited by 5 Pith papers
-
An Exploratory Study on Multi-modal Generative AI in AR Storytelling
The paper maps how storytellers prefer to use AI-generated text, audio, images, videos, and 3D content to augment AR stories, based on a 223-video analysis and two user studies with 30 participants.
-
CARING-AI: Towards Authoring Context-aware Augmented Reality INstruction through Generative Artificial Intelligence
CARING-AI combines ChatGPT text generation, environment scanning, and smoothed text-to-motion diffusion to let authors create spatially grounded AR avatar instructions without coding or motion capture.
-
Vision-Based Multimodal Interfaces: A Survey and Taxonomy for Enhanced Context-Aware System Design
A systematic survey and taxonomy of vision-based multimodal interfaces, organized around a Macro-Micro-Macro framework for context-aware system design.
-
Exploring Device-Oriented Video Encryption for Hierarchical Privacy Protection in AR Content Sharing
The paper sketches a device-oriented hierarchical ROI encryption scheme for AR sharing, but it is a preliminary position piece without new measurements.
-
MS2Mesh-XR: Multi-modal Sketch-to-Mesh Generation in XR Environments
A system that turns mid-air sketches plus voice into textured 3D meshes in XR by chaining ControlNet image generation with convolutional mesh reconstruction.
Discussion (0). Continue with ORCID to comment.