REVIEW 2 cited by
CLIP-CLOP: CLIP-Guided Collage and Photomontage
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The unabated mystique of large-scale neural networks, such as the CLIP dual image-and-text encoder, popularized automatically generated art. Increasingly more sophisticated generators enhanced the artworks' realism and visual appearance, and creative prompt engineering enabled stylistic expression. Guided by an artist-in-the-loop ideal, we design a gradient-based generator to produce collages. It requires the human artist to curate libraries of image patches and to describe (with prompts) the whole image composition, with the option to manually adjust the patches' positions during generation, thereby allowing humans to reclaim some control of the process and achieve greater creative freedom. We explore the aesthetic potentials of high-resolution collages, and provide an open-source Google Colab as an artistic tool.
Forward citations
Cited by 2 Pith papers
-
Empowering LLMs to Understand and Generate Complex Vector Graphics
LLM4SVG adds learnable SVG tokens and SFT data so LLMs can generate and describe scalable vector graphics much better than general-purpose LLMs.
-
SVGDreamer++: Advancing Editability and Diversity in Text-Guided SVG Generation
SVGDreamer++ uses SAM-based hierarchical masks and adaptive path control to generate text-guided SVGs that are more editable and visually detailed.
Discussion (0). Continue with ORCID to comment.