The authors connect a multimodal language model to a diffusion model so that multiple image regions can be inpainted simultaneously, each with a distinct automatically generated text prompt.
Gonzalez, Ion Stoica, and Eric P
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
I Dream My Painting: Connecting MLLMs and Diffusion Models via Prompt Generation for Text-Guided Multi-Mask Inpainting
The authors connect a multimodal language model to a diffusion model so that multiple image regions can be inpainted simultaneously, each with a distinct automatically generated text prompt.