A prompt-wrapper framework for vision-language models reports improved story and poetry generation on a private benchmark, without releasing code, data, or models.
COVID-19 detection in chest x-ray images using swin-transformer and transformer in transformer,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
REJECT 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
VisuCraft: Enhancing Large Vision-Language Models for Complex Visual-Guided Creative Content Generation via Structured Information Extraction
A prompt-wrapper framework for vision-language models reports improved story and poetry generation on a private benchmark, without releasing code, data, or models.