CrossWKV adapts RWKV-7's WKV state update into a cross-modal attention layer for diffusion text-to-image generation, but the reported benchmark results are explicitly labeled preliminary and in-progress.
Zero-shot text-to- image generation
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Cross-attention for State-based model RWKV-7
CrossWKV adapts RWKV-7's WKV state update into a cross-modal attention layer for diffusion text-to-image generation, but the reported benchmark results are explicitly labeled preliminary and in-progress.