A sparse set of massive activation channels in DiTs carries semantic information, proven critical by disruption tests, spatially aligned with image subjects via clustering, and transferable for semantic interpolation between prompts.
GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 2roles
background 1polarities
background 1representative citing papers
HPD v2 is the largest human preference dataset for text-to-image images with 798k choices, and HPS v2 is the resulting CLIP-based scorer that better predicts human judgments and responds to model improvements.
citing papers explorer
-
Few Channels Draw The Whole Picture: Revealing Massive Activations in Diffusion Transformers
A sparse set of massive activation channels in DiTs carries semantic information, proven critical by disruption tests, spatially aligned with image subjects via clustering, and transferable for semantic interpolation between prompts.
-
Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis
HPD v2 is the largest human preference dataset for text-to-image images with 798k choices, and HPS v2 is the resulting CLIP-based scorer that better predicts human judgments and responds to model improvements.