A unified feature codec for CNN and ViT features, built with format and value alignment, beats an architecture-specific baseline on ImageNet classification.
Feature Format Alignment We begin by introducing the feature extraction mechanisms in CNNs and Transformers, as illustrated in Fig
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Cross-architecture universal feature coding via distribution alignment
A unified feature codec for CNN and ViT features, built with format and value alignment, beats an architecture-specific baseline on ImageNet classification.