SpikeHQ uses differentiable neural architecture search to assign each layer of a spiking transformer a uniform or power-of-two quantizer with mixed bit widths, claiming large storage and energy savings at some accuracy cost.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.NE 1years
2024 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Trimming Down Large Spiking Vision Transformers via Heterogeneous Quantization Search
SpikeHQ uses differentiable neural architecture search to assign each layer of a spiking transformer a uniform or power-of-two quantizer with mixed bit widths, claiming large storage and energy savings at some accuracy cost.