Pith. sign in

Q-SNNs: Quantized Spiking Neural Networks

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Brain-inspired Spiking Neural Networks (SNNs) leverage sparse spikes to represent information and process them in an asynchronous event-driven manner, offering an energy-efficient paradigm for the next generation of machine intelligence. However, the current focus within the SNN community prioritizes accuracy optimization through the development of large-scale models, limiting their viability in resource-constrained and low-power edge devices. To address this challenge, we introduce a lightweight and hardware-friendly Quantized SNN (Q-SNN) that applies quantization to both synaptic weights and membrane potentials. By significantly compressing these two key elements, the proposed Q-SNNs substantially reduce both memory usage and computational complexity. Moreover, to prevent the performance degradation caused by this compression, we present a new Weight-Spike Dual Regulation (WS-DR) method inspired by information entropy theory. Experimental evaluations on various datasets, including static and neuromorphic, demonstrate that our Q-SNNs outperform existing methods in terms of both model size and accuracy. These state-of-the-art results in efficiency and efficacy suggest that the proposed method can significantly improve edge intelligent computing.

fields

cs.CV 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

QP-SNN: Quantized and Pruned Spiking Neural Networks

cs.CV · 2025-02-09 · conditional · novelty 5.0

A quantized and pruned spiking neural network with weight rescaling and singular-value-based pruning reaches comparable accuracy at roughly a tenth of the model size.

citing papers explorer

Showing 1 of 1 citing paper.

  • QP-SNN: Quantized and Pruned Spiking Neural Networks cs.CV · 2025-02-09 · conditional · none · ref 20 · internal anchor

    A quantized and pruned spiking neural network with weight rescaling and singular-value-based pruning reaches comparable accuracy at roughly a tenth of the model size.