Skip to content
Open access

Decoder-Aligned Cumulative-Prefix Training for Low-Latency Spiking Neural Network Classification

2026 · IEEE Access · Vol 14, pp. 103491-103507 · 0 citations · 46 references
Computer Science

TL;DR

The results support DACP as a decoder-aligned objective for cumulative anytime decoding when class evidence is temporally distributed, while identifying fast-saturating event-image tasks as boundary cases rather than domains of universal DACP dominance.

Abstract

Low-latency spiking neural network classification depends on the temporal decision state queried by the deployment decoder. For cumulative spike-count decoding, losses attached only to terminal counts, instantaneous spikes, or hidden readouts can optimize states that differ from the cumulative prefixes used for anytime decisions. This paper studies decoder-aligned cumulative-prefix training (DACP) for directly trained SNNs. DACP applies one loss-design rule: supervise the cumulative decoder state used at inference. The resulting curriculum begins with final-horizon count supervision and then adds cross-entropy on cumulative prefix logits in the same space consumed by the anytime decoder. The reported protocol compares Count, Per-step CE, Static Blend, TET, and DACP over five seeds on N-MNIST, CIFAR10-DVS, SHD, and SSC under fixed architectures, optimizer settings, checkpoint rules, and SynOp accounting; robustness checks cover schedules, prefix weights, recurrent depth and width, adaptive recurrent and transformer-style SHD presets, fast-sigmoid substitution, and TET native-readout re-evaluation. On the temporally rich SHD task, DACP raises $\gamma =0.80$ stop accuracy to 64.36%, compared with 47.10% for Static Blend and 36.51% for Count, and raises AATC to 51.84%, compared with 41.72% and 37.56%, respectively. On SSC, DACP gives the highest AATC among the tested losses (27.67% versus 26.64% for Count), but the margin is small. These results support DACP as a decoder-aligned objective for cumulative anytime decoding when class evidence is temporally distributed, while identifying fast-saturating event-image tasks as boundary cases rather than domains of universal DACP dominance.

Read PDF

Similar papers

Preprint Aug 2026

BASC : Behavior-Aligned Quantization and Pruning for Low-Bit Spiking Neural Networks

Extensive experiments on static and neuromorphic benchmarks show that lower-bit BASC models match or outperform higher-bit baselines and retain this accuracy advantage after structured pruning, while further reducing model storage and synaptic operations.

Linliang Chen, Yan Zhong, Xin Liu et al. · 0 citations
Preprint Aug 2026

Reducing ANN-SNN Conversion Error via Residual Membrane Potential Alignment

This work analyzes flaws of conventional conversion pipelines from residual membrane potential statistics and proposes a novel conversion strategy combining dynamic initial potential tuning and feature enhancement, which generalizes to ReLU CNNs, ANN Transformers, and multi-threshold SNN variants.

Zirui Chen, Zihan Huang, Tong Bu et al. · 0 citations
Preprint Aug 2026

PTQ4SNN: Membrane-Aware Post-Training Quantization for Spiking Neural Networks

PTQ4SNN is proposed, a membrane-aware post-training quantization framework that jointly quantizes weights and recurrent membrane states using only a small calibration set and effectively preserves model accuracy under W4 quantization and approximately 4-bit membrane precision.

Hui Xie, Tong Shi, Haotong Qin et al. · 0 citations
2025

Adaptive Fission: Post-training Encoding for Low-latency Spike Neural Networks

Adaptive Fission is proposed, a post-training encoding technique that selectively splits high-sensitivity neurons into groups with varying scales and weights that enables neuron-specific, on-demand precision and threshold allocation while introducing minimal spatial overhead.

Yizhou Jiang, Feng Chen, Yihan Li et al. · 2 citations
Preprint Aug 2026

SpikeWorld: Fast-State Adaptation for Frozen Spiking World Models

SpikeWorld, a 1.45M-parameter sparse spiking model jointly trained for heterogeneous sensory prediction, semantics, image-text binding and action-conditioned dynamics, is introduced, showing that the contribution is not superior linear identification, but its integration with a frozen multimodal spiking checkpoint.

Ziqiao Yu · 0 citations
Jul 2026

The Sparsity Ceiling: Where Spiking Networks Can and Cannot Trade Activity for Energy

It is argued the energy dividend of sparsity is not a property of SNNs but of the task, and the ceiling is formalized with an information-theoretic bound and confirmed: the floor rises with memory load, falls with state width, and (refuting a naive memory-only reading) rises with task difficulty.

Zeyu Wang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.