REVIEW 8 cited by
DIET-SNN: Direct Input Encoding With Leakage and Threshold Optimization in Deep Spiking Neural Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Bio-inspired spiking neural networks (SNNs), operating with asynchronous binary signals (or spikes) distributed over time, can potentially lead to greater computational efficiency on event-driven hardware. The state-of-the-art SNNs suffer from high inference latency, resulting from inefficient input encoding, and sub-optimal settings of the neuron parameters (firing threshold, and membrane leak). We propose DIET-SNN, a low-latency deep spiking network that is trained with gradient descent to optimize the membrane leak and the firing threshold along with other network parameters (weights). The membrane leak and threshold for each layer of the SNN are optimized with end-to-end backpropagation to achieve competitive accuracy at reduced latency. The analog pixel values of an image are directly applied to the input layer of DIET-SNN without the need to convert to spike-train. The first convolutional layer is trained to convert inputs into spikes where leaky-integrate-and-fire (LIF) neurons integrate the weighted inputs and generate an output spike when the membrane potential crosses the trained firing threshold. The trained membrane leak controls the flow of input information and attenuates irrelevant inputs to increase the activation sparsity in the convolutional and dense layers of the network. The reduced latency combined with high activation sparsity provides large improvements in computational efficiency. We evaluate DIET-SNN on image classification tasks from CIFAR and ImageNet datasets on VGG and ResNet architectures. We achieve top-1 accuracy of 69% with 5 timesteps (inference latency) on the ImageNet dataset with 12x less compute energy than an equivalent standard ANN. Additionally, DIET-SNN performs 20-500x faster inference compared to other state-of-the-art SNN models.
Forward citations
Cited by 8 Pith papers
-
SMM Transformer: Leveraging Spiking Neural Networks for Multimodal Tasks
SMM Transformer uses spiking neurons, spike-driven token mixing, and a spiking mixture of experts to reach ANN-comparable accuracy on vision and vision-language tasks with lower estimated compute energy.
-
TEFormer: Structured Bidirectional Temporal Enhancement Modeling in Spiking Transformers
A spiking transformer with forward temporal EMA in attention and backward gated recurrence in the MLP improves accuracy across static, neuromorphic, and temporally complex datasets.
-
ReverB-SNN: Reversing Bit of the Weight and Activation for Spiking Neural Networks
ReverB-SNN replaces binary spikes with real-valued spikes and real weights with binary weights, keeping SNN inference addition-only while improving accuracy.
-
Neuromorphic Optical Tracking and Imaging of Randomly Moving Targets through Strongly Scattering Media
An event camera plus a two-module spiking neural network tracks and reconstructs MNIST and Kanji characters hidden behind strongly scattering media in benchtop experiments.
-
Enhanced Temporal Processing in Spiking Neural Networks for Static Object Detection Using 3D Convolutions
A directly trained spiking YOLOv5n using 3D convolutions and a temporal recurrence mechanism reports mAP within 0.001 to 0.008 of a same-architecture ANN on COCO2017 and VOC at 224x224.
-
ALADE-SNN: Adaptive Logit Alignment in Dynamically Expandable Spiking Neural Networks for Class Incremental Learning
ALADE-SNN combines dynamic network expansion with adaptive logit alignment and OtoN weight suppression to improve class incremental learning in spiking neural networks.
-
The Neural Division of Labor: Biologically-Inspired Modular Architectures for Robust Neuromorphic Computing
A modular spiking network with isolated per-class experts and a hidden-activity loss is claimed to match dense accuracy with far fewer spikes and parameters, though key training details and internal numbers are inconsistent.
-
Spiking Neural Network Feature Discrimination Boosts Modality Fusion
Applying L2 normalization to the final hidden features of audio and visual spiking networks, followed by spiking MLP fusion, yields 98.6% on CIFAR10-AV and 97.2% on UrbanSound8K-AV, outperforming a transformer-based S...
Discussion (0). Continue with ORCID to comment.