Pith. sign in

REVIEW 3 cited by

DIET-SNN: Direct Input Encoding With Leakage and Threshold Optimization in Deep Spiking Neural Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2008.03658 v3 pith:323UMSSZ submitted 2020-08-09 cs.NE cs.LGstat.ML

classification cs.NEcs.LGstat.ML
keywords diet-snnmembranethresholdinputlatencyleaktrainedfiring
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Bio-inspired spiking neural networks (SNNs), operating with asynchronous binary signals (or spikes) distributed over time, can potentially lead to greater computational efficiency on event-driven hardware. The state-of-the-art SNNs suffer from high inference latency, resulting from inefficient input encoding, and sub-optimal settings of the neuron parameters (firing threshold, and membrane leak). We propose DIET-SNN, a low-latency deep spiking network that is trained with gradient descent to optimize the membrane leak and the firing threshold along with other network parameters (weights). The membrane leak and threshold for each layer of the SNN are optimized with end-to-end backpropagation to achieve competitive accuracy at reduced latency. The analog pixel values of an image are directly applied to the input layer of DIET-SNN without the need to convert to spike-train. The first convolutional layer is trained to convert inputs into spikes where leaky-integrate-and-fire (LIF) neurons integrate the weighted inputs and generate an output spike when the membrane potential crosses the trained firing threshold. The trained membrane leak controls the flow of input information and attenuates irrelevant inputs to increase the activation sparsity in the convolutional and dense layers of the network. The reduced latency combined with high activation sparsity provides large improvements in computational efficiency. We evaluate DIET-SNN on image classification tasks from CIFAR and ImageNet datasets on VGG and ResNet architectures. We achieve top-1 accuracy of 69% with 5 timesteps (inference latency) on the ImageNet dataset with 12x less compute energy than an equivalent standard ANN. Additionally, DIET-SNN performs 20-500x faster inference compared to other state-of-the-art SNN models.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SMM Transformer: Leveraging Spiking Neural Networks for Multimodal Tasks

    cs.NE 2026-08 reject novelty 6.0 of 10

    SMM Transformer uses spiking neurons, spike-driven token mixing, and a spiking mixture of experts to reach ANN-comparable accuracy on vision and vision-language tasks with lower estimated compute energy.

  2. TEFormer: Structured Bidirectional Temporal Enhancement Modeling in Spiking Transformers

    cs.NE 2026-01 conditional novelty 6.0 of 10

    A spiking transformer with forward temporal EMA in attention and backward gated recurrence in the MLP improves accuracy across static, neuromorphic, and temporally complex datasets.

  3. ReverB-SNN: Reversing Bit of the Weight and Activation for Spiking Neural Networks

    cs.CV 2025-06 conditional novelty 5.0 of 10

    ReverB-SNN replaces binary spikes with real-valued spikes and real weights with binary weights, keeping SNN inference addition-only while improving accuracy.

Pith tools