Abstract:Spiking Neural Networks have attracted significant attention in recent years due to their distinctive low-power characteristics. Meanwhile, Transformer models, known for their powerful self-attention mechanisms and parallel processing capabilities, have demonstrated exceptional performance across various domains, including natural language processing and computer vision. Despite the significant advantages of both SNNs and Transformers, directly combining the low-power benefits of SNNs with the high performance of Transformers remains challenging. Specifically, while the sparse computing mode of SNNs contributes to reduced energy consumption, traditional attention mechanisms depend on dense matrix computations and complex softmax operations. This reliance poses significant challenges for effective execution in low-power scenarios. Given the tremendous success of Transformers in deep learning, it is a necessary step to explore the integration of SNNs and Transformers to harness the strengths of both. In this paper, we propose a novel model architecture, Spike Aggregation Transformer (SAFormer), that integrates the low-power characteristics of SNNs with the high-performance advantages of Transformer models. The core contribution of SAFormer lies in the design of the Spike Aggregated Self-Attention (SASA) mechanism, which significantly simplifies the computation process by calculating attention weights using only the spike matrices query and key, thereby effectively reducing energy consumption. Additionally, we introduce a Depthwise Convolution Module (DWC) to enhance the feature extraction capabilities, further improving overall accuracy. We evaluated and demonstrated that SAFormer outperforms state-of-the-art SNNs in both accuracy and energy consumption, highlighting its significant advantages in low-power and high-performance computing.

Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training

Spikeformer: Training high-performance spiking neural network with transformer

Efficient Structure Slimming for Spiking Neural Networks

Spike Trains Encoding and Threshold Rescaling Method for Deep Spiking Neural Networks

Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic Chips

Towards High-performance Spiking Transformers from ANN to SNN Conversion

You Only Spike Once: Improving Energy-Efficient Neuromorphic Inference to ANN-Level Accuracy

SpikingMiniLM: Energy-Efficient Spiking Transformer for Natural Language Understanding

Spikingformer: Spike-driven Residual Learning for Transformer-based Spiking Neural Network

SpikingResformer: Bridging ResNet and Vision Transformer in Spiking Neural Networks

SpikeZIP-TF: Conversion is All You Need for Transformer-based SNN

Masked Spiking Transformer

Combining Aggregated Attention and Transformer Architecture for Accurate and Efficient Performance of Spiking Neural Networks

PT-Spike: A Precise-Time-Dependent Single Spike Neuromorphic Architecture with Efficient Supervised Learning

Training a General Spiking Neural Network with Improved Efficiency and Minimum Latency

SpikeConverter: an Efficient Conversion Framework Zipping the Gap Between Artificial Neural Networks and Spiking Neural Networks

Enabling Deep Spiking Neural Networks with Hybrid Conversion and Spike Timing Dependent Backpropagation

Integer-Valued Training and Spike-Driven Inference Spiking Neural Network for High-performance and Energy-efficient Object Detection

Spikformer: When Spiking Neural Network Meets Transformer

Direct Training for Spiking Neural Networks: Faster, Larger, Better

Toward High-Accuracy and Low-Latency Spiking Neural Networks With Two-Stage Optimization