cs.AIOct 1, 2026

Contrastive Attention Mitigates Spectral Bias in Spiking Transformers

Authors: Xiaoli Liu, Malu Zhang, Yang Yang

Organizations: School of Computer Science and Engineering, University of Electronic Science and Technology of China

Abstract

Spiking Transformers merge the energy-efficiency of spiking neural networks (SNNs) with the representational power of self-attention, creating a promising architecture for high-performance, energy-efficient computation. However, a performance gap persists versus its counterparts in artificial neural networks (ANNs). Unlike prior works attributing this to binary activations, we reveal that both spiking neurons and spiking self-attention (SSA) act as low-pass filters through multiscale spectral analysis. This characteristic leads to the dissipation of high-frequency components. To address this issue, we propose the Spiking Contrastive Attention (SCA) paradigm, which draw inspiration from the edge-detection and differential sensing properties of biological visual system. By extracting contrast prototypes via global contrastive aggregation and applying local differential refinement, SCA effectively enhances high-frequency information. Extensive experiments show that SCA is a general module that consistently boosts Spiking Transformers across image classification, semantic segmentation, and event-based tracking. Furthermore, it achieves lower complexity, offering superior efficiency over original SSA. These results establish its potential as a fundamental building block for energy-efficient Spiking Transformers.

Figures & tables

Explore similar work

CardsList
  1. SAFformer:Improving Spiking Transformer via Active Predictive Filtering

    May 8, 2026Zequan Xie, Weiming Zeng, Yunhua Chen +3Spiking Neural NetworksTransformer Architectures

  2. Breaking Global Self-Attention Bottlenecks in Transformer-based Spiking Neural Networks with Local Structure-Aware Self-Attention

    May 12, 2026Lingdong Li, Hangming Zhang, Qiang YuSpiking Neural NetworksTransformer Attention

  3. Rethinking Attention Locality in Spiking Transformers

    Aug 9, 2026Zeqi Zheng, Zizheng Zhu, Yuping Yan +3Transformer AttentionTime-To-First-Spike