BSViT: A Burst Spiking Vision Transformer for Expressive and Efficient Visual Representation Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Peng, Hongxiang, Bai, Dewei, Qu, Hong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models
von: Bai, Dewei, et al.
Veröffentlicht: (2026)
von: Bai, Dewei, et al.
Veröffentlicht: (2026)
QB-LIF: Learnable-Scale Quantized Burst Neurons for Efficient SNNs
von: Bai, Dewei, et al.
Veröffentlicht: (2026)
von: Bai, Dewei, et al.
Veröffentlicht: (2026)
Vision-TTT: Efficient and Expressive Visual Representation Learning with Test-Time Training
von: Kong, Quan, et al.
Veröffentlicht: (2026)
von: Kong, Quan, et al.
Veröffentlicht: (2026)
SpikeTrack: A Spike-driven Framework for Efficient Visual Tracking
von: Zhang, Qiuyang, et al.
Veröffentlicht: (2026)
von: Zhang, Qiuyang, et al.
Veröffentlicht: (2026)
Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training
von: Yao, Man, et al.
Veröffentlicht: (2024)
von: Yao, Man, et al.
Veröffentlicht: (2024)
Spiking Transformer:Introducing Accurate Addition-Only Spiking Self-Attention for Transformer
von: Guo, Yufei, et al.
Veröffentlicht: (2025)
von: Guo, Yufei, et al.
Veröffentlicht: (2025)
big.LITTLE Vision Transformer for Efficient Visual Recognition
von: Guo, He, et al.
Veröffentlicht: (2024)
von: Guo, He, et al.
Veröffentlicht: (2024)
StereoVGGT: A Training-Free Visual Geometry Transformer for Stereo Vision
von: Chen, Ziyang, et al.
Veröffentlicht: (2026)
von: Chen, Ziyang, et al.
Veröffentlicht: (2026)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
von: Li, Wenrui, et al.
Veröffentlicht: (2025)
von: Li, Wenrui, et al.
Veröffentlicht: (2025)
Spiking Vision Transformer with Saccadic Attention
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Adaptively Bypassing Vision Transformer Blocks for Efficient Visual Tracking
von: Yang, Xiangyang, et al.
Veröffentlicht: (2024)
von: Yang, Xiangyang, et al.
Veröffentlicht: (2024)
Distilling Vision Transformers for Distortion-Robust Representation Learning
von: Alexis, Konstantinos, et al.
Veröffentlicht: (2026)
von: Alexis, Konstantinos, et al.
Veröffentlicht: (2026)
Spike-EVPR: Deep Spiking Residual Networks with SNN-Tailored Representations for Event-Based Visual Place Recognition
von: Liu, Zuntao, et al.
Veröffentlicht: (2024)
von: Liu, Zuntao, et al.
Veröffentlicht: (2024)
SkyReels-A1: Expressive Portrait Animation in Video Diffusion Transformers
von: Qiu, Di, et al.
Veröffentlicht: (2025)
von: Qiu, Di, et al.
Veröffentlicht: (2025)
Parameter-Efficient and Memory-Efficient Tuning for Vision Transformer: A Disentangled Approach
von: Zhang, Taolin, et al.
Veröffentlicht: (2024)
von: Zhang, Taolin, et al.
Veröffentlicht: (2024)
VisionReasoner: Unified Reasoning-Integrated Visual Perception via Reinforcement Learning
von: Liu, Yuqi, et al.
Veröffentlicht: (2025)
von: Liu, Yuqi, et al.
Veröffentlicht: (2025)
Learning Visual Prompts for Guiding the Attention of Vision Transformers
von: Rezaei, Razieh, et al.
Veröffentlicht: (2024)
von: Rezaei, Razieh, et al.
Veröffentlicht: (2024)
Calibration Attention: Learning Reliability-Aware Representations for Vision Transformers
von: Liang, Wenhao, et al.
Veröffentlicht: (2025)
von: Liang, Wenhao, et al.
Veröffentlicht: (2025)
Revisiting Audio-Visual Segmentation with Vision-Centric Transformer
von: Huang, Shaofei, et al.
Veröffentlicht: (2025)
von: Huang, Shaofei, et al.
Veröffentlicht: (2025)
Efficient Visual Representation Learning with Heat Conduction Equation
von: Zhang, Zhemin, et al.
Veröffentlicht: (2024)
von: Zhang, Zhemin, et al.
Veröffentlicht: (2024)
Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
VAEmo: Efficient Representation Learning for Visual-Audio Emotion with Knowledge Injection
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
von: Cheng, Hao, et al.
Veröffentlicht: (2025)
Spiking-PhysFormer: Camera-Based Remote Photoplethysmography with Parallel Spike-driven Transformer
von: Liu, Mingxuan, et al.
Veröffentlicht: (2024)
von: Liu, Mingxuan, et al.
Veröffentlicht: (2024)
SVFormer: A Direct Training Spiking Transformer for Efficient Video Action Recognition
von: Yu, Liutao, et al.
Veröffentlicht: (2024)
von: Yu, Liutao, et al.
Veröffentlicht: (2024)
Transforming Vision Transformer: Towards Efficient Multi-Task Asynchronous Learning
von: Zhong, Hanwen, et al.
Veröffentlicht: (2025)
von: Zhong, Hanwen, et al.
Veröffentlicht: (2025)
Learning to Merge Tokens via Decoupled Embedding for Efficient Vision Transformers
von: Lee, Dong Hoon, et al.
Veröffentlicht: (2024)
von: Lee, Dong Hoon, et al.
Veröffentlicht: (2024)
SpikeGen: Decoupled "Rods and Cones" Visual Representation Processing with Latent Generative Framework
von: Dai, Gaole, et al.
Veröffentlicht: (2025)
von: Dai, Gaole, et al.
Veröffentlicht: (2025)
Representation Separation for Semantic Segmentation with Vision Transformers
von: Hong, Yuanduo, et al.
Veröffentlicht: (2022)
von: Hong, Yuanduo, et al.
Veröffentlicht: (2022)
Triage: Hierarchical Visual Budgeting for Efficient Video Reasoning in Vision-Language Models
von: Wang, Anmin, et al.
Veröffentlicht: (2026)
von: Wang, Anmin, et al.
Veröffentlicht: (2026)
Making Pose Representations More Expressive and Disentangled via Residual Vector Quantization
von: Jeong, Sukhyun, et al.
Veröffentlicht: (2025)
von: Jeong, Sukhyun, et al.
Veröffentlicht: (2025)
ToSA: Token Selective Attention for Efficient Vision Transformers
von: Singh, Manish Kumar, et al.
Veröffentlicht: (2024)
von: Singh, Manish Kumar, et al.
Veröffentlicht: (2024)
Efficient Unsupervised Visual Representation Learning with Explicit Cluster Balancing
von: Metaxas, Ioannis Maniadis, et al.
Veröffentlicht: (2024)
von: Metaxas, Ioannis Maniadis, et al.
Veröffentlicht: (2024)
Lightweight Pixel Difference Networks for Efficient Visual Representation Learning
von: Su, Zhuo, et al.
Veröffentlicht: (2024)
von: Su, Zhuo, et al.
Veröffentlicht: (2024)
Learning A Spiking Neural Network for Efficient Image Deraining
von: Song, Tianyu, et al.
Veröffentlicht: (2024)
von: Song, Tianyu, et al.
Veröffentlicht: (2024)
Exploring Image Representation with Decoupled Classical Visual Descriptors
von: Qu, Chenyuan, et al.
Veröffentlicht: (2025)
von: Qu, Chenyuan, et al.
Veröffentlicht: (2025)
PRISM: A Promptable and Robust Interactive Segmentation Model with Visual Prompts
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
SPFormer: Enhancing Vision Transformer with Superpixel Representation
von: Mei, Jieru, et al.
Veröffentlicht: (2024)
von: Mei, Jieru, et al.
Veröffentlicht: (2024)
Quantized Spike-driven Transformer
von: Qiu, Xuerui, et al.
Veröffentlicht: (2025)
von: Qiu, Xuerui, et al.
Veröffentlicht: (2025)
Efficient Spike-driven Transformer for High-performance Drone-View Geo-Localization
von: Chen, Zhongwei, et al.
Veröffentlicht: (2025)
von: Chen, Zhongwei, et al.
Veröffentlicht: (2025)
SpikeGS: Learning 3D Gaussian Fields from Continuous Spike Stream
von: Yu, Jinze, et al.
Veröffentlicht: (2024)
von: Yu, Jinze, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models
von: Bai, Dewei, et al.
Veröffentlicht: (2026) -
QB-LIF: Learnable-Scale Quantized Burst Neurons for Efficient SNNs
von: Bai, Dewei, et al.
Veröffentlicht: (2026) -
Vision-TTT: Efficient and Expressive Visual Representation Learning with Test-Time Training
von: Kong, Quan, et al.
Veröffentlicht: (2026) -
SpikeTrack: A Spike-driven Framework for Efficient Visual Tracking
von: Zhang, Qiuyang, et al.
Veröffentlicht: (2026) -
Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training
von: Yao, Man, et al.
Veröffentlicht: (2024)