SVFormer: A Direct Training Spiking Transformer for Efficient Video Action Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Yu, Liutao, Huang, Liwei, Zhou, Chenlin, Zhang, Han, Ma, Zhengyu, Zhou, Huihui, Tian, Yonghong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
QKFormer: Hierarchical Spiking Transformer using Q-K Attention
di: Zhou, Chenlin, et al.
Pubblicazione: (2024)
di: Zhou, Chenlin, et al.
Pubblicazione: (2024)
Direct Training High-Performance Deep Spiking Neural Networks: A Review of Theories and Methods
di: Zhou, Chenlin, et al.
Pubblicazione: (2024)
di: Zhou, Chenlin, et al.
Pubblicazione: (2024)
Long-Range Feedback Spiking Network Captures Dynamic and Static Representations of the Visual Cortex under Movie Stimuli
di: Huang, Liwei, et al.
Pubblicazione: (2023)
di: Huang, Liwei, et al.
Pubblicazione: (2023)
Spikingformer: A Key Foundation Model for Spiking Neural Networks
di: Zhou, Chenlin, et al.
Pubblicazione: (2023)
di: Zhou, Chenlin, et al.
Pubblicazione: (2023)
SpikePoint: An Efficient Point-based Spiking Neural Network for Event Cameras Action Recognition
di: Ren, Hongwei, et al.
Pubblicazione: (2023)
di: Ren, Hongwei, et al.
Pubblicazione: (2023)
High-speed and High-quality Vision Reconstruction of Spike Camera with Spike Stability Theorem
di: Zhang, Wei, et al.
Pubblicazione: (2024)
di: Zhang, Wei, et al.
Pubblicazione: (2024)
Parallel Spiking Neurons with High Efficiency and Ability to Learn Long-term Dependencies
di: Fang, Wei, et al.
Pubblicazione: (2023)
di: Fang, Wei, et al.
Pubblicazione: (2023)
Fire on Motion: Optimizing Video Pass-bands for Efficient Spiking Action Recognition
di: Ye, Shuhan, et al.
Pubblicazione: (2026)
di: Ye, Shuhan, et al.
Pubblicazione: (2026)
MSVIT: Improving Spiking Vision Transformer Using Multi-scale Attention Fusion
di: Hua, Wei, et al.
Pubblicazione: (2025)
di: Hua, Wei, et al.
Pubblicazione: (2025)
Scaling Spike-driven Transformer with Efficient Spike Firing Approximation Training
di: Yao, Man, et al.
Pubblicazione: (2024)
di: Yao, Man, et al.
Pubblicazione: (2024)
Efficient Speech Command Recognition Leveraging Spiking Neural Network and Curriculum Learning-based Knowledge Distillation
di: Wang, Jiaqi, et al.
Pubblicazione: (2024)
di: Wang, Jiaqi, et al.
Pubblicazione: (2024)
MoCrop: Training Free Motion Guided Cropping for Efficient Video Action Recognition
di: Huang, Binhua, et al.
Pubblicazione: (2025)
di: Huang, Binhua, et al.
Pubblicazione: (2025)
Human-Centric Transformer for Domain Adaptive Action Recognition
di: Lin, Kun-Yu, et al.
Pubblicazione: (2024)
di: Lin, Kun-Yu, et al.
Pubblicazione: (2024)
ActionHub: A Large-scale Action Video Description Dataset for Zero-shot Action Recognition
di: Zhou, Jiaming, et al.
Pubblicazione: (2024)
di: Zhou, Jiaming, et al.
Pubblicazione: (2024)
Multi-Timescale Motion-Decoupled Spiking Transformer for Audio-Visual Zero-Shot Learning
di: Li, Wenrui, et al.
Pubblicazione: (2025)
di: Li, Wenrui, et al.
Pubblicazione: (2025)
Dark Transformer: A Video Transformer for Action Recognition in the Dark
di: Ulhaq, Anwaar
Pubblicazione: (2024)
di: Ulhaq, Anwaar
Pubblicazione: (2024)
Spike-driven Transformer V2: Meta Spiking Neural Network Architecture Inspiring the Design of Next-generation Neuromorphic Chips
di: Yao, Man, et al.
Pubblicazione: (2024)
di: Yao, Man, et al.
Pubblicazione: (2024)
Leveraging Temporal Contextualization for Video Action Recognition
di: Kim, Minji, et al.
Pubblicazione: (2024)
di: Kim, Minji, et al.
Pubblicazione: (2024)
SSTFormer: Bridging Spiking Neural Network and Memory Support Transformer for Frame-Event based Recognition
di: Wang, Xiao, et al.
Pubblicazione: (2023)
di: Wang, Xiao, et al.
Pubblicazione: (2023)
Efficient 3D Recognition with Event-driven Spike Sparse Convolution
di: Qiu, Xuerui, et al.
Pubblicazione: (2024)
di: Qiu, Xuerui, et al.
Pubblicazione: (2024)
TP-Spikformer: Token Pruned Spiking Transformer
di: Wei, Wenjie, et al.
Pubblicazione: (2026)
di: Wei, Wenjie, et al.
Pubblicazione: (2026)
ReSpike: Residual Frames-based Hybrid Spiking Neural Networks for Efficient Action Recognition
di: Xiao, Shiting, et al.
Pubblicazione: (2024)
di: Xiao, Shiting, et al.
Pubblicazione: (2024)
Unleashing the Power of CNN and Transformer for Balanced RGB-Event Video Recognition
di: Wang, Xiao, et al.
Pubblicazione: (2023)
di: Wang, Xiao, et al.
Pubblicazione: (2023)
Binary Event-Driven Spiking Transformer
di: Cao, Honglin, et al.
Pubblicazione: (2025)
di: Cao, Honglin, et al.
Pubblicazione: (2025)
HoloTime: Taming Video Diffusion Models for Panoramic 4D Scene Generation
di: Zhou, Haiyang, et al.
Pubblicazione: (2025)
di: Zhou, Haiyang, et al.
Pubblicazione: (2025)
Adversarial Robustness in RGB-Skeleton Action Recognition: Leveraging Attention Modality Reweighter
di: Liu, Chao, et al.
Pubblicazione: (2024)
di: Liu, Chao, et al.
Pubblicazione: (2024)
Efficient-VLN: A Training-Efficient Vision-Language Navigation Model
di: Zheng, Duo, et al.
Pubblicazione: (2025)
di: Zheng, Duo, et al.
Pubblicazione: (2025)
Towards Scalable Modeling of Compressed Videos for Efficient Action Recognition
di: Biswas, Shristi Das, et al.
Pubblicazione: (2025)
di: Biswas, Shristi Das, et al.
Pubblicazione: (2025)
Enhancing Video Transformers for Action Understanding with VLM-aided Training
di: Lu, Hui, et al.
Pubblicazione: (2024)
di: Lu, Hui, et al.
Pubblicazione: (2024)
TDS-CLIP: Temporal Difference Side Network for Efficient VideoAction Recognition
di: Wang, Bin, et al.
Pubblicazione: (2024)
di: Wang, Bin, et al.
Pubblicazione: (2024)
Spiking Transformer:Introducing Accurate Addition-Only Spiking Self-Attention for Transformer
di: Guo, Yufei, et al.
Pubblicazione: (2025)
di: Guo, Yufei, et al.
Pubblicazione: (2025)
CT-SDM: A Sampling Diffusion Model for Sparse-View CT Reconstruction across All Sampling Rates
di: Yang, Liutao, et al.
Pubblicazione: (2024)
di: Yang, Liutao, et al.
Pubblicazione: (2024)
Masked Video and Body-worn IMU Autoencoder for Egocentric Action Recognition
di: Zhang, Mingfang, et al.
Pubblicazione: (2024)
di: Zhang, Mingfang, et al.
Pubblicazione: (2024)
Rethinking CLIP-based Video Learners in Cross-Domain Open-Vocabulary Action Recognition
di: Lin, Kun-Yu, et al.
Pubblicazione: (2024)
di: Lin, Kun-Yu, et al.
Pubblicazione: (2024)
SMTrack: End-to-End Trained Spiking Neural Networks for Multi-Object Tracking in RGB Videos
di: Zhong, Pengzhi, et al.
Pubblicazione: (2025)
di: Zhong, Pengzhi, et al.
Pubblicazione: (2025)
PruneVid: Visual Token Pruning for Efficient Video Large Language Models
di: Huang, Xiaohu, et al.
Pubblicazione: (2024)
di: Huang, Xiaohu, et al.
Pubblicazione: (2024)
View while Moving: Efficient Video Recognition in Long-untrimmed Videos
di: Tian, Ye, et al.
Pubblicazione: (2023)
di: Tian, Ye, et al.
Pubblicazione: (2023)
Enhanced Self-Distillation Framework for Efficient Spiking Neural Network Training
di: Zhao, Xiaochen, et al.
Pubblicazione: (2025)
di: Zhao, Xiaochen, et al.
Pubblicazione: (2025)
StegaVAR: Privacy-Preserving Video Action Recognition via Steganographic Domain Analysis
di: Chen, Lixin, et al.
Pubblicazione: (2025)
di: Chen, Lixin, et al.
Pubblicazione: (2025)
FROSTER: Frozen CLIP Is A Strong Teacher for Open-Vocabulary Action Recognition
di: Huang, Xiaohu, et al.
Pubblicazione: (2024)
di: Huang, Xiaohu, et al.
Pubblicazione: (2024)
Documenti analoghi
-
QKFormer: Hierarchical Spiking Transformer using Q-K Attention
di: Zhou, Chenlin, et al.
Pubblicazione: (2024) -
Direct Training High-Performance Deep Spiking Neural Networks: A Review of Theories and Methods
di: Zhou, Chenlin, et al.
Pubblicazione: (2024) -
Long-Range Feedback Spiking Network Captures Dynamic and Static Representations of the Visual Cortex under Movie Stimuli
di: Huang, Liwei, et al.
Pubblicazione: (2023) -
Spikingformer: A Key Foundation Model for Spiking Neural Networks
di: Zhou, Chenlin, et al.
Pubblicazione: (2023) -
SpikePoint: An Efficient Point-based Spiking Neural Network for Event Cameras Action Recognition
di: Ren, Hongwei, et al.
Pubblicazione: (2023)