SPACT18: Spiking Human Action Recognition Benchmark Dataset with Complementary RGB and Thermal Modalities
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ashraf, Yasser, Sharshar, Ahmed, Bojkovic, Velibor, Gu, Bin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Temporal Misalignment in ANN-SNN Conversion and Its Mitigation via Probabilistic Spiking Neurons
von: Bojković, Velibor, et al.
Veröffentlicht: (2025)
von: Bojković, Velibor, et al.
Veröffentlicht: (2025)
FTBC: Forward Temporal Bias Correction for Optimizing ANN-SNN Conversion
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2024)
PulmoFusion: Advancing Pulmonary Health with Efficient Multi-Modal Fusion
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025)
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025)
Not Only Grey Matter: OmniBrain for Robust Multimodal Classification of Alzheimer's Disease
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025)
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025)
Complementary Random Masking for RGB-Thermal Semantic Segmentation
von: Shin, Ukcheol, et al.
Veröffentlicht: (2023)
von: Shin, Ukcheol, et al.
Veröffentlicht: (2023)
Human Activity Recognition using RGB-Event based Sensors: A Multi-modal Heat Conduction Model and A Benchmark Dataset
von: Wang, Shiao, et al.
Veröffentlicht: (2025)
von: Wang, Shiao, et al.
Veröffentlicht: (2025)
GeoLLaVA: Efficient Fine-Tuned Vision-Language Models for Temporal Change Detection in Remote Sensing
von: Elgendy, Hosam, et al.
Veröffentlicht: (2024)
von: Elgendy, Hosam, et al.
Veröffentlicht: (2024)
An End-to-End Two-Stream Network Based on RGB Flow and Representation Flow for Human Action Recognition
von: Lai, Song-Jiang, et al.
Veröffentlicht: (2024)
von: Lai, Song-Jiang, et al.
Veröffentlicht: (2024)
RGB-Event based Pedestrian Attribute Recognition: A Benchmark Dataset and An Asymmetric RWKV Fusion Framework
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
Fire on Motion: Optimizing Video Pass-bands for Efficient Spiking Action Recognition
von: Ye, Shuhan, et al.
Veröffentlicht: (2026)
von: Ye, Shuhan, et al.
Veröffentlicht: (2026)
SigFormer: Sparse Signal-Guided Transformer for Multi-Modal Human Action Segmentation
von: Liu, Qi, et al.
Veröffentlicht: (2023)
von: Liu, Qi, et al.
Veröffentlicht: (2023)
Event Stream based Human Action Recognition: A High-Definition Benchmark Dataset and Algorithms
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
von: Wang, Xiao, et al.
Veröffentlicht: (2024)
Vision-Language Models for Edge Networks: A Comprehensive Survey
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025)
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025)
Signal-SGN++: Topology-Enhanced Time-Frequency Spiking Graph Network for Skeleton-Based Action Recognition
von: Zheng, Naichuan, et al.
Veröffentlicht: (2025)
von: Zheng, Naichuan, et al.
Veröffentlicht: (2025)
S3T-Former: A Purely Spike-Driven State-Space Topology Transformer for Skeleton Action Recognition
von: Zheng, Naichuan, et al.
Veröffentlicht: (2026)
von: Zheng, Naichuan, et al.
Veröffentlicht: (2026)
MMeViT: Multi-Modal ensemble ViT for Post-Stroke Rehabilitation Action Recognition
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)
Conformal Predictions for Human Action Recognition with Vision-Language Models
von: Tim, Bary, et al.
Veröffentlicht: (2025)
von: Tim, Bary, et al.
Veröffentlicht: (2025)
Temporal-Guided Spiking Neural Networks for Event-Based Human Action Recognition
von: Yang, Siyuan, et al.
Veröffentlicht: (2025)
von: Yang, Siyuan, et al.
Veröffentlicht: (2025)
Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
von: Tse, Tze Ho Elden, et al.
Veröffentlicht: (2025)
von: Tse, Tze Ho Elden, et al.
Veröffentlicht: (2025)
CMD-HAR: Cross-Modal Disentanglement for Wearable Human Activity Recognition
von: Yu, Ying, et al.
Veröffentlicht: (2025)
von: Yu, Ying, et al.
Veröffentlicht: (2025)
Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion
von: Singh, Shivam, et al.
Veröffentlicht: (2026)
von: Singh, Shivam, et al.
Veröffentlicht: (2026)
Complex Mathematical Expression Recognition: Benchmark, Large-Scale Dataset and Strong Baseline
von: Bai, Weikang, et al.
Veröffentlicht: (2025)
von: Bai, Weikang, et al.
Veröffentlicht: (2025)
Segment Any RGB-Thermal Model with Language-aided Distillation
von: Xing, Dong, et al.
Veröffentlicht: (2025)
von: Xing, Dong, et al.
Veröffentlicht: (2025)
TAG: Thinking with Action Unit Grounding for Facial Expression Recognition
von: Lin, Haobo, et al.
Veröffentlicht: (2026)
von: Lin, Haobo, et al.
Veröffentlicht: (2026)
EPAM-Net: An Efficient Pose-driven Attention-guided Multimodal Network for Video Action Recognition
von: Abdelkawy, Ahmed, et al.
Veröffentlicht: (2024)
von: Abdelkawy, Ahmed, et al.
Veröffentlicht: (2024)
Interpretable Action Recognition on Hard to Classify Actions
von: Anichenko, Anastasia, et al.
Veröffentlicht: (2024)
von: Anichenko, Anastasia, et al.
Veröffentlicht: (2024)
Rethinking RGB-D Fusion for Semantic Segmentation in Surgical Datasets
von: Jamal, Muhammad Abdullah, et al.
Veröffentlicht: (2024)
von: Jamal, Muhammad Abdullah, et al.
Veröffentlicht: (2024)
Pedestrian Attribute Recognition via Hierarchical Cross-Modality HyperGraph Learning
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
HalluciDet: Hallucinating RGB Modality for Person Detection Through Privileged Information
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2023)
von: Medeiros, Heitor Rapela, et al.
Veröffentlicht: (2023)
MVIP -- A Dataset and Methods for Application Oriented Multi-View and Multi-Modal Industrial Part Recognition
von: Koch, Paul, et al.
Veröffentlicht: (2025)
von: Koch, Paul, et al.
Veröffentlicht: (2025)
Lens Privacy Sealing: A New Benchmark and Method for Physical Privacy-Preserving Action Recognition
von: Liu, Mengyuan, et al.
Veröffentlicht: (2026)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2026)
Unleashing the Power of CNN and Transformer for Balanced RGB-Event Video Recognition
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
von: Wang, Xiao, et al.
Veröffentlicht: (2023)
Neurosymbolic Framework for Concept-Driven Logical Reasoning in Skeleton-Based Human Action Recognition
von: Ilyas, Talha, et al.
Veröffentlicht: (2026)
von: Ilyas, Talha, et al.
Veröffentlicht: (2026)
Enhanced Sparse Point Cloud Data Processing for Privacy-aware Human Action Recognition
von: Tunau, Maimunatu, et al.
Veröffentlicht: (2025)
von: Tunau, Maimunatu, et al.
Veröffentlicht: (2025)
SpikMamba: When SNN meets Mamba in Event-based Human Action Recognition
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
von: Chen, Jiaqi, et al.
Veröffentlicht: (2024)
Human-AI Divergence in Ego-centric Action Recognition under Spatial and Spatiotemporal Manipulations
von: Rahmaniboldaji, Sadegh, et al.
Veröffentlicht: (2026)
von: Rahmaniboldaji, Sadegh, et al.
Veröffentlicht: (2026)
IMPACT: A Dataset for Multi-Granularity Human Procedural Action Understanding in Industrial Assembly
von: Wen, Di, et al.
Veröffentlicht: (2026)
von: Wen, Di, et al.
Veröffentlicht: (2026)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
von: Zahid, Azizul, et al.
Veröffentlicht: (2025)
von: Zahid, Azizul, et al.
Veröffentlicht: (2025)
Real-Time Human Action Recognition on Embedded Platforms
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
von: Wang, Ruiqi, et al.
Veröffentlicht: (2024)
Grounding Foundational Vision Models with 3D Human Poses for Robust Action Recognition
von: Babey, Nicholas, et al.
Veröffentlicht: (2025)
von: Babey, Nicholas, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Temporal Misalignment in ANN-SNN Conversion and Its Mitigation via Probabilistic Spiking Neurons
von: Bojković, Velibor, et al.
Veröffentlicht: (2025) -
FTBC: Forward Temporal Bias Correction for Optimizing ANN-SNN Conversion
von: Wu, Xiaofeng, et al.
Veröffentlicht: (2024) -
PulmoFusion: Advancing Pulmonary Health with Efficient Multi-Modal Fusion
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025) -
Not Only Grey Matter: OmniBrain for Robust Multimodal Classification of Alzheimer's Disease
von: Sharshar, Ahmed, et al.
Veröffentlicht: (2025) -
Complementary Random Masking for RGB-Thermal Semantic Segmentation
von: Shin, Ukcheol, et al.
Veröffentlicht: (2023)