Efficient Spatial-Temporal Focal Adapter with SSM for Temporal Action Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qiu, Yicheng, Yanai, Keiji |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AM Flow: Adapters for Temporal Processing in Action Recognition
von: Agrawal, Tanay, et al.
Veröffentlicht: (2024)
von: Agrawal, Tanay, et al.
Veröffentlicht: (2024)
Point-MF: One-step Point Cloud Generation from a Single Image via Mean Flows
von: Baba, Yuta, et al.
Veröffentlicht: (2026)
von: Baba, Yuta, et al.
Veröffentlicht: (2026)
FDDet: Frequency-Decoupling for Boundary Refinement in Temporal Action Detection
von: Zhu, Xinnan, et al.
Veröffentlicht: (2025)
von: Zhu, Xinnan, et al.
Veröffentlicht: (2025)
StretchySnake: Flexible SSM Training Unlocks Action Recognition Across Spatio-Temporal Scales
von: Siddiqui, Nyle, et al.
Veröffentlicht: (2025)
von: Siddiqui, Nyle, et al.
Veröffentlicht: (2025)
D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition
von: Pei, Wenjie, et al.
Veröffentlicht: (2023)
von: Pei, Wenjie, et al.
Veröffentlicht: (2023)
Described Spatial-Temporal Video Detection
von: Ji, Wei, et al.
Veröffentlicht: (2024)
von: Ji, Wei, et al.
Veröffentlicht: (2024)
HOI-R1: Exploring the Potential of Multimodal Large Language Models for Human-Object Interaction Detection
von: Chen, Junwen, et al.
Veröffentlicht: (2025)
von: Chen, Junwen, et al.
Veröffentlicht: (2025)
Focusing on what to decode and what to train: SOV Decoding with Specific Target Guided DeNoising and Vision Language Advisor
von: Chen, Junwen, et al.
Veröffentlicht: (2023)
von: Chen, Junwen, et al.
Veröffentlicht: (2023)
Harnessing Temporal Causality for Advanced Temporal Action Detection
von: Liu, Shuming, et al.
Veröffentlicht: (2024)
von: Liu, Shuming, et al.
Veröffentlicht: (2024)
SceneTextStylizer: A Training-Free Scene Text Style Transfer Framework with Diffusion Model
von: Yuan, Honghui, et al.
Veröffentlicht: (2025)
von: Yuan, Honghui, et al.
Veröffentlicht: (2025)
LiquidTAD: Efficient Temporal Action Detection via Parallel Liquid-Inspired Temporal Relaxation
von: Sun, Zepeng, et al.
Veröffentlicht: (2026)
von: Sun, Zepeng, et al.
Veröffentlicht: (2026)
DVFL-Net: A Lightweight Distilled Video Focal Modulation Network for Spatio-Temporal Action Recognition
von: Ullah, Hayat, et al.
Veröffentlicht: (2025)
von: Ullah, Hayat, et al.
Veröffentlicht: (2025)
LoSA: Long-Short-range Adapter for Scaling End-to-End Temporal Action Localization
von: Gupta, Akshita, et al.
Veröffentlicht: (2024)
von: Gupta, Akshita, et al.
Veröffentlicht: (2024)
Benchmarking the Robustness of Temporal Action Detection Models Against Temporal Corruptions
von: Zeng, Runhao, et al.
Veröffentlicht: (2024)
von: Zeng, Runhao, et al.
Veröffentlicht: (2024)
T-Gated Adapter: A Lightweight Temporal Adapter for Vision-Language Medical Segmentation
von: Khadka, Pranjal
Veröffentlicht: (2026)
von: Khadka, Pranjal
Veröffentlicht: (2026)
Introducing Gating and Context into Temporal Action Detection
von: Reka, Aglind, et al.
Veröffentlicht: (2024)
von: Reka, Aglind, et al.
Veröffentlicht: (2024)
Prediction-Feedback DETR for Temporal Action Detection
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
Boundary-Recovering Network for Temporal Action Detection
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
Open-Vocabulary Spatio-Temporal Action Detection
von: Wu, Tao, et al.
Veröffentlicht: (2024)
von: Wu, Tao, et al.
Veröffentlicht: (2024)
Decoupling Spatio-Temporal Adapter for Fine-Grained Badminton Action Localization
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
von: Wang, Tianyu, et al.
Veröffentlicht: (2026)
BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing
von: Xiong, Peilin, et al.
Veröffentlicht: (2026)
von: Xiong, Peilin, et al.
Veröffentlicht: (2026)
PosBridge: Multi-View Positional Embedding Transplant for Identity-Aware Image Editing
von: Xiong, Peilin, et al.
Veröffentlicht: (2025)
von: Xiong, Peilin, et al.
Veröffentlicht: (2025)
One-Stage Open-Vocabulary Temporal Action Detection Leveraging Temporal Multi-scale and Action Label Features
von: Nguyen, Trung Thanh, et al.
Veröffentlicht: (2024)
von: Nguyen, Trung Thanh, et al.
Veröffentlicht: (2024)
SEM-Net: Efficient Pixel Modelling for image inpainting with Spatially Enhanced SSM
von: Chen, Shuang, et al.
Veröffentlicht: (2024)
von: Chen, Shuang, et al.
Veröffentlicht: (2024)
When Spatial meets Temporal in Action Recognition
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
von: Chen, Huilin, et al.
Veröffentlicht: (2024)
Discovering Intrinsic Spatial-Temporal Logic Rules to Explain Human Actions
von: Cao, Chengzhi, et al.
Veröffentlicht: (2023)
von: Cao, Chengzhi, et al.
Veröffentlicht: (2023)
Spatial-Temporal Perception with Causal Inference for Naturalistic Driving Action Recognition
von: Chang, Qing, et al.
Veröffentlicht: (2025)
von: Chang, Qing, et al.
Veröffentlicht: (2025)
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
von: Su, Rui, et al.
Veröffentlicht: (2019)
von: Su, Rui, et al.
Veröffentlicht: (2019)
Improving Viewpoint-Invariance and Temporal Consistency for Action Detection
von: Porto, Yannick, et al.
Veröffentlicht: (2026)
von: Porto, Yannick, et al.
Veröffentlicht: (2026)
Dual DETRs for Multi-Label Temporal Action Detection
von: Zhu, Yuhan, et al.
Veröffentlicht: (2024)
von: Zhu, Yuhan, et al.
Veröffentlicht: (2024)
FocalOrder: Focal Preference Optimization for Reading Order Detection
von: Liu, Fuyuan, et al.
Veröffentlicht: (2026)
von: Liu, Fuyuan, et al.
Veröffentlicht: (2026)
Are Spatial-Temporal Graph Convolution Networks for Human Action Recognition Over-Parameterized?
von: Xie, Jianyang, et al.
Veröffentlicht: (2025)
von: Xie, Jianyang, et al.
Veröffentlicht: (2025)
Joint Image-Instance Spatial-Temporal Attention for Few-shot Action Recognition
von: Qian, Zefeng, et al.
Veröffentlicht: (2025)
von: Qian, Zefeng, et al.
Veröffentlicht: (2025)
Temporal Action Detection Model Compression by Progressive Block Drop
von: Chen, Xiaoyong, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyong, et al.
Veröffentlicht: (2025)
Boundary Discretization and Reliable Classification Network for Temporal Action Detection
von: Fang, Zhenying, et al.
Veröffentlicht: (2023)
von: Fang, Zhenying, et al.
Veröffentlicht: (2023)
DyFADet: Dynamic Feature Aggregation for Temporal Action Detection
von: Yang, Le, et al.
Veröffentlicht: (2024)
von: Yang, Le, et al.
Veröffentlicht: (2024)
Long-term Pre-training for Temporal Action Detection with Transformers
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
Spatial-Temporal Human-Object Interaction Detection
von: Sun, Xu, et al.
Veröffentlicht: (2025)
von: Sun, Xu, et al.
Veröffentlicht: (2025)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
von: Han, Jiwook, et al.
Veröffentlicht: (2026)
von: Han, Jiwook, et al.
Veröffentlicht: (2026)
EEdit: Rethinking the Spatial and Temporal Redundancy for Efficient Image Editing
von: Yan, Zexuan, et al.
Veröffentlicht: (2025)
von: Yan, Zexuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
AM Flow: Adapters for Temporal Processing in Action Recognition
von: Agrawal, Tanay, et al.
Veröffentlicht: (2024) -
Point-MF: One-step Point Cloud Generation from a Single Image via Mean Flows
von: Baba, Yuta, et al.
Veröffentlicht: (2026) -
FDDet: Frequency-Decoupling for Boundary Refinement in Temporal Action Detection
von: Zhu, Xinnan, et al.
Veröffentlicht: (2025) -
StretchySnake: Flexible SSM Training Unlocks Action Recognition Across Spatio-Temporal Scales
von: Siddiqui, Nyle, et al.
Veröffentlicht: (2025) -
D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition
von: Pei, Wenjie, et al.
Veröffentlicht: (2023)