Action-Dynamics Modeling and Cross-Temporal Interaction for Online Action Understanding
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yang, Xinyu, Jiang, Zheheng, Zhou, Feixiang, Zhu, Yihang, Lv, Na, Xing, Nan, Canagarajah, Nishan, Zhou, Huiyu |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
SMC-NCA: Semantic-guided Multi-level Contrast for Semi-supervised Temporal Action Segmentation
par: Zhou, Feixiang, et autres
Publié: (2023)
par: Zhou, Feixiang, et autres
Publié: (2023)
Cross-Skeleton Interaction Graph Aggregation Network for Representation Learning of Mouse Social Behaviour
par: Zhou, Feixiang, et autres
Publié: (2022)
par: Zhou, Feixiang, et autres
Publié: (2022)
Towards Adaptive Pseudo-label Learning for Semi-Supervised Temporal Action Localization
par: Zhou, Feixiang, et autres
Publié: (2024)
par: Zhou, Feixiang, et autres
Publié: (2024)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
par: Xie, Liping, et autres
Publié: (2025)
par: Xie, Liping, et autres
Publié: (2025)
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
par: Su, Rui, et autres
Publié: (2019)
par: Su, Rui, et autres
Publié: (2019)
OnlineTAS: An Online Baseline for Temporal Action Segmentation
par: Zhong, Qing, et autres
Publié: (2024)
par: Zhong, Qing, et autres
Publié: (2024)
Modelling Spatio-Temporal Interactions For Compositional Action Recognition
par: Rajendiran, Ramanathan, et autres
Publié: (2023)
par: Rajendiran, Ramanathan, et autres
Publié: (2023)
GigaWorld-Policy: An Efficient Action-Centered World--Action Model
par: Ye, Angen, et autres
Publié: (2026)
par: Ye, Angen, et autres
Publié: (2026)
Online Temporal Action Localization with Memory-Augmented Transformer
par: Song, Youngkil, et autres
Publié: (2024)
par: Song, Youngkil, et autres
Publié: (2024)
TEMPURA: Temporal Event Masked Prediction and Understanding for Reasoning in Action
par: Cheng, Jen-Hao, et autres
Publié: (2025)
par: Cheng, Jen-Hao, et autres
Publié: (2025)
F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions
par: Lv, Qi, et autres
Publié: (2025)
par: Lv, Qi, et autres
Publié: (2025)
Recovering Complete Actions for Cross-dataset Skeleton Action Recognition
par: Liu, Hanchao, et autres
Publié: (2024)
par: Liu, Hanchao, et autres
Publié: (2024)
OZ-TAL: Online Zero-Shot Temporal Action Localization
par: Han, Chaolei, et autres
Publié: (2026)
par: Han, Chaolei, et autres
Publié: (2026)
Interaction-via-Actions: Cattle Interaction Detection with Joint Learning of Action-Interaction Latent Space
par: Nakagawa, Ren, et autres
Publié: (2025)
par: Nakagawa, Ren, et autres
Publié: (2025)
The Role of Video Generation in Enhancing Data-Limited Action Understanding
par: Li, Wei, et autres
Publié: (2025)
par: Li, Wei, et autres
Publié: (2025)
DyFADet: Dynamic Feature Aggregation for Temporal Action Detection
par: Yang, Le, et autres
Publié: (2024)
par: Yang, Le, et autres
Publié: (2024)
Temporal Action Localization with Cross Layer Task Decoupling and Refinement
par: Li, Qiang, et autres
Publié: (2024)
par: Li, Qiang, et autres
Publié: (2024)
Repetitive Action Counting with Hybrid Temporal Relation Modeling
par: Li, Kun, et autres
Publié: (2024)
par: Li, Kun, et autres
Publié: (2024)
HAT: History-Augmented Anchor Transformer for Online Temporal Action Localization
par: Reza, Sakib, et autres
Publié: (2024)
par: Reza, Sakib, et autres
Publié: (2024)
Probing Fine-Grained Action Understanding and Cross-View Generalization of Foundation Models
par: Ponbagavathi, Thinesh Thiyakesan, et autres
Publié: (2024)
par: Ponbagavathi, Thinesh Thiyakesan, et autres
Publié: (2024)
ActAvatar: Temporally-Aware Precise Action Control for Talking Avatars
par: Peng, Ziqiao, et autres
Publié: (2025)
par: Peng, Ziqiao, et autres
Publié: (2025)
MALT: Multi-scale Action Learning Transformer for Online Action Detection
par: Yang, Zhipeng, et autres
Publié: (2024)
par: Yang, Zhipeng, et autres
Publié: (2024)
RotVLA: Rotational Latent Action for Vision-Language-Action Model
par: Li, Qiwei, et autres
Publié: (2026)
par: Li, Qiwei, et autres
Publié: (2026)
Improving Weakly Supervised Temporal Action Localization by Exploiting Multi-resolution Information in Temporal Domain
par: Su, Rui, et autres
Publié: (2025)
par: Su, Rui, et autres
Publié: (2025)
SkeFi: Cross-Modal Knowledge Transfer for Wireless Skeleton-Based Action Recognition
par: Huang, Shunyu, et autres
Publié: (2026)
par: Huang, Shunyu, et autres
Publié: (2026)
TAMT: Temporal-Aware Model Tuning for Cross-Domain Few-Shot Action Recognition
par: Wang, Yilong, et autres
Publié: (2024)
par: Wang, Yilong, et autres
Publié: (2024)
ActionHub: A Large-scale Action Video Description Dataset for Zero-shot Action Recognition
par: Zhou, Jiaming, et autres
Publié: (2024)
par: Zhou, Jiaming, et autres
Publié: (2024)
The DAWN of World-Action Interactive Models
par: Lu, Hongbo, et autres
Publié: (2026)
par: Lu, Hongbo, et autres
Publié: (2026)
Understanding the Cross-Domain Capabilities of Video-Based Few-Shot Action Recognition Models
par: Markham, Georgia, et autres
Publié: (2024)
par: Markham, Georgia, et autres
Publié: (2024)
CLIP-AE: CLIP-assisted Cross-view Audio-Visual Enhancement for Unsupervised Temporal Action Localization
par: Xia, Rui, et autres
Publié: (2025)
par: Xia, Rui, et autres
Publié: (2025)
Insights from Visual Cognition: Understanding Human Action Dynamics with Overall Glance and Refined Gaze Transformer
par: Xing, Bohao, et autres
Publié: (2026)
par: Xing, Bohao, et autres
Publié: (2026)
FDDet: Frequency-Decoupling for Boundary Refinement in Temporal Action Detection
par: Zhu, Xinnan, et autres
Publié: (2025)
par: Zhu, Xinnan, et autres
Publié: (2025)
Cross-view Action Recognition Understanding From Exocentric to Egocentric Perspective
par: Truong, Thanh-Dat, et autres
Publié: (2023)
par: Truong, Thanh-Dat, et autres
Publié: (2023)
Generative Hierarchical Temporal Transformer for Hand Pose and Action Modeling
par: Wen, Yilin, et autres
Publié: (2023)
par: Wen, Yilin, et autres
Publié: (2023)
Dual DETRs for Multi-Label Temporal Action Detection
par: Zhu, Yuhan, et autres
Publié: (2024)
par: Zhu, Yuhan, et autres
Publié: (2024)
Benchmarking the Robustness of Temporal Action Detection Models Against Temporal Corruptions
par: Zeng, Runhao, et autres
Publié: (2024)
par: Zeng, Runhao, et autres
Publié: (2024)
PrevPredMap: Exploring Temporal Modeling with Previous Predictions for Online Vectorized HD Map Construction
par: Peng, Nan, et autres
Publié: (2024)
par: Peng, Nan, et autres
Publié: (2024)
ActionArt: Advancing Multimodal Large Models for Fine-Grained Human-Centric Video Understanding
par: Peng, Yi-Xing, et autres
Publié: (2025)
par: Peng, Yi-Xing, et autres
Publié: (2025)
The Solution for Temporal Action Localisation Task of Perception Test Challenge 2024
par: Han, Yinan, et autres
Publié: (2024)
par: Han, Yinan, et autres
Publié: (2024)
Aligning Neuronal Coding of Dynamic Visual Scenes with Foundation Vision Models
par: Wu, Rining, et autres
Publié: (2024)
par: Wu, Rining, et autres
Publié: (2024)
Documents similaires
-
SMC-NCA: Semantic-guided Multi-level Contrast for Semi-supervised Temporal Action Segmentation
par: Zhou, Feixiang, et autres
Publié: (2023) -
Cross-Skeleton Interaction Graph Aggregation Network for Representation Learning of Mouse Social Behaviour
par: Zhou, Feixiang, et autres
Publié: (2022) -
Towards Adaptive Pseudo-label Learning for Semi-Supervised Temporal Action Localization
par: Zhou, Feixiang, et autres
Publié: (2024) -
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
par: Xie, Liping, et autres
Publié: (2025) -
Progressive Cross-Stream Cooperation in Spatial and Temporal Domain for Action Localization
par: Su, Rui, et autres
Publié: (2019)