Action-Guided Attention for Video Action Anticipation
Fuente:
arXiv
Guardado en:
| Autores principales: | Tai, Tsung-Ming, Casarin, Sofia, Pilzer, Andrea, Nutt, Werner, Lanz, Oswald |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
L-SWAG: Layer-Sample Wise Activation with Gradients information for Zero-Shot NAS on Vision Transformers
por: Casarin, Sofia, et al.
Publicado: (2025)
por: Casarin, Sofia, et al.
Publicado: (2025)
GRASP-GCN: Graph-Shape Prioritization for Neural Architecture Search under Distribution Shifts
por: Casarin, Sofia, et al.
Publicado: (2024)
por: Casarin, Sofia, et al.
Publicado: (2024)
Your Image is My Video: Reshaping the Receptive Field via Image-To-Video Differentiable AutoAugmentation and Fusion
por: Casarin, Sofia, et al.
Publicado: (2024)
por: Casarin, Sofia, et al.
Publicado: (2024)
Gate-Shift-Pose: Enhancing Action Recognition in Sports with Skeleton Information
por: Bianchi, Edoardo, et al.
Publicado: (2025)
por: Bianchi, Edoardo, et al.
Publicado: (2025)
Semantically Guided Action Anticipation
por: Diko, Anxhelo, et al.
Publicado: (2024)
por: Diko, Anxhelo, et al.
Publicado: (2024)
Intention Action Anticipation Model with Guide-Feedback Loop Mechanism
por: Ma, Zongnan, et al.
Publicado: (2024)
por: Ma, Zongnan, et al.
Publicado: (2024)
Fractals as Pre-training Datasets for Anomaly Detection and Localization
por: Ugwu, C. I., et al.
Publicado: (2024)
por: Ugwu, C. I., et al.
Publicado: (2024)
Semantically Guided Representation Learning For Action Anticipation
por: Diko, Anxhelo, et al.
Publicado: (2024)
por: Diko, Anxhelo, et al.
Publicado: (2024)
Multimodal Large Models Are Effective Action Anticipators
por: Wang, Binglu, et al.
Publicado: (2025)
por: Wang, Binglu, et al.
Publicado: (2025)
Action-guided generation of 3D functionality segmentation data
por: Corsetti, Jaime, et al.
Publicado: (2025)
por: Corsetti, Jaime, et al.
Publicado: (2025)
Action Anticipation at a Glimpse: To What Extent Can Multimodal Cues Replace Video?
por: Benavent-Lledo, Manuel, et al.
Publicado: (2025)
por: Benavent-Lledo, Manuel, et al.
Publicado: (2025)
Bidirectional Action Sequence Learning for Long-term Action Anticipation with Large Language Models
por: Sato, Yuji, et al.
Publicado: (2025)
por: Sato, Yuji, et al.
Publicado: (2025)
Intention-Guided Cognitive Reasoning for Egocentric Long-Term Action Anticipation
por: Chu, Qiaohui, et al.
Publicado: (2025)
por: Chu, Qiaohui, et al.
Publicado: (2025)
Understanding Multimodal Complementarity for Single-Frame Action Anticipation
por: Benavent-Lledo, Manuel, et al.
Publicado: (2026)
por: Benavent-Lledo, Manuel, et al.
Publicado: (2026)
Interaction Region Visual Transformer for Egocentric Action Anticipation
por: Roy, Debaditya, et al.
Publicado: (2022)
por: Roy, Debaditya, et al.
Publicado: (2022)
A Survey on Deep Learning Techniques for Action Anticipation
por: Zhong, Zeyun, et al.
Publicado: (2023)
por: Zhong, Zeyun, et al.
Publicado: (2023)
Human Action Anticipation: A Survey
por: Lai, Bolin, et al.
Publicado: (2024)
por: Lai, Bolin, et al.
Publicado: (2024)
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
por: Liu, Xin, et al.
Publicado: (2024)
por: Liu, Xin, et al.
Publicado: (2024)
Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition
por: Guo, Hanyu, et al.
Publicado: (2024)
por: Guo, Hanyu, et al.
Publicado: (2024)
AntGPT: Can Large Language Models Help Long-term Action Anticipation from Videos?
por: Zhao, Qi, et al.
Publicado: (2023)
por: Zhao, Qi, et al.
Publicado: (2023)
Multi-level and Multi-modal Action Anticipation
por: Kim, Seulgi, et al.
Publicado: (2025)
por: Kim, Seulgi, et al.
Publicado: (2025)
Gaze-Guided Graph Neural Network for Action Anticipation Conditioned on Intention
por: Ozdel, Suleyman, et al.
Publicado: (2024)
por: Ozdel, Suleyman, et al.
Publicado: (2024)
Language Model Guided Interpretable Video Action Reasoning
por: Wang, Ning, et al.
Publicado: (2024)
por: Wang, Ning, et al.
Publicado: (2024)
Can't make an Omelette without Breaking some Eggs: Plausible Action Anticipation using Large Video-Language Models
por: Mittal, Himangi, et al.
Publicado: (2024)
por: Mittal, Himangi, et al.
Publicado: (2024)
MixANT: Observation-dependent Memory Propagation for Stochastic Dense Action Anticipation
por: Wasim, Syed Talal, et al.
Publicado: (2025)
por: Wasim, Syed Talal, et al.
Publicado: (2025)
See It Before You Grab It: Deep Learning-based Action Anticipation in Basketball
por: Roy, Arnau Barrera, et al.
Publicado: (2025)
por: Roy, Arnau Barrera, et al.
Publicado: (2025)
Vision and Intention Boost Large Language Model in Long-Term Action Anticipation
por: Cao, Congqi, et al.
Publicado: (2025)
por: Cao, Congqi, et al.
Publicado: (2025)
CAST: Cross-Attention in Space and Time for Video Action Recognition
por: Lee, Dongho, et al.
Publicado: (2023)
por: Lee, Dongho, et al.
Publicado: (2023)
Recovering Complete Actions for Cross-dataset Skeleton Action Recognition
por: Liu, Hanchao, et al.
Publicado: (2024)
por: Liu, Hanchao, et al.
Publicado: (2024)
Action Anticipation from SoccerNet Football Video Broadcasts
por: Dalal, Mohamad, et al.
Publicado: (2025)
por: Dalal, Mohamad, et al.
Publicado: (2025)
ATA: Bridging Implicit Reasoning with Attention-Guided and Action-Guided Inference for Vision-Language Action Models
por: Yang, Cheng, et al.
Publicado: (2026)
por: Yang, Cheng, et al.
Publicado: (2026)
Generative Model-Based Feature Attention Module for Video Action Analysis
por: Wang, Guiqin, et al.
Publicado: (2025)
por: Wang, Guiqin, et al.
Publicado: (2025)
Classification Matters: Improving Video Action Detection with Class-Specific Attention
por: Lee, Jinsung, et al.
Publicado: (2024)
por: Lee, Jinsung, et al.
Publicado: (2024)
ActionVOS: Actions as Prompts for Video Object Segmentation
por: Ouyang, Liangyang, et al.
Publicado: (2024)
por: Ouyang, Liangyang, et al.
Publicado: (2024)
Differentiable Frequency-based Disentanglement for Aerial Video Action Recognition
por: Kothandaraman, Divya, et al.
Publicado: (2022)
por: Kothandaraman, Divya, et al.
Publicado: (2022)
JFAA: Technical Report for the EPIC-KITCHENS-100 Action Anticipation Challenge at EgoVis 2026
por: Chu, Qiaohui, et al.
Publicado: (2026)
por: Chu, Qiaohui, et al.
Publicado: (2026)
Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling
por: Kang, Taewon, et al.
Publicado: (2025)
por: Kang, Taewon, et al.
Publicado: (2025)
ActFusion: a Unified Diffusion Model for Action Segmentation and Anticipation
por: Gong, Dayoung, et al.
Publicado: (2024)
por: Gong, Dayoung, et al.
Publicado: (2024)
QueryMamba: A Mamba-Based Encoder-Decoder Architecture with a Statistical Verb-Noun Interaction Module for Video Action Forecasting @ Ego4D Long-Term Action Anticipation Challenge 2024
por: Zhong, Zeyun, et al.
Publicado: (2024)
por: Zhong, Zeyun, et al.
Publicado: (2024)
Ensembling Pruned Attention Heads For Uncertainty-Aware Efficient Transformers
por: Gabetni, Firas, et al.
Publicado: (2025)
por: Gabetni, Firas, et al.
Publicado: (2025)
Ejemplares similares
-
L-SWAG: Layer-Sample Wise Activation with Gradients information for Zero-Shot NAS on Vision Transformers
por: Casarin, Sofia, et al.
Publicado: (2025) -
GRASP-GCN: Graph-Shape Prioritization for Neural Architecture Search under Distribution Shifts
por: Casarin, Sofia, et al.
Publicado: (2024) -
Your Image is My Video: Reshaping the Receptive Field via Image-To-Video Differentiable AutoAugmentation and Fusion
por: Casarin, Sofia, et al.
Publicado: (2024) -
Gate-Shift-Pose: Enhancing Action Recognition in Sports with Skeleton Information
por: Bianchi, Edoardo, et al.
Publicado: (2025) -
Semantically Guided Action Anticipation
por: Diko, Anxhelo, et al.
Publicado: (2024)