HAtt-Flow: Hierarchical Attention-Flow Mechanism for Group Activity Scene Graph Generation in Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Chappa, Naga VS Raviteja, Nguyen, Pha, Le, Thi Hoang Ngan, Luu, Khoa |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
REACT: Recognize Every Action Everywhere All At Once
by: Chappa, Naga VS Raviteja, et al.
Published: (2023)
by: Chappa, Naga VS Raviteja, et al.
Published: (2023)
FLAASH: Flow-Attention Adaptive Semantic Hierarchical Fusion for Multi-Modal Tobacco Content Analysis
by: Chappa, Naga VS Raviteja, et al.
Published: (2024)
by: Chappa, Naga VS Raviteja, et al.
Published: (2024)
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024)
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024)
SoGAR: Self-supervised Spatiotemporal Attention-based Social Group Activity Recognition
by: Chappa, Naga VS Raviteja, et al.
Published: (2023)
by: Chappa, Naga VS Raviteja, et al.
Published: (2023)
HIG: Hierarchical Interlacement Graph Approach to Scene Graph Generation in Video Understanding
by: Nguyen, Trong-Thuan, et al.
Published: (2023)
by: Nguyen, Trong-Thuan, et al.
Published: (2023)
Public Health Advocacy Dataset: A Dataset of Tobacco Usage Videos from Social Media
by: Chappa, Naga VS Raviteja, et al.
Published: (2024)
by: Chappa, Naga VS Raviteja, et al.
Published: (2024)
HyperGLM: HyperGraph for Video Scene Graph Generation and Anticipation
by: Nguyen, Trong-Thuan, et al.
Published: (2024)
by: Nguyen, Trong-Thuan, et al.
Published: (2024)
THYME: Temporal Hierarchical-Cyclic Interactivity Modeling for Video Scene Graphs in Aerial Footage
by: Nguyen, Trong-Thuan, et al.
Published: (2025)
by: Nguyen, Trong-Thuan, et al.
Published: (2025)
DINTR: Tracking via Diffusion-based Interpolation
by: Nguyen, Pha, et al.
Published: (2024)
by: Nguyen, Pha, et al.
Published: (2024)
DEFEND: A Large-scale 1M Dataset and Foundation Model for Tobacco Addiction Prevention
by: Chappa, Naga VS Raviteja, et al.
Published: (2025)
by: Chappa, Naga VS Raviteja, et al.
Published: (2025)
Z-GMOT: Zero-shot Generic Multiple Object Tracking
by: Tran, Kim Hoang, et al.
Published: (2023)
by: Tran, Kim Hoang, et al.
Published: (2023)
CYCLO: Cyclic Graph Transformer Approach to Multi-Object Relationship Modeling in Aerial Videos
by: Nguyen, Trong-Thuan, et al.
Published: (2024)
by: Nguyen, Trong-Thuan, et al.
Published: (2024)
Micro-DualNet: Dual-Path Spatio-Temporal Network for Micro-Action Recognition
by: Chappa, Naga VS Raviteja, et al.
Published: (2026)
by: Chappa, Naga VS Raviteja, et al.
Published: (2026)
QLAM: A Quantum Long-Attention Memory Approach to Long-Sequence Token Modeling
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
by: Nguyen, Hoang-Quan, et al.
Published: (2026)
HENASY: Learning to Assemble Scene-Entities for Egocentric Video-Language Model
by: Vo, Khoa, et al.
Published: (2024)
by: Vo, Khoa, et al.
Published: (2024)
MANGO: Multimodal Attention-based Normalizing Flow Approach to Fusion Learning
by: Truong, Thanh-Dat, et al.
Published: (2025)
by: Truong, Thanh-Dat, et al.
Published: (2025)
Linguistically Informed Multimodal Fusion for Vietnamese Scene-Text Image Captioning: Dataset, Graph Framework, and Phonological Attention
by: Nguyen, Nhi Ngoc-Yen, et al.
Published: (2026)
by: Nguyen, Nhi Ngoc-Yen, et al.
Published: (2026)
Dual-View Optical Flow for 4D Micro-Expression Recognition - A Multi-Stream Fusion Attention Approach
by: Nguyen, Luu Tu, et al.
Published: (2026)
by: Nguyen, Luu Tu, et al.
Published: (2026)
FMANet: A Novel Dual-Phase Optical Flow Approach with Fusion Motion Attention Network for Robust Micro-expression Recognition
by: Nguyen, Luu Tu, et al.
Published: (2025)
by: Nguyen, Luu Tu, et al.
Published: (2025)
MOOSE: Pay Attention to Temporal Dynamics for Video Understanding via Optical Flows
by: Nguyen, Hong, et al.
Published: (2025)
by: Nguyen, Hong, et al.
Published: (2025)
Hierarchical Quantum Control Gates for Functional MRI Understanding
by: Nguyen, Xuan-Bac, et al.
Published: (2024)
by: Nguyen, Xuan-Bac, et al.
Published: (2024)
Multi-view Action Recognition via Directed Gromov-Wasserstein Discrepancy
by: Nguyen, Hoang-Quan, et al.
Published: (2024)
by: Nguyen, Hoang-Quan, et al.
Published: (2024)
FrameDiT: Diffusion Transformer with Matrix Attention for Efficient Video Generation
by: Le, Minh Khoa, et al.
Published: (2026)
by: Le, Minh Khoa, et al.
Published: (2026)
Amodal Instance Segmentation with Diffusion Shape Prior Estimation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
UNO: Unifying One-stage Video Scene Graph Generation via Object-Centric Visual Representation Learning
by: Le, Huy, et al.
Published: (2025)
by: Le, Huy, et al.
Published: (2025)
A Novel Combined Optical Flow Approach for Comprehensive Micro-Expression Recognition
by: Khuong, Vu Tram Anh, et al.
Published: (2025)
by: Khuong, Vu Tram Anh, et al.
Published: (2025)
FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding
by: Truong, Thanh-Dat, et al.
Published: (2023)
by: Truong, Thanh-Dat, et al.
Published: (2023)
FlowScene: Style-Consistent Indoor Scene Generation with Multimodal Graph Rectified Flow
by: Yang, Zhifei, et al.
Published: (2026)
by: Yang, Zhifei, et al.
Published: (2026)
ViConsFormer: Constituting Meaningful Phrases of Scene Texts using Transformer-based Method in Vietnamese Text-based Visual Question Answering
by: Nguyen, Nghia Hieu, et al.
Published: (2024)
by: Nguyen, Nghia Hieu, et al.
Published: (2024)
AHAN: Asymmetric Hierarchical Attention Network for Identical Twin Face Verification
by: Nguyen, Hoang-Nhat
Published: (2026)
by: Nguyen, Hoang-Nhat
Published: (2026)
BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance
by: Le, Huy, et al.
Published: (2025)
by: Le, Huy, et al.
Published: (2025)
TP-GMOT: Tracking Generic Multiple Object by Textual Prompt with Motion-Appearance Cost (MAC) SORT
by: Anh, Duy Le Dinh, et al.
Published: (2024)
by: Anh, Duy Le Dinh, et al.
Published: (2024)
ShapeFormer: Shape Prior Visible-to-Amodal Transformer-based Amodal Instance Segmentation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
Can We Build Scene Graphs, Not Classify Them? FlowSG: Progressive Image-Conditioned Scene Graph Generation with Flow Matching
by: Hu, Xin, et al.
Published: (2026)
by: Hu, Xin, et al.
Published: (2026)
Aion: Towards Hierarchical 4D Scene Graphs with Temporal Flow Dynamics
by: Catalano, Iacopo, et al.
Published: (2025)
by: Catalano, Iacopo, et al.
Published: (2025)
Unifying Global and Local Scene Entities Modelling for Precise Action Spotting
by: Tran, Kim Hoang, et al.
Published: (2024)
by: Tran, Kim Hoang, et al.
Published: (2024)
Motion-aware Contrastive Learning for Temporal Panoptic Scene Graph Generation
by: Nguyen, Thong Thanh, et al.
Published: (2024)
by: Nguyen, Thong Thanh, et al.
Published: (2024)
NeurFlow: Interpreting Neural Networks through Neuron Groups and Functional Interactions
by: Cao, Tue M., et al.
Published: (2025)
by: Cao, Tue M., et al.
Published: (2025)
A Novel Dataset for Video-Based Neurodivergent Classification Leveraging Extra-Stimulatory Behavior
by: Serna-Aguilera, Manuel, et al.
Published: (2024)
by: Serna-Aguilera, Manuel, et al.
Published: (2024)
Apex-Centered Spatio-Temporal Rank Pooling and Gradient Attention for Micro-Expression Recognition
by: Nguyen, Luu Tu, et al.
Published: (2025)
by: Nguyen, Luu Tu, et al.
Published: (2025)
Similar Items
-
REACT: Recognize Every Action Everywhere All At Once
by: Chappa, Naga VS Raviteja, et al.
Published: (2023) -
FLAASH: Flow-Attention Adaptive Semantic Hierarchical Fusion for Multi-Modal Tobacco Content Analysis
by: Chappa, Naga VS Raviteja, et al.
Published: (2024) -
LiGAR: LiDAR-Guided Hierarchical Transformer for Multi-Modal Group Activity Recognition
by: Chappa, Naga Venkata Sai Raviteja, et al.
Published: (2024) -
SoGAR: Self-supervised Spatiotemporal Attention-based Social Group Activity Recognition
by: Chappa, Naga VS Raviteja, et al.
Published: (2023) -
HIG: Hierarchical Interlacement Graph Approach to Scene Graph Generation in Video Understanding
by: Nguyen, Trong-Thuan, et al.
Published: (2023)