Attention-Guided Dual-Stream Learning for Group Engagement Recognition: Fusing Transformer-Encoded Motion Dynamics with Scene Context via Adaptive Gating
Fuente:
arXiv
Saved in:
| Main Authors: | Chowdhury, Saniah Kayenat, Chowdhury, Muhammad E. H. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AnatomicalNets: A Multi-Structure Segmentation and Contour-Based Distance Estimation Pipeline for Clinically Grounded Lung Cancer T-Staging
by: Chowdhury, Saniah Kayenat, et al.
Published: (2025)
by: Chowdhury, Saniah Kayenat, et al.
Published: (2025)
Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition
by: Guo, Hanyu, et al.
Published: (2024)
by: Guo, Hanyu, et al.
Published: (2024)
Look Beyond Saliency: Low-Attention Guided Dual Encoding for Video Semantic Search
by: Aljehrai, Faisal, et al.
Published: (2026)
by: Aljehrai, Faisal, et al.
Published: (2026)
Deep Neural Network-Based Sign Language Recognition: A Comprehensive Approach Using Transfer Learning with Explainability
by: Ridwan, A. E. M, et al.
Published: (2024)
by: Ridwan, A. E. M, et al.
Published: (2024)
Automated Detection of Mutual Gaze and Joint Attention in Dual-Camera Settings via Dual-Stream Transformers
by: Kosmydel, Jakub, et al.
Published: (2026)
by: Kosmydel, Jakub, et al.
Published: (2026)
Learning Context-Adaptive Motion Priors for Masked Motion Diffusion Models with Efficient Kinematic Attention Aggregation
by: Jiang, Junkun, et al.
Published: (2026)
by: Jiang, Junkun, et al.
Published: (2026)
DSAGL: Dual-Stream Attention-Guided Learning for Weakly Supervised Whole Slide Image Classification
by: Cao, Daoxi, et al.
Published: (2025)
by: Cao, Daoxi, et al.
Published: (2025)
Incorporating Scene Context and Semantic Labels for Enhanced Group-level Emotion Recognition
by: Zhu, Qing, et al.
Published: (2025)
by: Zhu, Qing, et al.
Published: (2025)
Hyperspectral Image Classification via Transformer-based Spectral-Spatial Attention Decoupling and Adaptive Gating
by: Li, Guandong, et al.
Published: (2025)
by: Li, Guandong, et al.
Published: (2025)
Adversarial Machine Learning: Attacking and Safeguarding Image Datasets
by: Chowdhury, Koushik
Published: (2025)
by: Chowdhury, Koushik
Published: (2025)
Video-Based MPAA Rating Prediction: An Attention-Driven Hybrid Architecture Using Contrastive Learning
by: Neogi, Dipta, et al.
Published: (2025)
by: Neogi, Dipta, et al.
Published: (2025)
Uncertainty-Aware Diffusion Guided Refinement of 3D Scenes
by: Bose, Sarosij, et al.
Published: (2025)
by: Bose, Sarosij, et al.
Published: (2025)
AdaptoVision: A Multi-Resolution Image Recognition Model for Robust and Scalable Classification
by: Sabrin, Md. Sanaullah Chowdhury Lameya
Published: (2025)
by: Sabrin, Md. Sanaullah Chowdhury Lameya
Published: (2025)
Space-Time Forecasting of Dynamic Scenes with Motion-aware Gaussian Grouping
by: Lee, Junmyeong, et al.
Published: (2026)
by: Lee, Junmyeong, et al.
Published: (2026)
Making Every Frame Matter: Continuous Activity Recognition in Streaming Video via Adaptive Video Context Modeling
by: Wu, Hao, et al.
Published: (2024)
by: Wu, Hao, et al.
Published: (2024)
Content Adaptive Encoding For Interactive Game Streaming
by: Soltanayev, Shakarim, et al.
Published: (2025)
by: Soltanayev, Shakarim, et al.
Published: (2025)
MOSA: Motion-Guided Semantic Alignment for Dynamic Scene Graph Generation
by: Wang, Xuejiao, et al.
Published: (2026)
by: Wang, Xuejiao, et al.
Published: (2026)
HAMF: A Hybrid Attention-Mamba Framework for Joint Scene Context Understanding and Future Motion Representation Learning
by: Mei, Xiaodong, et al.
Published: (2025)
by: Mei, Xiaodong, et al.
Published: (2025)
TEACH: Text Encoding as Curriculum Hints for Scene Text Recognition
by: Yang, Xiahan, et al.
Published: (2025)
by: Yang, Xiahan, et al.
Published: (2025)
Masked and Permuted Implicit Context Learning for Scene Text Recognition
by: Yang, Xiaomeng, et al.
Published: (2023)
by: Yang, Xiaomeng, et al.
Published: (2023)
Adaptive Dual Residual U-Net with Attention Gate and Multiscale Spatial Attention Mechanisms (ADRUwAMS)
by: Suraki, Mohsen Yaghoubi
Published: (2026)
by: Suraki, Mohsen Yaghoubi
Published: (2026)
Flow-Assisted Motion Learning Network for Weakly-Supervised Group Activity Recognition
by: Nugroho, Muhammad Adi, et al.
Published: (2024)
by: Nugroho, Muhammad Adi, et al.
Published: (2024)
OUS: Scene-Guided Dynamic Facial Expression Recognition
by: Mai, Xinji, et al.
Published: (2024)
by: Mai, Xinji, et al.
Published: (2024)
DIANet: A Phase-Aware Dual-Stream Network for Micro-Expression Recognition via Dynamic Images
by: Khuong, Vu Tram Anh, et al.
Published: (2025)
by: Khuong, Vu Tram Anh, et al.
Published: (2025)
SketchFusion: Learning Universal Sketch Features through Fusing Foundation Models
by: Koley, Subhadeep, et al.
Published: (2025)
by: Koley, Subhadeep, et al.
Published: (2025)
AGCD-Net: Attention Guided Context Debiasing Network for Emotion Recognition
by: Devi, Varsha, et al.
Published: (2025)
by: Devi, Varsha, et al.
Published: (2025)
DCMorph: Face Morphing via Dual-Stream Cross-Attention Diffusion
by: Chettaoui, Tahar, et al.
Published: (2026)
by: Chettaoui, Tahar, et al.
Published: (2026)
Adaptive Temporal Motion Guided Graph Convolution Network for Micro-expression Recognition
by: Zhang, Fengyuan, et al.
Published: (2024)
by: Zhang, Fengyuan, et al.
Published: (2024)
Instruction-Guided Scene Text Recognition
by: Du, Yongkun, et al.
Published: (2024)
by: Du, Yongkun, et al.
Published: (2024)
S2M-Net: Spectral-Spatial Mixing for Medical Image Segmentation with Morphology-Aware Adaptive Loss
by: Sabrin, Md. Sanaullah Chowdhury Lameya
Published: (2026)
by: Sabrin, Md. Sanaullah Chowdhury Lameya
Published: (2026)
Instant Gaussian Stream: Fast and Generalizable Streaming of Dynamic Scene Reconstruction via Gaussian Splatting
by: Yan, Jinbo, et al.
Published: (2025)
by: Yan, Jinbo, et al.
Published: (2025)
Design and Analysis of Efficient Attention in Transformers for Social Group Activity Recognition
by: Tamura, Masato
Published: (2024)
by: Tamura, Masato
Published: (2024)
Vanishing Depth: A Depth Adapter with Positional Depth Encoding for Generalized Image Encoders
by: Koch, Paul, et al.
Published: (2025)
by: Koch, Paul, et al.
Published: (2025)
MoCLIP-Lite: Efficient Video Recognition by Fusing CLIP with Motion Vectors
by: Huang, Binhua, et al.
Published: (2025)
by: Huang, Binhua, et al.
Published: (2025)
MangoLeafViT: Leveraging Lightweight Vision Transformer with Runtime Augmentation for Efficient Mango Leaf Disease Classification
by: Chowdhury, Rafi Hassan, et al.
Published: (2025)
by: Chowdhury, Rafi Hassan, et al.
Published: (2025)
Convolutional Sparse Support Estimator Based Covid-19 Recognition from X-ray Images
by: Yamac, Mehmet, et al.
Published: (2020)
by: Yamac, Mehmet, et al.
Published: (2020)
DSXFormer: Dual-Pooling Spectral Squeeze-Expansion and Dynamic Context Attention Transformer for Hyperspectral Image Classification
by: Ullah, Farhan, et al.
Published: (2026)
by: Ullah, Farhan, et al.
Published: (2026)
Gated Fields: Learning Scene Reconstruction from Gated Videos
by: Ramazzina, Andrea, et al.
Published: (2024)
by: Ramazzina, Andrea, et al.
Published: (2024)
Segmentation of Ischemic Stroke Lesions using Transfer Learning on Multi-sequence MRI
by: Chowdhury, R. P., et al.
Published: (2025)
by: Chowdhury, R. P., et al.
Published: (2025)
Fusing Memory and Attention: A study on LSTM, Transformer and Hybrid Architectures for Symbolic Music Generation
by: Ghoshal, Soudeep, et al.
Published: (2026)
by: Ghoshal, Soudeep, et al.
Published: (2026)
Similar Items
-
AnatomicalNets: A Multi-Structure Segmentation and Contour-Based Distance Estimation Pipeline for Clinically Grounded Lung Cancer T-Staging
by: Chowdhury, Saniah Kayenat, et al.
Published: (2025) -
Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition
by: Guo, Hanyu, et al.
Published: (2024) -
Look Beyond Saliency: Low-Attention Guided Dual Encoding for Video Semantic Search
by: Aljehrai, Faisal, et al.
Published: (2026) -
Deep Neural Network-Based Sign Language Recognition: A Comprehensive Approach Using Transfer Learning with Explainability
by: Ridwan, A. E. M, et al.
Published: (2024) -
Automated Detection of Mutual Gaze and Joint Attention in Dual-Camera Settings via Dual-Stream Transformers
by: Kosmydel, Jakub, et al.
Published: (2026)