DMAGaze: Gaze Estimation Based on Feature Disentanglement and Multi-Scale Attention
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Haohan, Liu, Hongjia, Lan, Shiyong, Wang, Wenwu, Qiao, Yixin, Li, Yao, Deng, Guonan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Diffusion Model with Cross Attention as an Inductive Bias for Disentanglement
by: Yang, Tao, et al.
Published: (2024)
by: Yang, Tao, et al.
Published: (2024)
SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze
by: Sharma, Pavan Kumar, et al.
Published: (2026)
by: Sharma, Pavan Kumar, et al.
Published: (2026)
Is Geometry Enough? An Evaluation of Landmark-Based Gaze Estimation
by: Agostinelli, Daniele, et al.
Published: (2026)
by: Agostinelli, Daniele, et al.
Published: (2026)
Learning Spatio-Temporal Feature Representations for Video-Based Gaze Estimation
by: Personnic, Alexandre, et al.
Published: (2025)
by: Personnic, Alexandre, et al.
Published: (2025)
Gaze-VLM:Bridging Gaze and VLMs through Attention Regularization for Egocentric Understanding
by: Pani, Anupam, et al.
Published: (2025)
by: Pani, Anupam, et al.
Published: (2025)
SecureGaze: Defending Gaze Estimation Against Backdoor Attacks
by: Du, Lingyu, et al.
Published: (2025)
by: Du, Lingyu, et al.
Published: (2025)
MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
by: Liu, Hongjia, et al.
Published: (2025)
by: Liu, Hongjia, et al.
Published: (2025)
GazeOnce360: Fisheye-Based 360° Multi-Person Gaze Estimation with Global-Local Feature Fusion
by: Cai, Zhuojiang, et al.
Published: (2026)
by: Cai, Zhuojiang, et al.
Published: (2026)
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
by: Zhao, Xinyuan, et al.
Published: (2026)
by: Zhao, Xinyuan, et al.
Published: (2026)
Local-Global Attention: An Adaptive Mechanism for Multi-Scale Feature Integration
by: Shao, Yifan
Published: (2024)
by: Shao, Yifan
Published: (2024)
Domain-Adaptive Full-Face Gaze Estimation via Novel-View-Synthesis and Feature Disentanglement
by: Qin, Jiawei, et al.
Published: (2023)
by: Qin, Jiawei, et al.
Published: (2023)
Fast Disentangled Slim Tensor Learning for Multi-view Clustering
by: Xu, Deng, et al.
Published: (2024)
by: Xu, Deng, et al.
Published: (2024)
TripleFDS: Triple Feature Disentanglement and Synthesis for Scene Text Editing
by: Bao, Yuchen, et al.
Published: (2025)
by: Bao, Yuchen, et al.
Published: (2025)
GazeDETR: Gaze Detection using Disentangled Head and Gaze Representations
by: de Belen, Ryan Anthony Jalova, et al.
Published: (2025)
by: de Belen, Ryan Anthony Jalova, et al.
Published: (2025)
Cross-Paradigm Evaluation of Gaze-Based Semantic Object Identification for Intelligent Vehicles
by: Deng, Penghao, et al.
Published: (2026)
by: Deng, Penghao, et al.
Published: (2026)
RGBD Gaze Tracking Using Transformer for Feature Fusion
by: Bauer, Tobias J.
Published: (2025)
by: Bauer, Tobias J.
Published: (2025)
GazeVLM: A Vision-Language Model for Multi-Task Gaze Understanding
by: Mathew, Athul M., et al.
Published: (2025)
by: Mathew, Athul M., et al.
Published: (2025)
A Generalized Label Shift Perspective for Cross-Domain Gaze Estimation
by: Yang, Hao-Ran, et al.
Published: (2025)
by: Yang, Hao-Ran, et al.
Published: (2025)
Semi-Supervised Gaze Estimation via Disentangled Subspace Contrastive Learning
by: Tan, Qida, et al.
Published: (2026)
by: Tan, Qida, et al.
Published: (2026)
Distilling Cross-Modal Knowledge via Feature Disentanglement
by: Liu, Junhong, et al.
Published: (2025)
by: Liu, Junhong, et al.
Published: (2025)
Diffusion-Based Cross-Modal Feature Extraction for Multi-Label Classification
by: Lan, Tian, et al.
Published: (2025)
by: Lan, Tian, et al.
Published: (2025)
GazeMoE: Perception of Gaze Target with Mixture-of-Experts
by: Dai, Zhuangzhuang, et al.
Published: (2026)
by: Dai, Zhuangzhuang, et al.
Published: (2026)
Faithful Attention Explainer: Verbalizing Decisions Based on Discriminative Features
by: Rong, Yao, et al.
Published: (2024)
by: Rong, Yao, et al.
Published: (2024)
See Through the Noise: Improving Domain Generalization in Gaze Estimation
by: Peng, Yanming, et al.
Published: (2026)
by: Peng, Yanming, et al.
Published: (2026)
PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement
by: Zhan, Yu-Wei, et al.
Published: (2025)
by: Zhan, Yu-Wei, et al.
Published: (2025)
Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders
by: Ryan, Fiona, et al.
Published: (2024)
by: Ryan, Fiona, et al.
Published: (2024)
Tighnari: Multi-modal Plant Species Prediction Based on Hierarchical Cross-Attention Using Graph-Based and Vision Backbone-Extracted Features
by: Liu, Haixu, et al.
Published: (2025)
by: Liu, Haixu, et al.
Published: (2025)
OmniGaze: Reward-inspired Generalizable Gaze Estimation In The Wild
by: Qu, Hongyu, et al.
Published: (2025)
by: Qu, Hongyu, et al.
Published: (2025)
Cross-Layer Feature Self-Attention Module for Multi-Scale Object Detection
by: Xie, Dingzhou, et al.
Published: (2025)
by: Xie, Dingzhou, et al.
Published: (2025)
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
Cross-Camera Cow Identification via Disentangled Representation Learning
by: Wang, Runcheng, et al.
Published: (2026)
by: Wang, Runcheng, et al.
Published: (2026)
Gaze on the Prize: Shaping Visual Attention with Return-Guided Contrastive Learning
by: Lee, Andrew, et al.
Published: (2025)
by: Lee, Andrew, et al.
Published: (2025)
Nested-TNT: Hierarchical Vision Transformers with Multi-Scale Feature Processing
by: Liu, Yuang, et al.
Published: (2024)
by: Liu, Yuang, et al.
Published: (2024)
Multi-view Gaze Target Estimation
by: Miao, Qiaomu, et al.
Published: (2025)
by: Miao, Qiaomu, et al.
Published: (2025)
AFD: Mitigating Feature Gap for Adversarial Robustness by Feature Disentanglement
by: Zhou, Nuoyan, et al.
Published: (2024)
by: Zhou, Nuoyan, et al.
Published: (2024)
Image Forgery Localization via Guided Noise and Multi-Scale Feature Aggregation
by: Niu, Yakun, et al.
Published: (2024)
by: Niu, Yakun, et al.
Published: (2024)
GazeSearch: Radiology Findings Search Benchmark
by: Pham, Trong Thang, et al.
Published: (2024)
by: Pham, Trong Thang, et al.
Published: (2024)
DHECA-SuperGaze: Dual Head-Eye Cross-Attention and Super-Resolution for Unconstrained Gaze Estimation
by: Šikić, Franko, et al.
Published: (2025)
by: Šikić, Franko, et al.
Published: (2025)
PrivatEyes: Appearance-based Gaze Estimation Using Federated Secure Multi-Party Computation
by: Elfares, Mayar, et al.
Published: (2024)
by: Elfares, Mayar, et al.
Published: (2024)
3D Gaussian and Diffusion-Based Gaze Redirection
by: Panchalingam, Abiram, et al.
Published: (2025)
by: Panchalingam, Abiram, et al.
Published: (2025)
Similar Items
-
Diffusion Model with Cross Attention as an Inductive Bias for Disentanglement
by: Yang, Tao, et al.
Published: (2024) -
SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze
by: Sharma, Pavan Kumar, et al.
Published: (2026) -
Is Geometry Enough? An Evaluation of Landmark-Based Gaze Estimation
by: Agostinelli, Daniele, et al.
Published: (2026) -
Learning Spatio-Temporal Feature Representations for Video-Based Gaze Estimation
by: Personnic, Alexandre, et al.
Published: (2025) -
Gaze-VLM:Bridging Gaze and VLMs through Attention Regularization for Egocentric Understanding
by: Pani, Anupam, et al.
Published: (2025)