TransGOP: Transformer-Based Gaze Object Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Binglu, Guo, Chenxi, Jin, Yang, Xia, Haisheng, Liu, Nian |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GaTector+: A Unified Head-free Framework for Gaze Object and Gaze Following Prediction
by: Jin, Yang, et al.
Published: (2025)
by: Jin, Yang, et al.
Published: (2025)
Boosting Gaze Object Prediction via Pixel-level Supervision from Vision Foundation Model
by: Jin, Yang, et al.
Published: (2024)
by: Jin, Yang, et al.
Published: (2024)
OAT: Object-Level Attention Transformer for Gaze Scanpath Prediction
by: Fang, Yini, et al.
Published: (2024)
by: Fang, Yini, et al.
Published: (2024)
Vision-based Wearable Steering Assistance for People with Impaired Vision in Jogging
by: Liu, Xiaotong, et al.
Published: (2024)
by: Liu, Xiaotong, et al.
Published: (2024)
Multimodal Large Models Are Effective Action Anticipators
by: Wang, Binglu, et al.
Published: (2025)
by: Wang, Binglu, et al.
Published: (2025)
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
by: Lin, Xiao, et al.
Published: (2023)
by: Lin, Xiao, et al.
Published: (2023)
Towards Pixel-Level Prediction for Gaze Following: Benchmark and Approach
by: Liu, Feiyang, et al.
Published: (2024)
by: Liu, Feiyang, et al.
Published: (2024)
Learning from Observer Gaze:Zero-Shot Attention Prediction Oriented by Human-Object Interaction Recognition
by: Zhou, Yuchen, et al.
Published: (2024)
by: Zhou, Yuchen, et al.
Published: (2024)
From Scene to Object: Text-Guided Dual-Gaze Prediction
by: Ke, Zehong, et al.
Published: (2026)
by: Ke, Zehong, et al.
Published: (2026)
TextGaze: Gaze-Controllable Face Generation with Natural Language
by: Wang, Hengfei, et al.
Published: (2024)
by: Wang, Hengfei, et al.
Published: (2024)
Gaze-guided Hand-Object Interaction Synthesis: Dataset and Method
by: Tian, Jie, et al.
Published: (2024)
by: Tian, Jie, et al.
Published: (2024)
Eyes on VLM: Benchmarking Gaze Following and Social Gaze Prediction in Vision Language Models
by: Wang, Hengfei, et al.
Published: (2026)
by: Wang, Hengfei, et al.
Published: (2026)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
ViTGaze: Gaze Following with Interaction Features in Vision Transformers
by: Song, Yuehao, et al.
Published: (2024)
by: Song, Yuehao, et al.
Published: (2024)
VL4Gaze: Unleashing Vision-Language Models for Gaze Following
by: Wang, Shijing, et al.
Published: (2025)
by: Wang, Shijing, et al.
Published: (2025)
GazeProphet: Software-Only Gaze Prediction for VR Foveated Rendering
by: Ebadulla, Farhaan, et al.
Published: (2025)
by: Ebadulla, Farhaan, et al.
Published: (2025)
Enhancing Gaze Reasoning in Vision Foundation Models for Gaze Following
by: Wang, Shijing, et al.
Published: (2026)
by: Wang, Shijing, et al.
Published: (2026)
GazeProphetV2: Head-Movement-Based Gaze Prediction Enabling Efficient Foveated Rendering on Mobile VR
by: Ebadulla, Farhaan, et al.
Published: (2025)
by: Ebadulla, Farhaan, et al.
Published: (2025)
DualGazeNet: A Biologically Inspired Dual-Gaze Query Network for Salient Object Detection
by: Zhang, Yu, et al.
Published: (2025)
by: Zhang, Yu, et al.
Published: (2025)
SGAP-Gaze: Scene Grid Attention Based Point-of-Gaze Estimation Network for Driver Gaze
by: Sharma, Pavan Kumar, et al.
Published: (2026)
by: Sharma, Pavan Kumar, et al.
Published: (2026)
Cross-Paradigm Evaluation of Gaze-Based Semantic Object Identification for Intelligent Vehicles
by: Deng, Penghao, et al.
Published: (2026)
by: Deng, Penghao, et al.
Published: (2026)
FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
by: Su, Haisheng, et al.
Published: (2025)
by: Su, Haisheng, et al.
Published: (2025)
In the Eye of Transformer: Global-Local Correlation for Egocentric Gaze Estimation
by: Lai, Bolin, et al.
Published: (2022)
by: Lai, Bolin, et al.
Published: (2022)
Composed Object Retrieval: Object-level Retrieval via Composed Expressions
by: Wang, Tong, et al.
Published: (2025)
by: Wang, Tong, et al.
Published: (2025)
A Novel Framework for Multi-Person Temporal Gaze Following and Social Gaze Prediction
by: Gupta, Anshul, et al.
Published: (2024)
by: Gupta, Anshul, et al.
Published: (2024)
GazeMoDiff: Gaze-guided Diffusion Model for Stochastic Human Motion Prediction
by: Yan, Haodong, et al.
Published: (2023)
by: Yan, Haodong, et al.
Published: (2023)
Strong-TransCenter: Improved Multi-Object Tracking based on Transformers with Dense Representations
by: Galor, Amit, et al.
Published: (2022)
by: Galor, Amit, et al.
Published: (2022)
UniMamba: Unified Spatial-Channel Representation Learning with Group-Efficient Mamba for LiDAR-based 3D Object Detection
by: Jin, Xin, et al.
Published: (2025)
by: Jin, Xin, et al.
Published: (2025)
Object-Conditioned Energy-Based Attention Map Alignment in Text-to-Image Diffusion Models
by: Zhang, Yasi, et al.
Published: (2024)
by: Zhang, Yasi, et al.
Published: (2024)
SIGN: A Statistically-Informed Gaze Network for Gaze Time Prediction
by: Ye, Jianping, et al.
Published: (2025)
by: Ye, Jianping, et al.
Published: (2025)
Look Hear: Gaze Prediction for Speech-directed Human Attention
by: Mondal, Sounak, et al.
Published: (2024)
by: Mondal, Sounak, et al.
Published: (2024)
GMGaze: MoE-Based Context-Aware Gaze Estimation with CLIP and Multiscale Transformer
by: Zhao, Xinyuan, et al.
Published: (2026)
by: Zhao, Xinyuan, et al.
Published: (2026)
ARGaze: Autoregressive Transformers for Online Egocentric Gaze Estimation
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
PGNeXt: High-Resolution Salient Object Detection via Pyramid Grafting Network
by: Xia, Changqun, et al.
Published: (2024)
by: Xia, Changqun, et al.
Published: (2024)
TransBridge: Boost 3D Object Detection by Scene-Level Completion with Transformer Decoder
by: Meng, Qinghao, et al.
Published: (2025)
by: Meng, Qinghao, et al.
Published: (2025)
Differential Contrastive Training for Gaze Estimation
by: Zhang, Lin, et al.
Published: (2025)
by: Zhang, Lin, et al.
Published: (2025)
SecureGaze: Defending Gaze Estimation Against Backdoor Attacks
by: Du, Lingyu, et al.
Published: (2025)
by: Du, Lingyu, et al.
Published: (2025)
Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
by: Lall, Vishakha, et al.
Published: (2025)
by: Lall, Vishakha, et al.
Published: (2025)
Prime and Reach: Synthesising Body Motion for Gaze-Primed Object Reach
by: Hatano, Masashi, et al.
Published: (2025)
by: Hatano, Masashi, et al.
Published: (2025)
GazeDETR: Gaze Detection using Disentangled Head and Gaze Representations
by: de Belen, Ryan Anthony Jalova, et al.
Published: (2025)
by: de Belen, Ryan Anthony Jalova, et al.
Published: (2025)
Similar Items
-
GaTector+: A Unified Head-free Framework for Gaze Object and Gaze Following Prediction
by: Jin, Yang, et al.
Published: (2025) -
Boosting Gaze Object Prediction via Pixel-level Supervision from Vision Foundation Model
by: Jin, Yang, et al.
Published: (2024) -
OAT: Object-Level Attention Transformer for Gaze Scanpath Prediction
by: Fang, Yini, et al.
Published: (2024) -
Vision-based Wearable Steering Assistance for People with Impaired Vision in Jogging
by: Liu, Xiaotong, et al.
Published: (2024) -
Multimodal Large Models Are Effective Action Anticipators
by: Wang, Binglu, et al.
Published: (2025)