GazeQwen: Lightweight Gaze-Conditioned LLM Modulation for Streaming Video Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Pham, Trong Thang, Nguyen, Hien, Le, Ngan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GazeSearch: Radiology Findings Search Benchmark
by: Pham, Trong Thang, et al.
Published: (2024)
by: Pham, Trong Thang, et al.
Published: (2024)
MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
by: Pham, Trong-Thang, et al.
Published: (2026)
by: Pham, Trong-Thang, et al.
Published: (2026)
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation
by: Pham, Trong Thang, et al.
Published: (2024)
by: Pham, Trong Thang, et al.
Published: (2024)
StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos
by: Lee, Daeun, et al.
Published: (2025)
by: Lee, Daeun, et al.
Published: (2025)
Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis
by: Pham, Trong-Thang, et al.
Published: (2025)
by: Pham, Trong-Thang, et al.
Published: (2025)
DuFal: Dual-Frequency-Aware Learning for High-Fidelity Extremely Sparse-view CBCT Reconstruction
by: Van, Cuong Tran, et al.
Published: (2026)
by: Van, Cuong Tran, et al.
Published: (2026)
Gaze-VLM:Bridging Gaze and VLMs through Attention Regularization for Egocentric Understanding
by: Pani, Anupam, et al.
Published: (2025)
by: Pani, Anupam, et al.
Published: (2025)
GazeVLM: A Vision-Language Model for Multi-Task Gaze Understanding
by: Mathew, Athul M., et al.
Published: (2025)
by: Mathew, Athul M., et al.
Published: (2025)
GazeMoE: Perception of Gaze Target with Mixture-of-Experts
by: Dai, Zhuangzhuang, et al.
Published: (2026)
by: Dai, Zhuangzhuang, et al.
Published: (2026)
CT-ScanGaze: A Dataset and Baselines for 3D Volumetric Scanpath Modeling
by: Pham, Trong-Thang, et al.
Published: (2025)
by: Pham, Trong-Thang, et al.
Published: (2025)
Eyes on Target: Gaze-Aware Object Detection in Egocentric Video
by: Lall, Vishakha, et al.
Published: (2025)
by: Lall, Vishakha, et al.
Published: (2025)
TPP-Gaze: Modelling Gaze Dynamics in Space and Time with Neural Temporal Point Processes
by: D'Amelio, Alessandro, et al.
Published: (2024)
by: D'Amelio, Alessandro, et al.
Published: (2024)
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
by: Zhao, Xinyuan, et al.
Published: (2026)
by: Zhao, Xinyuan, et al.
Published: (2026)
Learning Gaze-aware Compositional GAN
by: Aranjuelo, Nerea, et al.
Published: (2024)
by: Aranjuelo, Nerea, et al.
Published: (2024)
InverFill: One-Step Inversion for Enhanced Few-Step Diffusion Inpainting
by: Vu, Duc, et al.
Published: (2026)
by: Vu, Duc, et al.
Published: (2026)
DeepFake Detection in Dyadic Video Calls using Point of Gaze Tracking
by: Kohler, Odin, et al.
Published: (2025)
by: Kohler, Odin, et al.
Published: (2025)
WAVER: Writing-style Agnostic Text-Video Retrieval via Distilling Vision-Language Models Through Open-Vocabulary Knowledge
by: Le, Huy, et al.
Published: (2023)
by: Le, Huy, et al.
Published: (2023)
GazeLLM: Multimodal LLMs incorporating Human Visual Attention
by: Rekimoto, Jun
Published: (2025)
by: Rekimoto, Jun
Published: (2025)
Is Geometry Enough? An Evaluation of Landmark-Based Gaze Estimation
by: Agostinelli, Daniele, et al.
Published: (2026)
by: Agostinelli, Daniele, et al.
Published: (2026)
RGBD Gaze Tracking Using Transformer for Feature Fusion
by: Bauer, Tobias J.
Published: (2025)
by: Bauer, Tobias J.
Published: (2025)
Weakly-supervised Medical Image Segmentation with Gaze Annotations
by: Zhong, Yuan, et al.
Published: (2024)
by: Zhong, Yuan, et al.
Published: (2024)
3D Gaussian and Diffusion-Based Gaze Redirection
by: Panchalingam, Abiram, et al.
Published: (2025)
by: Panchalingam, Abiram, et al.
Published: (2025)
Toward Gaze Target Detection of Young Autistic Children
by: Deng, Shijian, et al.
Published: (2025)
by: Deng, Shijian, et al.
Published: (2025)
Learning Spatio-Temporal Feature Representations for Video-Based Gaze Estimation
by: Personnic, Alexandre, et al.
Published: (2025)
by: Personnic, Alexandre, et al.
Published: (2025)
GazeFusion: Saliency-Guided Image Generation
by: Zhang, Yunxiang, et al.
Published: (2024)
by: Zhang, Yunxiang, et al.
Published: (2024)
See Through the Noise: Improving Domain Generalization in Gaze Estimation
by: Peng, Yanming, et al.
Published: (2026)
by: Peng, Yanming, et al.
Published: (2026)
Unveiling the Truth: Exploring Human Gaze Patterns in Fake Images
by: Cartella, Giuseppe, et al.
Published: (2024)
by: Cartella, Giuseppe, et al.
Published: (2024)
Toddlers' Active Gaze Behavior Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2024)
by: Yu, Zhengyang, et al.
Published: (2024)
Modeling Human Gaze Behavior with Diffusion Models for Unified Scanpath Prediction
by: Cartella, Giuseppe, et al.
Published: (2025)
by: Cartella, Giuseppe, et al.
Published: (2025)
A Generalized Label Shift Perspective for Cross-Domain Gaze Estimation
by: Yang, Hao-Ran, et al.
Published: (2025)
by: Yang, Hao-Ran, et al.
Published: (2025)
DMAGaze: Gaze Estimation Based on Feature Disentanglement and Multi-Scale Attention
by: Chen, Haohan, et al.
Published: (2025)
by: Chen, Haohan, et al.
Published: (2025)
BiMa: Towards Biases Mitigation for Text-Video Retrieval via Scene Element Guidance
by: Le, Huy, et al.
Published: (2025)
by: Le, Huy, et al.
Published: (2025)
A2VIS: Amodal-Aware Approach to Video Instance Segmentation
by: Tran, Minh, et al.
Published: (2024)
by: Tran, Minh, et al.
Published: (2024)
Lightweight Models for Emotional Analysis in Video
by: Nguyen, Quoc-Tien, et al.
Published: (2025)
by: Nguyen, Quoc-Tien, et al.
Published: (2025)
Modeling Subjective Urban Perception with Human Gaze
by: Che, Lin, et al.
Published: (2026)
by: Che, Lin, et al.
Published: (2026)
Thinking with Gaze: Sequential Eye-Tracking as Visual Reasoning Supervision for Medical VLMs
by: Li, Yiwei, et al.
Published: (2026)
by: Li, Yiwei, et al.
Published: (2026)
Cross-Paradigm Evaluation of Gaze-Based Semantic Object Identification for Intelligent Vehicles
by: Deng, Penghao, et al.
Published: (2026)
by: Deng, Penghao, et al.
Published: (2026)
LogicGaze: Benchmarking Causal Consistency in Visual Narratives via Counterfactual Verification
by: Driscoll, Rory, et al.
Published: (2026)
by: Driscoll, Rory, et al.
Published: (2026)
Glance-or-Gaze: Incentivizing LMMs to Adaptively Focus Search via Reinforcement Learning
by: Bai, Hongbo, et al.
Published: (2026)
by: Bai, Hongbo, et al.
Published: (2026)
3D Gaze Tracking for Studying Collaborative Interactions in Mixed-Reality Environments
by: Davalos, Eduardo, et al.
Published: (2024)
by: Davalos, Eduardo, et al.
Published: (2024)
Similar Items
-
GazeSearch: Radiology Findings Search Benchmark
by: Pham, Trong Thang, et al.
Published: (2024) -
MedSteer: Counterfactual Endoscopic Synthesis via Training-Free Activation Steering
by: Pham, Trong-Thang, et al.
Published: (2026) -
FG-CXR: A Radiologist-Aligned Gaze Dataset for Enhancing Interpretability in Chest X-Ray Report Generation
by: Pham, Trong Thang, et al.
Published: (2024) -
StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos
by: Lee, Daeun, et al.
Published: (2025) -
Interpreting Radiologist's Intention from Eye Movements in Chest X-ray Diagnosis
by: Pham, Trong-Thang, et al.
Published: (2025)