End-to-end Video Gaze Estimation via Capturing Head-face-eye Spatial-temporal Interaction Context
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Guan, Yiran, Chen, Zhuoguang, Zeng, Wenzheng, Cao, Zhiguo, Xiao, Yang |
|---|---|
| Format: | Preprint |
| Publié: |
2023
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
GazeHTA: End-to-end Gaze Target Detection with Head-Target Association
par: Lin, Zhi-Yi, et autres
Publié: (2024)
par: Lin, Zhi-Yi, et autres
Publié: (2024)
DICE: End-to-end Deformation Capture of Hand-Face Interactions from a Single Image
par: Wu, Qingxuan, et autres
Publié: (2024)
par: Wu, Qingxuan, et autres
Publié: (2024)
StyGazeTalk: Learning Stylized Generation of Gaze and Head Dynamics
par: Shi, Chengwei, et autres
Publié: (2025)
par: Shi, Chengwei, et autres
Publié: (2025)
CLIP-Gaze: Towards General Gaze Estimation via Visual-Linguistic Model
par: Yin, Pengwei, et autres
Publié: (2024)
par: Yin, Pengwei, et autres
Publié: (2024)
GazeDETR: Gaze Detection using Disentangled Head and Gaze Representations
par: de Belen, Ryan Anthony Jalova, et autres
Publié: (2025)
par: de Belen, Ryan Anthony Jalova, et autres
Publié: (2025)
Identity-Preserving Image-to-Video Generation via Reward-Guided Optimization
par: Shen, Liao, et autres
Publié: (2025)
par: Shen, Liao, et autres
Publié: (2025)
DHECA-SuperGaze: Dual Head-Eye Cross-Attention and Super-Resolution for Unconstrained Gaze Estimation
par: Šikić, Franko, et autres
Publié: (2025)
par: Šikić, Franko, et autres
Publié: (2025)
HOIGaze: Gaze Estimation During Hand-Object Interactions in Extended Reality Exploiting Eye-Hand-Head Coordination
par: Hu, Zhiming, et autres
Publié: (2025)
par: Hu, Zhiming, et autres
Publié: (2025)
CrossGLG: LLM Guides One-shot Skeleton-based 3D Action Recognition in a Cross-level Manner
par: Yan, Tingbing, et autres
Publié: (2024)
par: Yan, Tingbing, et autres
Publié: (2024)
Enhancing Space-time Video Super-resolution via Spatial-temporal Feature Interaction
par: Yue, Zijie, et autres
Publié: (2022)
par: Yue, Zijie, et autres
Publié: (2022)
Gaze Label Alignment: Alleviating Domain Shift for Gaze Estimation
par: Zeng, Guanzhong, et autres
Publié: (2024)
par: Zeng, Guanzhong, et autres
Publié: (2024)
PPAD: Iterative Interactions of Prediction and Planning for End-to-end Autonomous Driving
par: Chen, Zhili, et autres
Publié: (2023)
par: Chen, Zhili, et autres
Publié: (2023)
GazeFormer-MoE: Context-Aware Gaze Estimation via CLIP and MoE Transformer
par: Zhao, Xinyuan, et autres
Publié: (2026)
par: Zhao, Xinyuan, et autres
Publié: (2026)
UncAD: Towards Safe End-to-end Autonomous Driving via Online Map Uncertainty
par: Yang, Pengxuan, et autres
Publié: (2025)
par: Yang, Pengxuan, et autres
Publié: (2025)
GaTector+: A Unified Head-free Framework for Gaze Object and Gaze Following Prediction
par: Jin, Yang, et autres
Publié: (2025)
par: Jin, Yang, et autres
Publié: (2025)
REST: Diffusion-based Real-time End-to-end Streaming Talking Head Generation via ID-Context Caching and Asynchronous Streaming Distillation
par: Wang, Haotian, et autres
Publié: (2025)
par: Wang, Haotian, et autres
Publié: (2025)
NVDS+: Towards Efficient and Versatile Neural Stabilizer for Video Depth Estimation
par: Wang, Yiran, et autres
Publié: (2023)
par: Wang, Yiran, et autres
Publié: (2023)
DepthVLA: Enhancing Vision-Language-Action Models with Depth-Aware Spatial Reasoning
par: Yuan, Tianyuan, et autres
Publié: (2025)
par: Yuan, Tianyuan, et autres
Publié: (2025)
COMICS: End-to-end Bi-grained Contrastive Learning for Multi-face Forgery Detection
par: Zhang, Cong, et autres
Publié: (2023)
par: Zhang, Cong, et autres
Publié: (2023)
LG-Gaze: Learning Geometry-aware Continuous Prompts for Language-Guided Gaze Estimation
par: Yin, Pengwei, et autres
Publié: (2024)
par: Yin, Pengwei, et autres
Publié: (2024)
GA3CE: Unconstrained 3D Gaze Estimation with Gaze-Aware 3D Context Encoding
par: Kawana, Yuki, et autres
Publié: (2025)
par: Kawana, Yuki, et autres
Publié: (2025)
DREAM: Document Reconstruction via End-to-end Autoregressive Model
par: Li, Xin, et autres
Publié: (2025)
par: Li, Xin, et autres
Publié: (2025)
Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders
par: Ryan, Fiona, et autres
Publié: (2024)
par: Ryan, Fiona, et autres
Publié: (2024)
End-to-end Semantic-centric Video-based Multimodal Affective Computing
par: Lin, Ronghao, et autres
Publié: (2024)
par: Lin, Ronghao, et autres
Publié: (2024)
DFIMat: Decoupled Flexible Interactive Matting in Multi-Person Scenarios
par: Jiao, Siyi, et autres
Publié: (2024)
par: Jiao, Siyi, et autres
Publié: (2024)
Towards Robust Monocular Depth Estimation in Non-Lambertian Surfaces
par: Zhang, Junrui, et autres
Publié: (2024)
par: Zhang, Junrui, et autres
Publié: (2024)
In-Context Matting
par: Guo, He, et autres
Publié: (2024)
par: Guo, He, et autres
Publié: (2024)
Capturing Head Avatar with Hand Contacts from a Monocular Video
par: He, Haonan, et autres
Publié: (2025)
par: He, Haonan, et autres
Publié: (2025)
UniGaze: Towards Universal Gaze Estimation via Large-scale Pre-Training
par: Qin, Jiawei, et autres
Publié: (2025)
par: Qin, Jiawei, et autres
Publié: (2025)
Semi-Supervised Gaze Estimation via Disentangled Subspace Contrastive Learning
par: Tan, Qida, et autres
Publié: (2026)
par: Tan, Qida, et autres
Publié: (2026)
GazeD: Context-Aware Diffusion for Accurate 3D Gaze Estimation
par: Catalini, Riccardo, et autres
Publié: (2026)
par: Catalini, Riccardo, et autres
Publié: (2026)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
par: Yu, Jiwen, et autres
Publié: (2025)
par: Yu, Jiwen, et autres
Publié: (2025)
HALO: Human-Aligned End-to-end Image Retargeting with Layered Transformations
par: Xu, Yiran, et autres
Publié: (2025)
par: Xu, Yiran, et autres
Publié: (2025)
RoboTAG: End-to-end Robot Configuration Estimation via Topological Alignment Graph
par: Liu, Yifan, et autres
Publié: (2025)
par: Liu, Yifan, et autres
Publié: (2025)
Data-driven Head Motion Generation through Natural Gaze-Head Coordination
par: Liu, Xiaohan, et autres
Publié: (2026)
par: Liu, Xiaohan, et autres
Publié: (2026)
End-to-End Multi-Person Pose Estimation with Pose-Aware Video Transformer
par: Yu, Yonghui, et autres
Publié: (2025)
par: Yu, Yonghui, et autres
Publié: (2025)
GazeShift: Unsupervised Gaze Estimation and Dataset for VR
par: Shapira, Gil, et autres
Publié: (2026)
par: Shapira, Gil, et autres
Publié: (2026)
Large Spatial Model: End-to-end Unposed Images to Semantic 3D
par: Fan, Zhiwen, et autres
Publié: (2024)
par: Fan, Zhiwen, et autres
Publié: (2024)
TacoDepth: Towards Efficient Radar-Camera Depth Estimation with One-stage Fusion
par: Wang, Yiran, et autres
Publié: (2025)
par: Wang, Yiran, et autres
Publié: (2025)
Spatio-Temporal Attention and Gaussian Processes for Personalized Video Gaze Estimation
par: Jindal, Swati, et autres
Publié: (2024)
par: Jindal, Swati, et autres
Publié: (2024)
Documents similaires
-
GazeHTA: End-to-end Gaze Target Detection with Head-Target Association
par: Lin, Zhi-Yi, et autres
Publié: (2024) -
DICE: End-to-end Deformation Capture of Hand-Face Interactions from a Single Image
par: Wu, Qingxuan, et autres
Publié: (2024) -
StyGazeTalk: Learning Stylized Generation of Gaze and Head Dynamics
par: Shi, Chengwei, et autres
Publié: (2025) -
CLIP-Gaze: Towards General Gaze Estimation via Visual-Linguistic Model
par: Yin, Pengwei, et autres
Publié: (2024) -
GazeDETR: Gaze Detection using Disentangled Head and Gaze Representations
par: de Belen, Ryan Anthony Jalova, et autres
Publié: (2025)