Multi-view Gaze Target Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miao, Qiaomu, Golani, Vivek Raju, Xu, Jingyi, Dutta, Progga Paromita, Hoai, Minh, Samaras, Dimitris |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Diffusion-Refined VQA Annotations for Semi-Supervised Gaze Following
von: Miao, Qiaomu, et al.
Veröffentlicht: (2024)
von: Miao, Qiaomu, et al.
Veröffentlicht: (2024)
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
von: Wu, Haoyu, et al.
Veröffentlicht: (2025)
Look Hear: Gaze Prediction for Speech-directed Human Attention
von: Mondal, Sounak, et al.
Veröffentlicht: (2024)
von: Mondal, Sounak, et al.
Veröffentlicht: (2024)
Few-shot Personalized Scanpath Prediction
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
Assessing Sample Quality via the Latent Space of Generative Models
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
Personalized Image Descriptions from Attention Sequences
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025)
Importance-Based Token Merging for Efficient Image and Video Generation
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
von: Wu, Haoyu, et al.
Veröffentlicht: (2024)
Talking Head Generation via AU-Guided Landmark Prediction
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
von: Chang, Shao-Yu, et al.
Veröffentlicht: (2025)
Embedding Physical Reasoning into Diffusion-Based Shadow Generation
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Cast and Attached Shadow Detection via Iterative Light and Geometry Reasoning
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
von: Hu, Shilin, et al.
Veröffentlicht: (2025)
Self-supervised co-salient object detection via feature correspondence at multiple scales
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
von: Chakraborty, Souradeep, et al.
Veröffentlicht: (2024)
Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers
von: Yang, Zhibo, et al.
Veröffentlicht: (2023)
von: Yang, Zhibo, et al.
Veröffentlicht: (2023)
What about gravity in video generation? Post-Training Newton's Laws with Verifiable Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2025)
MIGS: Multi-Identity Gaussian Splatting via Tensor Decomposition
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
Beyond Pixels: Semi-Supervised Semantic Segmentation with a Multi-scale Patch-based Multi-Label Classifier
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Improving Contrastive Learning for Referring Expression Counting
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
von: Triaridis, Kostas, et al.
Veröffentlicht: (2025)
CORA: Consistency-Guided Semi-Supervised Framework for Reasoning Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2025)
von: Howlader, Prantik, et al.
Veröffentlicht: (2025)
TopoDiffusionNet: A Topology-aware Diffusion Model
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
von: Gupta, Saumya, et al.
Veröffentlicht: (2024)
Weighting Pseudo-Labels via High-Activation Feature Index Similarity and Object Detection for Semi-Supervised Segmentation
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
von: Howlader, Prantik, et al.
Veröffentlicht: (2024)
Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders
von: Ryan, Fiona, et al.
Veröffentlicht: (2024)
von: Ryan, Fiona, et al.
Veröffentlicht: (2024)
DualMat: PBR Material Estimation via Coherent Dual-Path Diffusion
von: Huang, Yifeng, et al.
Veröffentlicht: (2025)
von: Huang, Yifeng, et al.
Veröffentlicht: (2025)
MI-NeRF: Learning a Single Face NeRF from Multiple Identities
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
von: Chatziagapi, Aggelina, et al.
Veröffentlicht: (2024)
GazeTarget360: Towards Gaze Target Estimation in 360-Degree for Robot Perception
von: Dai, Zhuangzhuang, et al.
Veröffentlicht: (2025)
von: Dai, Zhuangzhuang, et al.
Veröffentlicht: (2025)
PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards
von: Le, Minh-Quan, et al.
Veröffentlicht: (2026)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2026)
$\infty$-Brush: Controllable Large Image Synthesis with Diffusion Models in Infinite Dimensions
von: Le, Minh-Quan, et al.
Veröffentlicht: (2024)
von: Le, Minh-Quan, et al.
Veröffentlicht: (2024)
Poppy: Polarization-based Plug-and-Play Guidance for Enhancing Monocular Normal Estimation
von: Kim, Irene, et al.
Veröffentlicht: (2026)
von: Kim, Irene, et al.
Veröffentlicht: (2026)
Upper-Body Pose-based Gaze Estimation for Privacy-Preserving 3D Gaze Target Detection
von: Toaiari, Andrea, et al.
Veröffentlicht: (2024)
von: Toaiari, Andrea, et al.
Veröffentlicht: (2024)
Learning 3D Reconstruction with Priors in Test Time
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
von: Zhou, Lei, et al.
Veröffentlicht: (2026)
JEAN: Joint Expression and Audio-guided NeRF-based Talking Face Generation
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
von: Chakkera, Sai Tanmay Reddy, et al.
Veröffentlicht: (2024)
Fast constrained sampling in pre-trained diffusion models
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2024)
GazeHTA: End-to-end Gaze Target Detection with Head-Target Association
von: Lin, Zhi-Yi, et al.
Veröffentlicht: (2024)
von: Lin, Zhi-Yi, et al.
Veröffentlicht: (2024)
Gaze Label Alignment: Alleviating Domain Shift for Gaze Estimation
von: Zeng, Guanzhong, et al.
Veröffentlicht: (2024)
von: Zeng, Guanzhong, et al.
Veröffentlicht: (2024)
MLI-NeRF: Multi-Light Intrinsic-Aware Neural Radiance Fields
von: Yang, Yixiong, et al.
Veröffentlicht: (2024)
von: Yang, Yixiong, et al.
Veröffentlicht: (2024)
In the Eye of Transformer: Global-Local Correlation for Egocentric Gaze Estimation
von: Lai, Bolin, et al.
Veröffentlicht: (2022)
von: Lai, Bolin, et al.
Veröffentlicht: (2022)
Leveraging Multi-Modal Saliency and Fusion for Gaze Target Detection
von: Mathew, Athul M., et al.
Veröffentlicht: (2025)
von: Mathew, Athul M., et al.
Veröffentlicht: (2025)
Learned representation-guided diffusion models for large-image generation
von: Graikos, Alexandros, et al.
Veröffentlicht: (2023)
von: Graikos, Alexandros, et al.
Veröffentlicht: (2023)
PathSegDiff: Pathology Segmentation using Diffusion model representations
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
von: Danisetty, Sachin Kumar, et al.
Veröffentlicht: (2025)
Rig3DGS: Creating Controllable Portraits from Casual Monocular Videos
von: Rivero, Alfredo, et al.
Veröffentlicht: (2024)
von: Rivero, Alfredo, et al.
Veröffentlicht: (2024)
Multi-task Gaze Estimation Via Unidirectional Convolution
von: Cheng, Zhang, et al.
Veröffentlicht: (2024)
von: Cheng, Zhang, et al.
Veröffentlicht: (2024)
GazeShift: Unsupervised Gaze Estimation and Dataset for VR
von: Shapira, Gil, et al.
Veröffentlicht: (2026)
von: Shapira, Gil, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Diffusion-Refined VQA Annotations for Semi-Supervised Gaze Following
von: Miao, Qiaomu, et al.
Veröffentlicht: (2024) -
One Attention, One Scale: Phase-Aligned Rotary Positional Embeddings for Mixed-Resolution Diffusion Transformer
von: Wu, Haoyu, et al.
Veröffentlicht: (2025) -
Look Hear: Gaze Prediction for Speech-directed Human Attention
von: Mondal, Sounak, et al.
Veröffentlicht: (2024) -
Few-shot Personalized Scanpath Prediction
von: Xue, Ruoyu, et al.
Veröffentlicht: (2025) -
Assessing Sample Quality via the Latent Space of Generative Models
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)