GazeD: Context-Aware Diffusion for Accurate 3D Gaze Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Catalini, Riccardo, Di Nucci, Davide, Borghi, Guido, Davoli, Davide, Garattoni, Lorenzo, Francesca, Gianpiero, Kawana, Yuki, Vezzani, Roberto |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SnapPose3D: Diffusion-Based Single-Frame 2D-to-3D Lifting of Human Poses
por: Simoni, Alessandro, et al.
Publicado: (2026)
por: Simoni, Alessandro, et al.
Publicado: (2026)
Fake3DGS: A Benchmark for 3D Manipulation Detection in Neural Rendering
por: Di Nucci, Davide, et al.
Publicado: (2026)
por: Di Nucci, Davide, et al.
Publicado: (2026)
Depth-based Privileged Information for Boosting 3D Human Pose Estimation on RGB
por: Simoni, Alessandro, et al.
Publicado: (2024)
por: Simoni, Alessandro, et al.
Publicado: (2024)
GA3CE: Unconstrained 3D Gaze Estimation with Gaze-Aware 3D Context Encoding
por: Kawana, Yuki, et al.
Publicado: (2025)
por: Kawana, Yuki, et al.
Publicado: (2025)
From Gaze to Insight: Bridging Human Visual Attention and Vision Language Model Explanation for Weakly-Supervised Medical Image Segmentation
por: Chen, Jingkun, et al.
Publicado: (2025)
por: Chen, Jingkun, et al.
Publicado: (2025)
Yolo-Key-6D: Single Stage Monocular 6D Pose Estimation with Keypoint Enhancements
por: Çetiner, Kemal Alperen, et al.
Publicado: (2026)
por: Çetiner, Kemal Alperen, et al.
Publicado: (2026)
On the Model Theory of Second-Order Objects
por: Hyttinen, Tapani, et al.
Publicado: (2024)
por: Hyttinen, Tapani, et al.
Publicado: (2024)
Detecting 3D Line Segments for 6DoF Pose Estimation with Limited Data
por: Mok, Matej, et al.
Publicado: (2026)
por: Mok, Matej, et al.
Publicado: (2026)
ReFlow6D: Refraction-Guided Transparent Object 6D Pose Estimation via Intermediate Representation Learning
por: Gupta, Hrishikesh, et al.
Publicado: (2024)
por: Gupta, Hrishikesh, et al.
Publicado: (2024)
Diffusion Features for Zero-Shot 6DoF Object Pose Estimation
por: Von Gimborn, Bernd, et al.
Publicado: (2024)
por: Von Gimborn, Bernd, et al.
Publicado: (2024)
BRUM: Robust 3D Vehicle Reconstruction from 360 Sparse Images
por: Di Nucci, Davide, et al.
Publicado: (2025)
por: Di Nucci, Davide, et al.
Publicado: (2025)
LISA: Language-guided Interference-aware Spatial-Frequency Attention for Driver Gaze Estimation
por: Ma, Jun, et al.
Publicado: (2026)
por: Ma, Jun, et al.
Publicado: (2026)
Gaussian Alignment for Relative Camera Pose Estimation via Single-View Reconstruction
por: Li, Yumin, et al.
Publicado: (2025)
por: Li, Yumin, et al.
Publicado: (2025)
Splat and Distill: Augmenting Teachers with Feed-Forward 3D Reconstruction For 3D-Aware Distillation
por: Shavin, David, et al.
Publicado: (2026)
por: Shavin, David, et al.
Publicado: (2026)
A straightening-unstraightening equivalence for $\infty$-operads
por: Pratali, Francesca
Publicado: (2025)
por: Pratali, Francesca
Publicado: (2025)
BID: Boundary-Interior Decoding for Unsupervised Temporal Action Localization Pre-Trainin
por: Fang, Qihang, et al.
Publicado: (2024)
por: Fang, Qihang, et al.
Publicado: (2024)
On the spectrum of limit models
por: Beard, Jeremy, et al.
Publicado: (2025)
por: Beard, Jeremy, et al.
Publicado: (2025)
Learning through Creation: A Hash-Free Framework for On-the-Fly Category Discovery
por: Zhang, Bohan, et al.
Publicado: (2026)
por: Zhang, Bohan, et al.
Publicado: (2026)
Rectification of dendroidal left fibrations
por: Pratali, Francesca
Publicado: (2025)
por: Pratali, Francesca
Publicado: (2025)
Gaze into the Heart: A Multi-View Video Dataset for rPPG and Health Biomarkers Estimation
por: Egorov, Konstantin, et al.
Publicado: (2025)
por: Egorov, Konstantin, et al.
Publicado: (2025)
Modal Semantics for Reasoning with Probability and Uncertainty
por: Guallart, Nino
Publicado: (2024)
por: Guallart, Nino
Publicado: (2024)
POC-SLT: Partial Object Completion with SDF Latent Transformers
por: Zakeri, Faezeh, et al.
Publicado: (2024)
por: Zakeri, Faezeh, et al.
Publicado: (2024)
3DreamBooth: High-Fidelity 3D Subject-Driven Video Generation Model
por: Ko, Hyun-kyu, et al.
Publicado: (2026)
por: Ko, Hyun-kyu, et al.
Publicado: (2026)
A new space of generalised vector-valued functions of bounded variation
por: Donati, Davide
Publicado: (2023)
por: Donati, Davide
Publicado: (2023)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
por: Zinnen, Mathias, et al.
Publicado: (2025)
por: Zinnen, Mathias, et al.
Publicado: (2025)
DSER: Spectral Epipolar Representation for Efficient Light Field Depth Estimation
por: Mohammad, Noor Islam S., et al.
Publicado: (2025)
por: Mohammad, Noor Islam S., et al.
Publicado: (2025)
Order Relations of the Wasserstein mean and the spectral geometric mean
por: Gan, Luyining, et al.
Publicado: (2023)
por: Gan, Luyining, et al.
Publicado: (2023)
Digital analysis of early color photographs taken using regular color screen processes
por: Hubička, Jan, et al.
Publicado: (2023)
por: Hubička, Jan, et al.
Publicado: (2023)
Improving Object Detection for Time-Lapse Imagery Using Temporal Features in Wildlife Monitoring
por: Jenkins, Marcus, et al.
Publicado: (2024)
por: Jenkins, Marcus, et al.
Publicado: (2024)
A Riemannian gradient descent method for optimization on the indefinite Stiefel manifold
por: Van Tiep, Dinh, et al.
Publicado: (2024)
por: Van Tiep, Dinh, et al.
Publicado: (2024)
Language-Based Swarm Perception: Decentralized Person Re-Identification via Natural Language Descriptions
por: Kegeleirs, Miquel, et al.
Publicado: (2026)
por: Kegeleirs, Miquel, et al.
Publicado: (2026)
AUTHENTICATION: Identifying Rare Failure Modes in Autonomous Vehicle Perception Systems using Adversarially Guided Diffusion Models
por: Zarei, Mohammad, et al.
Publicado: (2025)
por: Zarei, Mohammad, et al.
Publicado: (2025)
Long limit models are isomorphic assuming a splitting-like relation
por: Beard, Jeremy
Publicado: (2025)
por: Beard, Jeremy
Publicado: (2025)
VDPP: Video Depth Post-Processing for Speed and Scalability
por: Yoon, Daewon, et al.
Publicado: (2026)
por: Yoon, Daewon, et al.
Publicado: (2026)
An NIP-like Notion in Abstract Elementary Classes
por: Yang, Wentao
Publicado: (2023)
por: Yang, Wentao
Publicado: (2023)
Disjoint non-forking amalgamation in stable AECs
por: Beard, Jeremy
Publicado: (2026)
por: Beard, Jeremy
Publicado: (2026)
Camera Pose Revisited
por: Skarbek, Władysław, et al.
Publicado: (2026)
por: Skarbek, Władysław, et al.
Publicado: (2026)
DNRSelect: Active Best View Selection for Deferred Neural Rendering
por: Wu, Dongli, et al.
Publicado: (2025)
por: Wu, Dongli, et al.
Publicado: (2025)
RailSafeNet: Visual Scene Understanding for Tram Safety
por: Valach, Ondřej, et al.
Publicado: (2025)
por: Valach, Ondřej, et al.
Publicado: (2025)
HEDGE: Hallucination Estimation via Dense Geometric Entropy for VQA with Vision-Language Models
por: Gautam, Sushant, et al.
Publicado: (2025)
por: Gautam, Sushant, et al.
Publicado: (2025)
Ejemplares similares
-
SnapPose3D: Diffusion-Based Single-Frame 2D-to-3D Lifting of Human Poses
por: Simoni, Alessandro, et al.
Publicado: (2026) -
Fake3DGS: A Benchmark for 3D Manipulation Detection in Neural Rendering
por: Di Nucci, Davide, et al.
Publicado: (2026) -
Depth-based Privileged Information for Boosting 3D Human Pose Estimation on RGB
por: Simoni, Alessandro, et al.
Publicado: (2024) -
GA3CE: Unconstrained 3D Gaze Estimation with Gaze-Aware 3D Context Encoding
por: Kawana, Yuki, et al.
Publicado: (2025) -
From Gaze to Insight: Bridging Human Visual Attention and Vision Language Model Explanation for Weakly-Supervised Medical Image Segmentation
por: Chen, Jingkun, et al.
Publicado: (2025)