Gesture-Aware Pretraining and Token Fusion for 3D Hand Pose Estimation
Fuente:
arXiv
Guardado en:
| Autores principales: | Hong, Rui, Kosecka, Jana |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Toward Phonology-Guided Sign Language Motion Generation: A Diffusion Baseline and Conditioning Analysis
por: Hong, Rui, et al.
Publicado: (2026)
por: Hong, Rui, et al.
Publicado: (2026)
Conditional Collapse in Sign Language Production: A Diagnostic and a Scaling Argument
por: Hong, Rui, et al.
Publicado: (2026)
por: Hong, Rui, et al.
Publicado: (2026)
VarSplat: Uncertainty-aware 3D Gaussian Splatting for Robust RGB-D SLAM
por: Tran, Anh Thuan, et al.
Publicado: (2026)
por: Tran, Anh Thuan, et al.
Publicado: (2026)
PointSplat: Efficient Geometry-Driven Pruning and Transformer Refinement for 3D Gaussian Splatting
por: Tran, Anh Thuan, et al.
Publicado: (2026)
por: Tran, Anh Thuan, et al.
Publicado: (2026)
EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras
por: Wang, Luming, et al.
Publicado: (2026)
por: Wang, Luming, et al.
Publicado: (2026)
SPLite Hand: Sparsity-Aware Lightweight 3D Hand Pose Estimation
por: Hao, Yeh Keng, et al.
Publicado: (2025)
por: Hao, Yeh Keng, et al.
Publicado: (2025)
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
por: Yang, Hui, et al.
Publicado: (2025)
por: Yang, Hui, et al.
Publicado: (2025)
Direction-Aware Hybrid Representation Learning for 3D Hand Pose and Shape Estimation
por: Liu, Shiyong, et al.
Publicado: (2025)
por: Liu, Shiyong, et al.
Publicado: (2025)
Multi-temporal Adaptive Red-Green-Blue and Long-Wave Infrared Fusion for You Only Look Once-Based Landmine Detection from Unmanned Aerial Systems
por: Gallagher, James E., et al.
Publicado: (2025)
por: Gallagher, James E., et al.
Publicado: (2025)
mmEgoHand: Egocentric Hand Pose Estimation and Gesture Recognition with Head-mounted Millimeter-wave Radar and IMU
por: Lv, Yizhe, et al.
Publicado: (2025)
por: Lv, Yizhe, et al.
Publicado: (2025)
Gloss2Text: Sign Language Gloss translation using LLMs and Semantically Aware Label Smoothing
por: Fayyazsanavi, Pooya, et al.
Publicado: (2024)
por: Fayyazsanavi, Pooya, et al.
Publicado: (2024)
Multi-view Pose Fusion for Occlusion-Aware 3D Human Pose Estimation
por: Bragagnolo, Laura, et al.
Publicado: (2024)
por: Bragagnolo, Laura, et al.
Publicado: (2024)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
por: Rajabi, Navid, et al.
Publicado: (2024)
por: Rajabi, Navid, et al.
Publicado: (2024)
Towards Grounded Visual Spatial Reasoning in Multi-Modal Vision Language Models
por: Rajabi, Navid, et al.
Publicado: (2023)
por: Rajabi, Navid, et al.
Publicado: (2023)
Q-GroundCAM: Quantifying Grounding in Vision Language Models via GradCAM
por: Rajabi, Navid, et al.
Publicado: (2024)
por: Rajabi, Navid, et al.
Publicado: (2024)
Enhancing Hands in 3D Whole-Body Pose Estimation with Conditional Hands Modulator
por: Moon, Gyeongsik
Publicado: (2026)
por: Moon, Gyeongsik
Publicado: (2026)
HandDiff: 3D Hand Pose Estimation with Diffusion on Image-Point Cloud
por: Cheng, Wencan, et al.
Publicado: (2024)
por: Cheng, Wencan, et al.
Publicado: (2024)
6D Pose Estimation on Spoons and Hands
por: Tan, Kevin, et al.
Publicado: (2025)
por: Tan, Kevin, et al.
Publicado: (2025)
The Invisible EgoHand: 3D Hand Forecasting through EgoBody Pose Estimation
por: Hatano, Masashi, et al.
Publicado: (2025)
por: Hatano, Masashi, et al.
Publicado: (2025)
Towards Egocentric 3D Hand Pose Estimation in Unseen Domains
por: Mucha, Wiktor, et al.
Publicado: (2026)
por: Mucha, Wiktor, et al.
Publicado: (2026)
Structured Spatial Reasoning with Open Vocabulary Object Detectors
por: Nejatishahidin, Negar, et al.
Publicado: (2024)
por: Nejatishahidin, Negar, et al.
Publicado: (2024)
Compositional Image-Text Matching and Retrieval by Grounding Entities
por: Vongala, Madhukar Reddy, et al.
Publicado: (2025)
por: Vongala, Madhukar Reddy, et al.
Publicado: (2025)
GenHOI: Generalized Hand-Object Pose Estimation with Occlusion Awareness
por: Yang, Hui, et al.
Publicado: (2026)
por: Yang, Hui, et al.
Publicado: (2026)
Automated Patient Positioning with Learned 3D Hand Gestures
por: Gao, Zhongpai, et al.
Publicado: (2024)
por: Gao, Zhongpai, et al.
Publicado: (2024)
Single-to-Dual-View Adaptation for Egocentric 3D Hand Pose Estimation
por: Liu, Ruicong, et al.
Publicado: (2024)
por: Liu, Ruicong, et al.
Publicado: (2024)
Learning Correlation-aware Aleatoric Uncertainty for 3D Hand Pose Estimation
por: Chae-Yeon, Lee, et al.
Publicado: (2025)
por: Chae-Yeon, Lee, et al.
Publicado: (2025)
Analyzing the Synthetic-to-Real Domain Gap in 3D Hand Pose Estimation
por: Zhao, Zhuoran, et al.
Publicado: (2025)
por: Zhao, Zhuoran, et al.
Publicado: (2025)
Pre-Training for 3D Hand Pose Estimation with Contrastive Learning on Large-Scale Hand Images in the Wild
por: Lin, Nie, et al.
Publicado: (2024)
por: Lin, Nie, et al.
Publicado: (2024)
TRAVEL: Training-Free Retrieval and Alignment for Vision-and-Language Navigation
por: Rajabi, Navid, et al.
Publicado: (2025)
por: Rajabi, Navid, et al.
Publicado: (2025)
LA-Pose: Latent Action Pretraining Meets Pose Estimation
por: Wang, Zhengqing, et al.
Publicado: (2026)
por: Wang, Zhengqing, et al.
Publicado: (2026)
AnyHand: A Large-Scale Synthetic Dataset for RGB(-D) Hand Pose Estimation
por: Si, Chen, et al.
Publicado: (2026)
por: Si, Chen, et al.
Publicado: (2026)
HandDAGT: A Denoising Adaptive Graph Transformer for 3D Hand Pose Estimation
por: Cheng, Wencan, et al.
Publicado: (2024)
por: Cheng, Wencan, et al.
Publicado: (2024)
HandMCM: Multi-modal Point Cloud-based Correspondence State Space Model for 3D Hand Pose Estimation
por: Cheng, Wencan, et al.
Publicado: (2026)
por: Cheng, Wencan, et al.
Publicado: (2026)
A Multi-View Pipeline and Benchmark Dataset for 3D Hand Pose Estimation in Surgery
por: Fischer, Valery, et al.
Publicado: (2026)
por: Fischer, Valery, et al.
Publicado: (2026)
HOISDF: Constraining 3D Hand-Object Pose Estimation with Global Signed Distance Fields
por: Qi, Haozhe, et al.
Publicado: (2024)
por: Qi, Haozhe, et al.
Publicado: (2024)
PCIE_EgoHandPose Solution for EgoExo4D Hand Pose Challenge
por: Chen, Feng, et al.
Publicado: (2024)
por: Chen, Feng, et al.
Publicado: (2024)
Mamba-Driven Topology Fusion for Monocular 3D Human Pose Estimation
por: Zheng, Zenghao, et al.
Publicado: (2025)
por: Zheng, Zenghao, et al.
Publicado: (2025)
SHARP: Segmentation of Hands and Arms by Range using Pseudo-Depth for Enhanced Egocentric 3D Hand Pose Estimation and Action Recognition
por: Mucha, Wiktor, et al.
Publicado: (2024)
por: Mucha, Wiktor, et al.
Publicado: (2024)
REACH: Hand Pose Estimation from Room Corners
por: Nakamura, Shu, et al.
Publicado: (2026)
por: Nakamura, Shu, et al.
Publicado: (2026)
UniHOPE: A Unified Approach for Hand-Only and Hand-Object Pose Estimation
por: Wang, Yinqiao, et al.
Publicado: (2025)
por: Wang, Yinqiao, et al.
Publicado: (2025)
Ejemplares similares
-
Toward Phonology-Guided Sign Language Motion Generation: A Diffusion Baseline and Conditioning Analysis
por: Hong, Rui, et al.
Publicado: (2026) -
Conditional Collapse in Sign Language Production: A Diagnostic and a Scaling Argument
por: Hong, Rui, et al.
Publicado: (2026) -
VarSplat: Uncertainty-aware 3D Gaussian Splatting for Robust RGB-D SLAM
por: Tran, Anh Thuan, et al.
Publicado: (2026) -
PointSplat: Efficient Geometry-Driven Pruning and Transformer Refinement for 3D Gaussian Splatting
por: Tran, Anh Thuan, et al.
Publicado: (2026) -
EgoEV-HandPose: Egocentric 3D Hand Pose Estimation and Gesture Recognition with Stereo Event Cameras
por: Wang, Luming, et al.
Publicado: (2026)