Novel Diffusion Models for Multimodal 3D Hand Trajectory Prediction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Junyi, Bao, Wentao, Xu, Jingyi, Sun, Guanzhong, Chen, Xieyuanli, Wang, Hesheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views
von: Ma, Junyi, et al.
Veröffentlicht: (2025)
von: Ma, Junyi, et al.
Veröffentlicht: (2025)
Cam4DOcc: Benchmark for Camera-Only 4D Occupancy Forecasting in Autonomous Driving Applications
von: Ma, Junyi, et al.
Veröffentlicht: (2023)
von: Ma, Junyi, et al.
Veröffentlicht: (2023)
Spatiotemporal Decoupling for Efficient Vision-Based Occupancy Forecasting
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
HandDiffuse: Generative Controllers for Two-Hand Interactions via Diffusion Models
von: Lin, Pei, et al.
Veröffentlicht: (2023)
von: Lin, Pei, et al.
Veröffentlicht: (2023)
VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction
von: Chen, Xun, et al.
Veröffentlicht: (2026)
von: Chen, Xun, et al.
Veröffentlicht: (2026)
Explicit Interaction for Fusion-Based Place Recognition
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)
GSPR: Multimodal Place Recognition Using 3D Gaussian Splatting for Autonomous Driving
von: Qi, Zhangshuo, et al.
Veröffentlicht: (2024)
von: Qi, Zhangshuo, et al.
Veröffentlicht: (2024)
SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
EADReg: Probabilistic Correspondence Generation with Efficient Autoregressive Diffusion Model for Outdoor Point Cloud Registration
von: Gong, Linrui, et al.
Veröffentlicht: (2024)
von: Gong, Linrui, et al.
Veröffentlicht: (2024)
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
von: Bao, Chen, et al.
Veröffentlicht: (2024)
von: Bao, Chen, et al.
Veröffentlicht: (2024)
3D Hand Mesh-Guided AI-Generated Malformed Hand Refinement with Hand Pose Transformation via Diffusion Model
von: Feng, Chen-Bin, et al.
Veröffentlicht: (2025)
von: Feng, Chen-Bin, et al.
Veröffentlicht: (2025)
Efficient Multimodal 3D Object Detector via Instance-Level Contrastive Distillation
von: Su, Zhuoqun, et al.
Veröffentlicht: (2025)
von: Su, Zhuoqun, et al.
Veröffentlicht: (2025)
SemAlign3D: Semantic Correspondence between RGB-Images through Aligning 3D Object-Class Representations
von: Wandel, Krispin, et al.
Veröffentlicht: (2025)
von: Wandel, Krispin, et al.
Veröffentlicht: (2025)
BOTH2Hands: Inferring 3D Hands from Both Text Prompts and Body Dynamics
von: Zhang, Wenqian, et al.
Veröffentlicht: (2023)
von: Zhang, Wenqian, et al.
Veröffentlicht: (2023)
Intention Enhanced Diffusion Model for Multimodal Pedestrian Trajectory Prediction
von: Liu, Yu, et al.
Veröffentlicht: (2025)
von: Liu, Yu, et al.
Veröffentlicht: (2025)
Predicting 4D Hand Trajectory from Monocular Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2025)
von: Ye, Yufei, et al.
Veröffentlicht: (2025)
Ego-centric Predictive Model Conditioned on Hand Trajectories
von: Zhang, Binjie, et al.
Veröffentlicht: (2025)
von: Zhang, Binjie, et al.
Veröffentlicht: (2025)
VGGT-MPR: VGGT-Enhanced Multimodal Place Recognition in Autonomous Driving Environments
von: Xu, Jingyi, et al.
Veröffentlicht: (2026)
von: Xu, Jingyi, et al.
Veröffentlicht: (2026)
Sortblock: Similarity-Aware Feature Reuse for Diffusion Model
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
GDTS: Goal-Guided Diffusion Model with Tree Sampling for Multi-Modal Pedestrian Trajectory Prediction
von: Sun, Ge, et al.
Veröffentlicht: (2023)
von: Sun, Ge, et al.
Veröffentlicht: (2023)
LinK3D: Linear Keypoints Representation for 3D LiDAR Point Cloud
von: Cui, Yunge, et al.
Veröffentlicht: (2022)
von: Cui, Yunge, et al.
Veröffentlicht: (2022)
Optimizing Diffusion Models for Joint Trajectory Prediction and Controllable Generation
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
von: Wang, Yixiao, et al.
Veröffentlicht: (2024)
Fusion4CA: Boosting 3D Object Detection via Comprehensive Image Exploitation
von: Luo, Kang, et al.
Veröffentlicht: (2026)
von: Luo, Kang, et al.
Veröffentlicht: (2026)
Diffusion^2: Dual Diffusion Model with Uncertainty-Aware Adaptive Noise for Momentary Trajectory Prediction
von: Luo, Yuhao, et al.
Veröffentlicht: (2025)
von: Luo, Yuhao, et al.
Veröffentlicht: (2025)
HandDiff: 3D Hand Pose Estimation with Diffusion on Image-Point Cloud
von: Cheng, Wencan, et al.
Veröffentlicht: (2024)
von: Cheng, Wencan, et al.
Veröffentlicht: (2024)
DifFlow3D: Toward Robust Uncertainty-Aware Scene Flow Estimation with Diffusion Model
von: Liu, Jiuming, et al.
Veröffentlicht: (2023)
von: Liu, Jiuming, et al.
Veröffentlicht: (2023)
Robot Learning from Human Videos: A Survey
von: Ma, Junyi, et al.
Veröffentlicht: (2026)
von: Ma, Junyi, et al.
Veröffentlicht: (2026)
EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2025)
von: Ma, Junyi, et al.
Veröffentlicht: (2025)
LCPR: A Multi-Scale Attention-Based LiDAR-Camera Fusion Network for Place Recognition
von: Zhou, Zijie, et al.
Veröffentlicht: (2023)
von: Zhou, Zijie, et al.
Veröffentlicht: (2023)
SIGHT: Synthesizing Image-Text Conditioned and Geometry-Guided 3D Hand-Object Trajectories
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
GeoHand: Unlocking Prior Geometry Knowledge for Monocular 3D Hand Reconstruction
von: Lin, Weiquan, et al.
Veröffentlicht: (2026)
von: Lin, Weiquan, et al.
Veröffentlicht: (2026)
HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions
von: Xu, Hao, et al.
Veröffentlicht: (2024)
von: Xu, Hao, et al.
Veröffentlicht: (2024)
MID: A Self-supervised Multimodal Iterative Denoising Framework
von: Nie, Chang, et al.
Veröffentlicht: (2025)
von: Nie, Chang, et al.
Veröffentlicht: (2025)
BEVDiffLoc: End-to-End LiDAR Global Localization in BEV View based on Diffusion Model
von: Wang, Ziyue, et al.
Veröffentlicht: (2025)
von: Wang, Ziyue, et al.
Veröffentlicht: (2025)
Affordance-Guided Diffusion Prior for 3D Hand Reconstruction
von: Suzuki, Naru, et al.
Veröffentlicht: (2025)
von: Suzuki, Naru, et al.
Veröffentlicht: (2025)
Uncovering the human motion pattern: Pattern Memory-based Diffusion Model for Trajectory Prediction
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
von: Yang, Yuxin, et al.
Veröffentlicht: (2024)
Teacher-Student Diffusion Model for Text-Driven 3D Hand Motion Generation
von: Cheng, Ching-Lam, et al.
Veröffentlicht: (2026)
von: Cheng, Ching-Lam, et al.
Veröffentlicht: (2026)
NL2Contact: Natural Language Guided 3D Hand-Object Contact Modeling with Diffusion Model
von: Zhang, Zhongqun, et al.
Veröffentlicht: (2024)
von: Zhang, Zhongqun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024) -
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024) -
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views
von: Ma, Junyi, et al.
Veröffentlicht: (2025) -
Cam4DOcc: Benchmark for Camera-Only 4D Occupancy Forecasting in Autonomous Driving Applications
von: Ma, Junyi, et al.
Veröffentlicht: (2023) -
Spatiotemporal Decoupling for Efficient Vision-Based Occupancy Forecasting
von: Xu, Jingyi, et al.
Veröffentlicht: (2024)