MADiff: Motion-Aware Mamba Diffusion Models for Hand Trajectory Prediction on Egocentric Videos
Fuente:
arXiv
Salvato in:
| Autori principali: | Ma, Junyi, Chen, Xieyuanli, Bao, Wentao, Xu, Jingyi, Wang, Hesheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
di: Ma, Junyi, et al.
Pubblicazione: (2024)
di: Ma, Junyi, et al.
Pubblicazione: (2024)
Novel Diffusion Models for Multimodal 3D Hand Trajectory Prediction
di: Ma, Junyi, et al.
Pubblicazione: (2025)
di: Ma, Junyi, et al.
Pubblicazione: (2025)
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views
di: Ma, Junyi, et al.
Pubblicazione: (2025)
di: Ma, Junyi, et al.
Pubblicazione: (2025)
EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos
di: Ma, Junyi, et al.
Pubblicazione: (2025)
di: Ma, Junyi, et al.
Pubblicazione: (2025)
Zero-Shot Temporal Interaction Localization for Egocentric Videos
di: Zhang, Erhang, et al.
Pubblicazione: (2025)
di: Zhang, Erhang, et al.
Pubblicazione: (2025)
Cam4DOcc: Benchmark for Camera-Only 4D Occupancy Forecasting in Autonomous Driving Applications
di: Ma, Junyi, et al.
Pubblicazione: (2023)
di: Ma, Junyi, et al.
Pubblicazione: (2023)
TrajMamba: An Ego-Motion-Guided Mamba Model for Pedestrian Trajectory Prediction from an Egocentric Perspective
di: Peng, Yusheng, et al.
Pubblicazione: (2026)
di: Peng, Yusheng, et al.
Pubblicazione: (2026)
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
di: Zhan, Zechao, et al.
Pubblicazione: (2024)
di: Zhan, Zechao, et al.
Pubblicazione: (2024)
Flowing from Reasoning to Motion: Learning 3D Hand Trajectory Prediction from Egocentric Human Interaction Videos
di: Chen, Mingfei, et al.
Pubblicazione: (2025)
di: Chen, Mingfei, et al.
Pubblicazione: (2025)
HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos
di: Zhang, Jinglei, et al.
Pubblicazione: (2025)
di: Zhang, Jinglei, et al.
Pubblicazione: (2025)
Spatiotemporal Decoupling for Efficient Vision-Based Occupancy Forecasting
di: Xu, Jingyi, et al.
Pubblicazione: (2024)
di: Xu, Jingyi, et al.
Pubblicazione: (2024)
HandDiffuse: Generative Controllers for Two-Hand Interactions via Diffusion Models
di: Lin, Pei, et al.
Pubblicazione: (2023)
di: Lin, Pei, et al.
Pubblicazione: (2023)
Modeling Fine-Grained Hand-Object Dynamics for Egocentric Video Representation Learning
di: Pei, Baoqi, et al.
Pubblicazione: (2025)
di: Pei, Baoqi, et al.
Pubblicazione: (2025)
Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints
di: Zhang, Chenyangguang, et al.
Pubblicazione: (2026)
di: Zhang, Chenyangguang, et al.
Pubblicazione: (2026)
StableHand: Quality-Aware Flow Matching for World-Space Dual-Hand Motion Estimation from Egocentric Video
di: Zeng, Huajian, et al.
Pubblicazione: (2026)
di: Zeng, Huajian, et al.
Pubblicazione: (2026)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
di: Zhou, Bohan, et al.
Pubblicazione: (2025)
di: Zhou, Bohan, et al.
Pubblicazione: (2025)
Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?
di: Xu, Boshen, et al.
Pubblicazione: (2024)
di: Xu, Boshen, et al.
Pubblicazione: (2024)
PAD-Hand: Physics-Aware Diffusion for Hand Motion Recovery
di: Ismayilzada, Elkhan, et al.
Pubblicazione: (2026)
di: Ismayilzada, Elkhan, et al.
Pubblicazione: (2026)
VGGT-Occ: Geometry-Grounded and Density-Aware Gated Fusion for 3D Occupancy Prediction
di: Chen, Xun, et al.
Pubblicazione: (2026)
di: Chen, Xun, et al.
Pubblicazione: (2026)
Explicit Interaction for Fusion-Based Place Recognition
di: Xu, Jingyi, et al.
Pubblicazione: (2024)
di: Xu, Jingyi, et al.
Pubblicazione: (2024)
Recognizing Hand Use and Hand Role at Home After Stroke from Egocentric Video
di: Tsai, Meng-Fen, et al.
Pubblicazione: (2022)
di: Tsai, Meng-Fen, et al.
Pubblicazione: (2022)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
di: Zhang, Tianqi, et al.
Pubblicazione: (2026)
di: Zhang, Tianqi, et al.
Pubblicazione: (2026)
EgoMotion: Hierarchical Reasoning and Diffusion for Egocentric Vision-Language Motion Generation
di: Hou, Ruibing, et al.
Pubblicazione: (2026)
di: Hou, Ruibing, et al.
Pubblicazione: (2026)
Detecting Precise Hand Touch Moments in Egocentric Video
di: Nguyen, Huy Anh, et al.
Pubblicazione: (2026)
di: Nguyen, Huy Anh, et al.
Pubblicazione: (2026)
Incentivizing Temporal-Awareness in Egocentric Video Understanding Models
di: Xu, Zhiyang, et al.
Pubblicazione: (2026)
di: Xu, Zhiyang, et al.
Pubblicazione: (2026)
EMAG: Ego-motion Aware and Generalizable 2D Hand Forecasting from Egocentric Videos
di: Hatano, Masashi, et al.
Pubblicazione: (2024)
di: Hatano, Masashi, et al.
Pubblicazione: (2024)
Mamba-VGGT: Persistent Long-Sequence Video Geometry Grounded Transformer via External Sliding Window Mamba Memory
di: Deng, Tianchen, et al.
Pubblicazione: (2026)
di: Deng, Tianchen, et al.
Pubblicazione: (2026)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
di: Zhu, Zhifan, et al.
Pubblicazione: (2025)
di: Zhu, Zhifan, et al.
Pubblicazione: (2025)
Robot Learning from Human Videos: A Survey
di: Ma, Junyi, et al.
Pubblicazione: (2026)
di: Ma, Junyi, et al.
Pubblicazione: (2026)
Predicting 4D Hand Trajectory from Monocular Videos
di: Ye, Yufei, et al.
Pubblicazione: (2025)
di: Ye, Yufei, et al.
Pubblicazione: (2025)
Motion Focus Recognition in Fast-Moving Egocentric Video
di: Hong, Si-En, et al.
Pubblicazione: (2026)
di: Hong, Si-En, et al.
Pubblicazione: (2026)
Cross-Modal Action Recognition in Egocentric Video Using Mamba: Integrating RGB and Hand Skeleton Streams via CLS Token Fusion Strategies
di: Gorostegui, Juan Ignacio Bustos, et al.
Pubblicazione: (2026)
di: Gorostegui, Juan Ignacio Bustos, et al.
Pubblicazione: (2026)
EADReg: Probabilistic Correspondence Generation with Efficient Autoregressive Diffusion Model for Outdoor Point Cloud Registration
di: Gong, Linrui, et al.
Pubblicazione: (2024)
di: Gong, Linrui, et al.
Pubblicazione: (2024)
Diffusion^2: Dual Diffusion Model with Uncertainty-Aware Adaptive Noise for Momentary Trajectory Prediction
di: Luo, Yuhao, et al.
Pubblicazione: (2025)
di: Luo, Yuhao, et al.
Pubblicazione: (2025)
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
di: Bao, Chen, et al.
Pubblicazione: (2024)
di: Bao, Chen, et al.
Pubblicazione: (2024)
WHOLE: World-Grounded Hand-Object Lifted from Egocentric Videos
di: Ye, Yufei, et al.
Pubblicazione: (2026)
di: Ye, Yufei, et al.
Pubblicazione: (2026)
EgoHandICL: Egocentric 3D Hand Reconstruction with In-Context Learning
di: Xie, Binzhu, et al.
Pubblicazione: (2026)
di: Xie, Binzhu, et al.
Pubblicazione: (2026)
Vamba: Understanding Hour-Long Videos with Hybrid Mamba-Transformers
di: Ren, Weiming, et al.
Pubblicazione: (2025)
di: Ren, Weiming, et al.
Pubblicazione: (2025)
EgoNav: Egocentric Scene-aware Human Trajectory Prediction
di: Wang, Weizhuo, et al.
Pubblicazione: (2024)
di: Wang, Weizhuo, et al.
Pubblicazione: (2024)
MambaControl: Anatomy Graph-Enhanced Mamba ControlNet with Fourier Refinement for Diffusion-Based Disease Trajectory Prediction
di: Yang, Hao, et al.
Pubblicazione: (2025)
di: Yang, Hao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
di: Ma, Junyi, et al.
Pubblicazione: (2024) -
Novel Diffusion Models for Multimodal 3D Hand Trajectory Prediction
di: Ma, Junyi, et al.
Pubblicazione: (2025) -
Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views
di: Ma, Junyi, et al.
Pubblicazione: (2025) -
EgoLoc: A Generalizable Solution for Temporal Interaction Localization in Egocentric Videos
di: Ma, Junyi, et al.
Pubblicazione: (2025) -
Zero-Shot Temporal Interaction Localization for Egocentric Videos
di: Zhang, Erhang, et al.
Pubblicazione: (2025)