Open-world Hand-Object Interaction Video Generation Based on Structure and Contact-aware Representation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Yan, Haodong, Yu, Hang, Zhong, Zhide, Yuan, Weilin, Gong, Xin, Luo, Zehang, Heyu, Chengxi, Li, Junfeng, Song, Wenxuan, Zhou, Shunbo, Li, Haoang |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
VLA-OPD: Bridging Offline SFT and Online RL for Vision-Language-Action Models via On-Policy Distillation
par: Zhong, Zhide, et autres
Publié: (2026)
par: Zhong, Zhide, et autres
Publié: (2026)
FlowVLA: Visual Chain of Thought-based Motion Reasoning for Vision-Language-Action Models
par: Zhong, Zhide, et autres
Publié: (2025)
par: Zhong, Zhide, et autres
Publié: (2025)
S-VAM: Shortcut Video-Action Model by Self-Distilling Geometric and Semantic Foresight
par: Yan, Haodong, et autres
Publié: (2026)
par: Yan, Haodong, et autres
Publié: (2026)
Dynamic Reconstruction of Hand-Object Interaction with Distributed Force-aware Contact Representation
par: Yu, Zhenjun, et autres
Publié: (2024)
par: Yu, Zhenjun, et autres
Publié: (2024)
CHOIR: Contact-aware 4D Hand-Object Interaction Reconstruction
par: Xu, Hao, et autres
Publié: (2026)
par: Xu, Hao, et autres
Publié: (2026)
DualCoT-VLA: Visual-Linguistic Chain of Thought via Parallel Reasoning for Vision-Language-Action Models
par: Zhong, Zhide, et autres
Publié: (2026)
par: Zhong, Zhide, et autres
Publié: (2026)
Physics-aware Hand-object Interaction Denoising
par: Luo, Haowen, et autres
Publié: (2024)
par: Luo, Haowen, et autres
Publié: (2024)
HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness
par: Xue, Zihui, et autres
Publié: (2024)
par: Xue, Zihui, et autres
Publié: (2024)
Interaction-aware Representation Modeling with Co-occurrence Consistency for Egocentric Hand-Object Parsing
par: Su, Yuejiao, et autres
Publié: (2026)
par: Su, Yuejiao, et autres
Publié: (2026)
VE2VF: Vision-Enabled to Vision-Free Distillation via Real-world Reinforcement Learning for Robust Contact-Rich Manipulation
par: Kowalski, Victor, et autres
Publié: (2026)
par: Kowalski, Victor, et autres
Publié: (2026)
SpriteHand: Real-Time Versatile Hand-Object Interaction with Autoregressive Video Generation
par: Li, Zisu, et autres
Publié: (2025)
par: Li, Zisu, et autres
Publié: (2025)
Hand-Object Interaction Pretraining from Videos
par: Singh, Himanshu Gaurav, et autres
Publié: (2024)
par: Singh, Himanshu Gaurav, et autres
Publié: (2024)
CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation
par: Luo, Xiangyang, et autres
Publié: (2026)
par: Luo, Xiangyang, et autres
Publié: (2026)
Interactive Perception for Deformable Object Manipulation
par: Weng, Zehang, et autres
Publié: (2024)
par: Weng, Zehang, et autres
Publié: (2024)
Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model
par: Fan, Yingying, et autres
Publié: (2025)
par: Fan, Yingying, et autres
Publié: (2025)
A Versatile and Differentiable Hand-Object Interaction Representation
par: Morales, Théo, et autres
Publié: (2024)
par: Morales, Théo, et autres
Publié: (2024)
iDiT-HOI: Inpainting-based Hand Object Interaction Reenactment via Video Diffusion Transformer
par: Shen, Zhelun, et autres
Publié: (2025)
par: Shen, Zhelun, et autres
Publié: (2025)
Grasping a Handful: Sequential Multi-Object Dexterous Grasp Generation
par: Lu, Haofei, et autres
Publié: (2025)
par: Lu, Haofei, et autres
Publié: (2025)
PD-VLA: Accelerating Vision-Language-Action Model Integrated with Action Chunking via Parallel Decoding
par: Song, Wenxuan, et autres
Publié: (2025)
par: Song, Wenxuan, et autres
Publié: (2025)
CaRe-Ego: Contact-aware Relationship Modeling for Egocentric Interactive Hand-object Segmentation
par: Su, Yuejiao, et autres
Publié: (2024)
par: Su, Yuejiao, et autres
Publié: (2024)
SynHLMA:Synthesizing Hand Language Manipulation for Articulated Object with Discrete Human Object Interaction Representation
par: zhi, Wang, et autres
Publié: (2025)
par: zhi, Wang, et autres
Publié: (2025)
Rethinking the Practicality of Vision-language-action Model: A Comprehensive Benchmark and An Improved Baseline
par: Song, Wenxuan, et autres
Publié: (2026)
par: Song, Wenxuan, et autres
Publié: (2026)
DyGeoVLN: Infusing Dynamic Geometry Foundation Model into Vision-Language Navigation
par: Liu, Xiangchen, et autres
Publié: (2026)
par: Liu, Xiangchen, et autres
Publié: (2026)
NCRF: Neural Contact Radiance Fields for Free-Viewpoint Rendering of Hand-Object Interaction
par: Zhang, Zhongqun, et autres
Publié: (2024)
par: Zhang, Zhongqun, et autres
Publié: (2024)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
par: Zhu, Zhifan, et autres
Publié: (2025)
par: Zhu, Zhifan, et autres
Publié: (2025)
SCAR: Semantic Cardiac Adversarial Representation via Spatiotemporal Manifold Optimization in ECG
par: Jia, Shunbo, et autres
Publié: (2025)
par: Jia, Shunbo, et autres
Publié: (2025)
DexTwist: Dexterous Hand Retargeting for Twist Motion via Mixed Reality-based Teleoperation
par: Lee, Dongmyoung, et autres
Publié: (2026)
par: Lee, Dongmyoung, et autres
Publié: (2026)
CPR: Causal Physiological Representation Learning for Robust ECG Analysis under Distribution Shifts
par: Jia, Shunbo, et autres
Publié: (2025)
par: Jia, Shunbo, et autres
Publié: (2025)
Continual Hand-Eye Calibration for Open-world Robotic Manipulation
par: Li, Fazeng, et autres
Publié: (2026)
par: Li, Fazeng, et autres
Publié: (2026)
SU-YOLO: Spiking Neural Network for Efficient Underwater Object Detection
par: Li, Chenyang, et autres
Publié: (2025)
par: Li, Chenyang, et autres
Publié: (2025)
ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction Videos
par: Chen, Yuantao, et autres
Publié: (2026)
par: Chen, Yuantao, et autres
Publié: (2026)
Spatial Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model
par: Li, Fuhao, et autres
Publié: (2025)
par: Li, Fuhao, et autres
Publié: (2025)
PEAR: Phrase-Based Hand-Object Interaction Anticipation
par: Zhang, Zichen, et autres
Publié: (2024)
par: Zhang, Zichen, et autres
Publié: (2024)
OpenHOI: Open-World Hand-Object Interaction Synthesis with Multimodal Large Language Model
par: Zhang, Zhenhao, et autres
Publié: (2025)
par: Zhang, Zhenhao, et autres
Publié: (2025)
Modeling Fine-Grained Hand-Object Dynamics for Egocentric Video Representation Learning
par: Pei, Baoqi, et autres
Publié: (2025)
par: Pei, Baoqi, et autres
Publié: (2025)
Hand-Object Contact Detection using Grasp Quality Metrics
par: Nguyen, Thanh Vinh, et autres
Publié: (2025)
par: Nguyen, Thanh Vinh, et autres
Publié: (2025)
DyTact: Capturing Dynamic Contacts in Hand-Object Manipulation
par: Cong, Xiaoyan, et autres
Publié: (2025)
par: Cong, Xiaoyan, et autres
Publié: (2025)
HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions
par: Xu, Hao, et autres
Publié: (2024)
par: Xu, Hao, et autres
Publié: (2024)
InterRVOS: Interaction-aware Referring Video Object Segmentation
par: Jin, Woojeong, et autres
Publié: (2025)
par: Jin, Woojeong, et autres
Publié: (2025)
Learning Explicit Contact for Implicit Reconstruction of Hand-held Objects from Monocular Images
par: Hu, Junxing, et autres
Publié: (2023)
par: Hu, Junxing, et autres
Publié: (2023)
Documents similaires
-
VLA-OPD: Bridging Offline SFT and Online RL for Vision-Language-Action Models via On-Policy Distillation
par: Zhong, Zhide, et autres
Publié: (2026) -
FlowVLA: Visual Chain of Thought-based Motion Reasoning for Vision-Language-Action Models
par: Zhong, Zhide, et autres
Publié: (2025) -
S-VAM: Shortcut Video-Action Model by Self-Distilling Geometric and Semantic Foresight
par: Yan, Haodong, et autres
Publié: (2026) -
Dynamic Reconstruction of Hand-Object Interaction with Distributed Force-aware Contact Representation
par: Yu, Zhenjun, et autres
Publié: (2024) -
CHOIR: Contact-aware 4D Hand-Object Interaction Reconstruction
par: Xu, Hao, et autres
Publié: (2026)