BridgeACT: Bridging Human Demonstrations to Robot Actions via Unified Tool-Target Affordances
Fuente:
arXiv
Saved in:
| Main Authors: | Han, Yifan, Liu, Jianxiang, Zhang, Haoyu, Gu, Yuqi, Guo, Yunhan, Lian, Wenzhao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FSAG: Enhancing Human-to-Dexterous-Hand Finger-Specific Affordance Grounding via Diffusion Models
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
Haptic-ACT: Bridging Human Intuition with Compliant Robotic Manipulation via Immersive VR
by: Li, Kelin, et al.
Published: (2024)
by: Li, Kelin, et al.
Published: (2024)
CLASH: Collision Learning via Augmented Sim-to-real Hybridization to Bridge the Reality Gap
by: He, Haotian, et al.
Published: (2026)
by: He, Haotian, et al.
Published: (2026)
Act, Sense, Act: Learning Non-Markovian Active Perception Strategies from Large-Scale Egocentric Human Data
by: Li, Jialiang, et al.
Published: (2026)
by: Li, Jialiang, et al.
Published: (2026)
Vision in Action: Learning Active Perception from Human Demonstrations
by: Xiong, Haoyu, et al.
Published: (2025)
by: Xiong, Haoyu, et al.
Published: (2025)
Learning Spatial Bimanual Action Models Based on Affordance Regions and Human Demonstrations
by: Plonka, Björn S., et al.
Published: (2024)
by: Plonka, Björn S., et al.
Published: (2024)
Look, Zoom, Understand: The Robotic Eyeball for Embodied Perception
by: Yang, Jiashu, et al.
Published: (2025)
by: Yang, Jiashu, et al.
Published: (2025)
BORA: Bridging Offline Reinforcement Learning and Online Residual Adaptation for Real-World Dexterous VLA Models
by: Chen, Zhongxi, et al.
Published: (2026)
by: Chen, Zhongxi, et al.
Published: (2026)
DexCanvas: Bridging Human Demonstrations and Robot Learning for Dexterous Manipulation
by: Xu, Xinyue, et al.
Published: (2025)
by: Xu, Xinyue, et al.
Published: (2025)
MagiClaw: A Dual-Use, Vision-Based Soft Gripper for Bridging the Human Demonstration to Robotic Deployment Gap
by: Wu, Tianyu, et al.
Published: (2025)
by: Wu, Tianyu, et al.
Published: (2025)
Bridging Scale Discrepancies in Robotic Control via Language-Based Action Representations
by: Zhang, Yuchi, et al.
Published: (2025)
by: Zhang, Yuchi, et al.
Published: (2025)
Towards Affordance-Aware Robotic Dexterous Grasping with Human-like Priors
by: Zhao, Haoyu, et al.
Published: (2025)
by: Zhao, Haoyu, et al.
Published: (2025)
BifrostUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation
by: Yu, Chenhao, et al.
Published: (2026)
by: Yu, Chenhao, et al.
Published: (2026)
Agentic Scene Policies: Unifying Space, Semantics, and Affordances for Robot Action
by: Morin, Sacha, et al.
Published: (2025)
by: Morin, Sacha, et al.
Published: (2025)
PoseDiff: A Unified Diffusion Model Bridging Robot Pose Estimation and Video-to-Action Control
by: Zhang, Haozhuo, et al.
Published: (2025)
by: Zhang, Haozhuo, et al.
Published: (2025)
Viewpoint Matters: Dynamically Optimizing Viewpoints with Masked Autoencoder for Visual Manipulation
by: Yi, Pengfei, et al.
Published: (2026)
by: Yi, Pengfei, et al.
Published: (2026)
DexHiL: A Human-in-the-Loop Framework for Vision-Language-Action Model Post-Training in Dexterous Manipulation
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
RoboPCA: Pose-centered Affordance Learning from Human Demonstrations for Robot Manipulation
by: Xiao, Zhanqi, et al.
Published: (2026)
by: Xiao, Zhanqi, et al.
Published: (2026)
FoAM: Foresight-Augmented Multi-Task Imitation Policy for Robotic Manipulation
by: Liu, Litao, et al.
Published: (2024)
by: Liu, Litao, et al.
Published: (2024)
Semantic-Geometric Task Representations for Bimanual Manipulation from Human Demonstrations to Robot Action Planning
by: Herbert, Franziska, et al.
Published: (2026)
by: Herbert, Franziska, et al.
Published: (2026)
Boosting Action-Information via a Variational Bottleneck on Unlabelled Robot Videos
by: Zhang, Haoyu, et al.
Published: (2025)
by: Zhang, Haoyu, et al.
Published: (2025)
AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
Learning Human-Intention Priors from Large-Scale Human Demonstrations for Robotic Manipulation
by: Xie, Yifan, et al.
Published: (2026)
by: Xie, Yifan, et al.
Published: (2026)
SAGE: Scene Graph-Aware Guidance and Execution for Long-Horizon Manipulation Tasks
by: Li, Jialiang, et al.
Published: (2025)
by: Li, Jialiang, et al.
Published: (2025)
Bridging Language and Action: A Survey of Language-Conditioned Robot Manipulation
by: Yao, Xiangtong, et al.
Published: (2023)
by: Yao, Xiangtong, et al.
Published: (2023)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
by: Zahid, Azizul, et al.
Published: (2025)
by: Zahid, Azizul, et al.
Published: (2025)
ToolEENet: Tool Affordance 6D Pose Estimation
by: Wang, Yunlong, et al.
Published: (2024)
by: Wang, Yunlong, et al.
Published: (2024)
3D Affordance Keypoint Detection for Robotic Manipulation
by: Liu, Zhiyang, et al.
Published: (2025)
by: Liu, Zhiyang, et al.
Published: (2025)
UniAff: A Unified Representation of Affordances for Tool Usage and Articulation with Vision-Language Models
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
DRAW2ACT: Turning Depth-Encoded Trajectories into Robotic Demonstration Videos
by: Bai, Yang, et al.
Published: (2025)
by: Bai, Yang, et al.
Published: (2025)
Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization
by: Yang, Jonathan, et al.
Published: (2025)
by: Yang, Jonathan, et al.
Published: (2025)
ExACT: An End-to-End Autonomous Excavator System Using Action Chunking With Transformers
by: Chen, Liangliang, et al.
Published: (2024)
by: Chen, Liangliang, et al.
Published: (2024)
From LLMs to Actions: Latent Codes as Bridges in Hierarchical Robot Control
by: Shentu, Yide, et al.
Published: (2024)
by: Shentu, Yide, et al.
Published: (2024)
Bridging Language, Vision and Action: Multimodal VAEs in Robotic Manipulation Tasks
by: Sejnova, Gabriela, et al.
Published: (2024)
by: Sejnova, Gabriela, et al.
Published: (2024)
The Wilhelm Tell Dataset of Affordance Demonstrations
by: Ringe, Rachel, et al.
Published: (2025)
by: Ringe, Rachel, et al.
Published: (2025)
Unified Learning from Demonstrations, Corrections, and Preferences during Physical Human-Robot Interaction
by: Mehta, Shaunak A., et al.
Published: (2022)
by: Mehta, Shaunak A., et al.
Published: (2022)
RwoR: Generating Robot Demonstrations from Human Hand Collection for Policy Learning without Robot
by: Heng, Liang, et al.
Published: (2025)
by: Heng, Liang, et al.
Published: (2025)
SymBridge: A Human-in-the-Loop Cyber-Physical Interactive System for Adaptive Human-Robot Symbiosis
by: Chen, Haoran, et al.
Published: (2025)
by: Chen, Haoran, et al.
Published: (2025)
PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation
by: Liu, Yuanzhe, et al.
Published: (2026)
by: Liu, Yuanzhe, et al.
Published: (2026)
HRP: Human Affordances for Robotic Pre-Training
by: Srirama, Mohan Kumar, et al.
Published: (2024)
by: Srirama, Mohan Kumar, et al.
Published: (2024)
Similar Items
-
FSAG: Enhancing Human-to-Dexterous-Hand Finger-Specific Affordance Grounding via Diffusion Models
by: Han, Yifan, et al.
Published: (2026) -
Haptic-ACT: Bridging Human Intuition with Compliant Robotic Manipulation via Immersive VR
by: Li, Kelin, et al.
Published: (2024) -
CLASH: Collision Learning via Augmented Sim-to-real Hybridization to Bridge the Reality Gap
by: He, Haotian, et al.
Published: (2026) -
Act, Sense, Act: Learning Non-Markovian Active Perception Strategies from Large-Scale Egocentric Human Data
by: Li, Jialiang, et al.
Published: (2026) -
Vision in Action: Learning Active Perception from Human Demonstrations
by: Xiong, Haoyu, et al.
Published: (2025)