Multi-Modal Manipulation via Multi-Modal Policy Consensus
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Haonan, Xu, Jiaming, Chen, Hongyu, Hong, Kaiwen, Huang, Binghao, Liu, Chaoqi, Mao, Jiayuan, Li, Yunzhu, Du, Yilun, Driggs-Campbell, Katherine |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Coordinated Bimanual Manipulation Policies using State Diffusion and Inverse Dynamics Models
by: Chen, Haonan, et al.
Published: (2025)
by: Chen, Haonan, et al.
Published: (2025)
Tool-as-Interface: Learning Robot Policies from Observing Human Tool Use
by: Chen, Haonan, et al.
Published: (2025)
by: Chen, Haonan, et al.
Published: (2025)
Flexible Multitask Learning with Factorized Diffusion Policy
by: Liu, Chaoqi, et al.
Published: (2025)
by: Liu, Chaoqi, et al.
Published: (2025)
Towards Uncertainty Unification: A Case Study for Preference Learning
by: Peng, Shaoting, et al.
Published: (2025)
by: Peng, Shaoting, et al.
Published: (2025)
DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors
by: Wang, Pengcheng, et al.
Published: (2026)
by: Wang, Pengcheng, et al.
Published: (2026)
Neural Informed RRT*: Learning-based Path Planning with Point Cloud State Representations under Admissible Ellipsoidal Constraints
by: Huang, Zhe, et al.
Published: (2023)
by: Huang, Zhe, et al.
Published: (2023)
Touch in the Wild: Learning Fine-Grained Manipulation with a Portable Visuo-Tactile Gripper
by: Zhu, Xinyue, et al.
Published: (2025)
by: Zhu, Xinyue, et al.
Published: (2025)
FlexiTac: A Low-Cost, Open-Source, Scalable Tactile Sensing Solution for Robotic Systems
by: Huang, Binghao, et al.
Published: (2026)
by: Huang, Binghao, et al.
Published: (2026)
Structured Graph Network for Constrained Robot Crowd Navigation with Low Fidelity Simulation
by: Liu, Shuijing, et al.
Published: (2024)
by: Liu, Shuijing, et al.
Published: (2024)
Localized Graph-Based Neural Dynamics Models for Terrain Manipulation
by: Liu, Chaoqi, et al.
Published: (2025)
by: Liu, Chaoqi, et al.
Published: (2025)
Learning Multi-Modal Trajectory Policies for Data-Efficient Robotic Manipulation
by: Chen, Zijia, et al.
Published: (2026)
by: Chen, Zijia, et al.
Published: (2026)
Do You Know the Way? Human-in-the-Loop Understanding for Fast Traversability Estimation in Mobile Robotics
by: Schreiber, Andre, et al.
Published: (2025)
by: Schreiber, Andre, et al.
Published: (2025)
Learning Tactile-Aware Quadrupedal Loco-Manipulation Policies
by: Zhou, Pokuang, et al.
Published: (2026)
by: Zhou, Pokuang, et al.
Published: (2026)
COMMET: A System for Human-Induced Conflicts in Mobile Manipulation of Everyday Tasks
by: Li, Dongping, et al.
Published: (2025)
by: Li, Dongping, et al.
Published: (2025)
Topology-Guided ORCA: Smooth Multi-Agent Motion Planning in Constrained Environments
by: Pouria, Fatemeh Cheraghi, et al.
Published: (2024)
by: Pouria, Fatemeh Cheraghi, et al.
Published: (2024)
3D-ViTac: Learning Fine-Grained Manipulation with Visuo-Tactile Sensing
by: Huang, Binghao, et al.
Published: (2024)
by: Huang, Binghao, et al.
Published: (2024)
Trace-Focused Diffusion Policy for Multi-Modal Action Disambiguation in Long-Horizon Robotic Manipulation
by: Hu, Yuxuan, et al.
Published: (2026)
by: Hu, Yuxuan, et al.
Published: (2026)
VFP: Variational Flow-Matching Policy for Multi-Modal Robot Manipulation
by: Zhai, Xuanran, et al.
Published: (2025)
by: Zhai, Xuanran, et al.
Published: (2025)
Learning Force-Regulated Manipulation with a Low-Cost Tactile-Force-Controlled Gripper
by: Kang, Xuhui, et al.
Published: (2026)
by: Kang, Xuhui, et al.
Published: (2026)
Trust-Aware Embodied Bayesian Persuasion for Mixed-Autonomy
by: Peng, Shaoting, et al.
Published: (2025)
by: Peng, Shaoting, et al.
Published: (2025)
A Brief Survey on Leveraging Large Scale Vision Models for Enhanced Robot Grasping
by: Kamboj, Abhi, et al.
Published: (2024)
by: Kamboj, Abhi, et al.
Published: (2024)
OAT: Ordered Action Tokenization
by: Liu, Chaoqi, et al.
Published: (2026)
by: Liu, Chaoqi, et al.
Published: (2026)
Hybrid Consistency Policy: Decoupling Multi-Modal Diversity and Real-Time Efficiency in Robotic Manipulation
by: Zhao, Qianyou, et al.
Published: (2025)
by: Zhao, Qianyou, et al.
Published: (2025)
Efficient and Reliable Teleoperation through Real-to-Sim-to-Real Shared Autonomy
by: Sha, Shuo, et al.
Published: (2026)
by: Sha, Shuo, et al.
Published: (2026)
SIMPACT: Simulation-Enabled Action Planning using Vision-Language Models
by: Liu, Haowen, et al.
Published: (2025)
by: Liu, Haowen, et al.
Published: (2025)
Rethinking Gaussian Trajectory Predictors: Calibrated Uncertainty for Safe Planning
by: Pouria, Fatemeh Cheraghi, et al.
Published: (2026)
by: Pouria, Fatemeh Cheraghi, et al.
Published: (2026)
TacUMI: A Multi-Modal Universal Manipulation Interface for Contact-Rich Tasks
by: Cheng, Tailai, et al.
Published: (2026)
by: Cheng, Tailai, et al.
Published: (2026)
Human-Agent Joint Learning for Efficient Robot Manipulation Skill Acquisition
by: Luo, Shengcheng, et al.
Published: (2024)
by: Luo, Shengcheng, et al.
Published: (2024)
MMScan: A Multi-Modal 3D Scene Dataset with Hierarchical Grounded Language Annotations
by: Lyu, Ruiyuan, et al.
Published: (2024)
by: Lyu, Ruiyuan, et al.
Published: (2024)
Hallucination Detection in Foundation Models for Decision-Making: A Flexible Definition and Review of the State of the Art
by: Chakraborty, Neeloy, et al.
Published: (2024)
by: Chakraborty, Neeloy, et al.
Published: (2024)
Tactile-Based Human Intent Recognition for Robot Assistive Navigation
by: Peng, Shaoting, et al.
Published: (2025)
by: Peng, Shaoting, et al.
Published: (2025)
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy
by: Wang, Yixuan, et al.
Published: (2024)
by: Wang, Yixuan, et al.
Published: (2024)
In-Situ Soil-Property Estimation and Bayesian Mapping with a Simulated Compact Track Loader
by: Wagner, W. Jacob, et al.
Published: (2025)
by: Wagner, W. Jacob, et al.
Published: (2025)
Hierarchical Intention Tracking with Switching Trees for Real-Time Adaptation to Dynamic Human Intentions during Collaboration
by: Huang, Zhe, et al.
Published: (2025)
by: Huang, Zhe, et al.
Published: (2025)
EMMOE: A Comprehensive Benchmark for Embodied Mobile Manipulation in Open Environments
by: Li, Dongping, et al.
Published: (2025)
by: Li, Dongping, et al.
Published: (2025)
LIT: Large Language Model Driven Intention Tracking for Proactive Human-Robot Collaboration -- A Robot Sous-Chef Application
by: Huang, Zhe, et al.
Published: (2024)
by: Huang, Zhe, et al.
Published: (2024)
Design and Benchmarking of A Multi-Modality Sensor for Robotic Manipulation with GAN-Based Cross-Modality Interpretation
by: Zhang, Dandan, et al.
Published: (2025)
by: Zhang, Dandan, et al.
Published: (2025)
ADM-DP: Adaptive Dynamic Modality Diffusion Policy through Vision-Tactile-Graph Fusion for Multi-Agent Manipulation
by: Wang, Enyi, et al.
Published: (2026)
by: Wang, Enyi, et al.
Published: (2026)
GR-MG: Leveraging Partially Annotated Data via Multi-Modal Goal-Conditioned Policy
by: Li, Peiyan, et al.
Published: (2024)
by: Li, Peiyan, et al.
Published: (2024)
RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation
by: Jiang, Hanxiao, et al.
Published: (2024)
by: Jiang, Hanxiao, et al.
Published: (2024)
Similar Items
-
Learning Coordinated Bimanual Manipulation Policies using State Diffusion and Inverse Dynamics Models
by: Chen, Haonan, et al.
Published: (2025) -
Tool-as-Interface: Learning Robot Policies from Observing Human Tool Use
by: Chen, Haonan, et al.
Published: (2025) -
Flexible Multitask Learning with Factorized Diffusion Policy
by: Liu, Chaoqi, et al.
Published: (2025) -
Towards Uncertainty Unification: A Case Study for Preference Learning
by: Peng, Shaoting, et al.
Published: (2025) -
DiscreteRTC: Discrete Diffusion Policies are Natural Asynchronous Executors
by: Wang, Pengcheng, et al.
Published: (2026)