Saved in:
| Main Authors: | Huang, Yongxi, Wang, Zhuohang, Tang, Wenjing, Lu, Cewu, Cai, Panpan |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2603.00600 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tru-POMDP: Task Planning Under Uncertainty via Tree of Hypotheses and Open-Ended POMDPs
by: Tang, Wenjing, et al.
Published: (2025)
by: Tang, Wenjing, et al.
Published: (2025)
Mimic Intent, Not Just Trajectories
by: Huang, Renming, et al.
Published: (2026)
by: Huang, Renming, et al.
Published: (2026)
UniPlan: Vision-Language Task Planning for Mobile Manipulation with Unified PDDL Formulation
by: Ye, Haoming, et al.
Published: (2026)
by: Ye, Haoming, et al.
Published: (2026)
UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning
by: Ye, Haoming, et al.
Published: (2025)
by: Ye, Haoming, et al.
Published: (2025)
Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks
by: Liu, Zhihong, et al.
Published: (2026)
by: Liu, Zhihong, et al.
Published: (2026)
RISE: 3D Perception Makes Real-World Robot Imitation Simple and Effective
by: Wang, Chenxi, et al.
Published: (2024)
by: Wang, Chenxi, et al.
Published: (2024)
DexTOG: Learning Task-Oriented Dexterous Grasp with Language
by: Zhang, Jieyi, et al.
Published: (2025)
by: Zhang, Jieyi, et al.
Published: (2025)
Active Semantic Perception
by: Tang, Huayi, et al.
Published: (2025)
by: Tang, Huayi, et al.
Published: (2025)
ActiveGlasses: Learning Manipulation with Active Vision from Ego-centric Human Demonstration
by: Zou, Yanwen, et al.
Published: (2026)
by: Zou, Yanwen, et al.
Published: (2026)
ForceVLA: Enhancing VLA Models with a Force-aware MoE for Contact-rich Manipulation
by: Yu, Jiawen, et al.
Published: (2025)
by: Yu, Jiawen, et al.
Published: (2025)
A Surprisingly Efficient Representation for Multi-Finger Grasping
by: Yan, Hengxu, et al.
Published: (2024)
by: Yan, Hengxu, et al.
Published: (2024)
HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid
by: Xu, Xinyu, et al.
Published: (2024)
by: Xu, Xinyu, et al.
Published: (2024)
FSGlove: An Inertial-Based Hand Tracking System with Shape-Aware Calibration
by: Li, Yutong, et al.
Published: (2025)
by: Li, Yutong, et al.
Published: (2025)
An Aerial Manipulator for Perception-Driven Flower Targeting Toward Contactless Pollination in Vertical Farming
by: Jin, Chenzhe, et al.
Published: (2026)
by: Jin, Chenzhe, et al.
Published: (2026)
Kalib: Easy Hand-Eye Calibration with Reference Point Tracking
by: Tang, Tutian, et al.
Published: (2024)
by: Tang, Tutian, et al.
Published: (2024)
RPMArt: Towards Robust Perception and Manipulation for Articulated Objects
by: Wang, Junbo, et al.
Published: (2024)
by: Wang, Junbo, et al.
Published: (2024)
Chat with UAV -- Human-UAV Interaction Based on Large Language Models
by: Wang, Haoran, et al.
Published: (2025)
by: Wang, Haoran, et al.
Published: (2025)
Digital Gene: Learning about the Physical World through Analytic Concepts
by: Sun, Jianhua, et al.
Published: (2025)
by: Sun, Jianhua, et al.
Published: (2025)
Neural Randomized Planning for Whole Body Robot Motion
by: Lu, Yunfan, et al.
Published: (2024)
by: Lu, Yunfan, et al.
Published: (2024)
Dexterous Manipulation Based on Prior Dexterous Grasp Pose Knowledge
by: Yan, Hengxu, et al.
Published: (2024)
by: Yan, Hengxu, et al.
Published: (2024)
ActiveVLA: Injecting Active Perception into Vision-Language-Action Models for Precise 3D Robotic Manipulation
by: Liu, Zhenyang, et al.
Published: (2026)
by: Liu, Zhenyang, et al.
Published: (2026)
ForceMimic: Force-Centric Imitation Learning with Force-Motion Capture System for Contact-Rich Manipulation
by: Liu, Wenhai, et al.
Published: (2024)
by: Liu, Wenhai, et al.
Published: (2024)
History-Aware Visuomotor Policy Learning via Point Tracking
by: Chen, Jingjing, et al.
Published: (2025)
by: Chen, Jingjing, et al.
Published: (2025)
DiPGrasp: Parallel Local Searching for Efficient Differentiable Grasp Planning
by: Xu, Wenqiang, et al.
Published: (2024)
by: Xu, Wenqiang, et al.
Published: (2024)
UniAff: A Unified Representation of Affordances for Tool Usage and Articulation with Vision-Language Models
by: Yu, Qiaojun, et al.
Published: (2024)
by: Yu, Qiaojun, et al.
Published: (2024)
DiffGen: Robot Demonstration Generation via Differentiable Physics Simulation, Differentiable Rendering, and Vision-Language Model
by: Jin, Yang, et al.
Published: (2024)
by: Jin, Yang, et al.
Published: (2024)
CLEVER: Stream-based Active Learning for Robust Semantic Perception from Human Instructions
by: Lee, Jongseok, et al.
Published: (2025)
by: Lee, Jongseok, et al.
Published: (2025)
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction
by: Xiong, Kai, et al.
Published: (2026)
by: Xiong, Kai, et al.
Published: (2026)
MS-MANO: Enabling Hand Pose Tracking with Biomechanical Constraints
by: Xie, Pengfei, et al.
Published: (2024)
by: Xie, Pengfei, et al.
Published: (2024)
CAGE: Causal Attention Enables Data-Efficient Generalizable Robotic Manipulation
by: Xia, Shangning, et al.
Published: (2024)
by: Xia, Shangning, et al.
Published: (2024)
Towards Effective Utilization of Mixed-Quality Demonstrations in Robotic Manipulation via Segment-Level Selection and Optimization
by: Chen, Jingjing, et al.
Published: (2024)
by: Chen, Jingjing, et al.
Published: (2024)
ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
by: Song, Wenxuan, et al.
Published: (2025)
by: Song, Wenxuan, et al.
Published: (2025)
FoundationGrasp: Generalizable Task-Oriented Grasping with Foundation Models
by: Tang, Chao, et al.
Published: (2024)
by: Tang, Chao, et al.
Published: (2024)
FBI: Learning Dexterous In-hand Manipulation with Dynamic Visuotactile Shortcut Policy
by: Chen, Yijin, et al.
Published: (2025)
by: Chen, Yijin, et al.
Published: (2025)
DriveSOTIF: Advancing Perception SOTIF Through Multimodal Large Language Models
by: Huang, Shucheng, et al.
Published: (2025)
by: Huang, Shucheng, et al.
Published: (2025)
AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation
by: Zhang, Yizheng, et al.
Published: (2025)
by: Zhang, Yizheng, et al.
Published: (2025)
Knowledge-Driven Imitation Learning: Enabling Generalization Across Diverse Conditions
by: Miao, Zhuochen, et al.
Published: (2025)
by: Miao, Zhuochen, et al.
Published: (2025)
AnyDexGrasp: General Dexterous Grasping for Different Hands with Human-level Learning Efficiency
by: Fang, Hao-Shu, et al.
Published: (2025)
by: Fang, Hao-Shu, et al.
Published: (2025)
FALCON: Actively Decoupled Visuomotor Policies for Loco-Manipulation with Foundation-Model-Based Coordination
by: He, Chengyang, et al.
Published: (2025)
by: He, Chengyang, et al.
Published: (2025)
Words to Wheels: Vision-Based Autonomous Driving Understanding Human Language Instructions Using Foundation Models
by: Ryu, Chanhoe, et al.
Published: (2024)
by: Ryu, Chanhoe, et al.
Published: (2024)
Similar Items
-
Tru-POMDP: Task Planning Under Uncertainty via Tree of Hypotheses and Open-Ended POMDPs
by: Tang, Wenjing, et al.
Published: (2025) -
Mimic Intent, Not Just Trajectories
by: Huang, Renming, et al.
Published: (2026) -
UniPlan: Vision-Language Task Planning for Mobile Manipulation with Unified PDDL Formulation
by: Ye, Haoming, et al.
Published: (2026) -
UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning
by: Ye, Haoming, et al.
Published: (2025) -
Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks
by: Liu, Zhihong, et al.
Published: (2026)