Contrastive Imitation Learning for Language-guided Multi-Task Robotic Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Teli, Zhou, Jiaming, Wang, Zifan, Qiu, Ronghe, Liang, Junwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mitigating the Human-Robot Domain Discrepancy in Visual Pre-training for Robotic Manipulation
von: Zhou, Jiaming, et al.
Veröffentlicht: (2024)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2024)
GLOVER++: Unleashing the Potential of Affordance Learning from Human Behaviors for Robotic Manipulation
von: Ma, Teli, et al.
Veröffentlicht: (2025)
von: Ma, Teli, et al.
Veröffentlicht: (2025)
Exploring the Limits of Vision-Language-Action Manipulations in Cross-task Generalization
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
von: Ma, Teli, et al.
Veröffentlicht: (2024)
von: Ma, Teli, et al.
Veröffentlicht: (2024)
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)
GenH2R: Learning Generalizable Human-to-Robot Handover via Scalable Simulation, Demonstration, and Imitation
von: Wang, Zifan, et al.
Veröffentlicht: (2024)
von: Wang, Zifan, et al.
Veröffentlicht: (2024)
Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with Imagination
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
von: Barcellona, Leonardo, et al.
Veröffentlicht: (2024)
Robot See Robot Do: Imitating Articulated Object Manipulation with Monocular 4D Reconstruction
von: Kerr, Justin, et al.
Veröffentlicht: (2024)
von: Kerr, Justin, et al.
Veröffentlicht: (2024)
Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
von: Bai, Yongjie, et al.
Veröffentlicht: (2025)
Two by Two: Learning Multi-Task Pairwise Objects Assembly for Generalizable Robot Manipulation
von: Qi, Yu, et al.
Veröffentlicht: (2025)
von: Qi, Yu, et al.
Veröffentlicht: (2025)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
von: Zheng, Junwei, et al.
Veröffentlicht: (2023)
Robust Instant Policy: Leveraging Student's t-Regression Model for Robust In-context Imitation Learning of Robot Manipulation
von: Oh, Hanbit, et al.
Veröffentlicht: (2025)
von: Oh, Hanbit, et al.
Veröffentlicht: (2025)
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
von: Yang, Dejie, et al.
Veröffentlicht: (2025)
Learning Sidewalk Autopilot from Multi-Scale Imitation with Corrective Behavior Expansion
von: He, Honglin, et al.
Veröffentlicht: (2026)
von: He, Honglin, et al.
Veröffentlicht: (2026)
OmniVLA: Physically-Grounded Multimodal VLA with Unified Multi-Sensor Perception for Robotic Manipulation
von: Guo, Heyu, et al.
Veröffentlicht: (2025)
von: Guo, Heyu, et al.
Veröffentlicht: (2025)
Open-vocabulary Mobile Manipulation in Unseen Dynamic Environments with 3D Semantic Maps
von: Qiu, Dicong, et al.
Veröffentlicht: (2024)
von: Qiu, Dicong, et al.
Veröffentlicht: (2024)
LOTUS: Continual Imitation Learning for Robot Manipulation Through Unsupervised Skill Discovery
von: Wan, Weikang, et al.
Veröffentlicht: (2023)
von: Wan, Weikang, et al.
Veröffentlicht: (2023)
VILP: Imitation Learning with Latent Video Planning
von: Xu, Zhengtong, et al.
Veröffentlicht: (2025)
von: Xu, Zhengtong, et al.
Veröffentlicht: (2025)
Towards Generalizable Vision-Language Robotic Manipulation: A Benchmark and LLM-guided 3D Policy
von: Garcia, Ricardo, et al.
Veröffentlicht: (2024)
von: Garcia, Ricardo, et al.
Veröffentlicht: (2024)
FUNCTO: Function-Centric One-Shot Imitation Learning for Tool Manipulation
von: Tang, Chao, et al.
Veröffentlicht: (2025)
von: Tang, Chao, et al.
Veröffentlicht: (2025)
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
von: Huang, Haifeng, et al.
Veröffentlicht: (2025)
Robotic Manipulation by Imitating Generated Videos Without Physical Demonstrations
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
von: Patel, Shivansh, et al.
Veröffentlicht: (2025)
Multi-Camera View Scaling for Data-Efficient Robot Imitation Learning
von: Xie, Yichen, et al.
Veröffentlicht: (2026)
von: Xie, Yichen, et al.
Veröffentlicht: (2026)
Multi-Task Learning for Robot Perception with Imbalanced Data
von: Erkent, Ozgur
Veröffentlicht: (2026)
von: Erkent, Ozgur
Veröffentlicht: (2026)
Play to the Score: Stage-Guided Dynamic Multi-Sensory Fusion for Robotic Manipulation
von: Feng, Ruoxuan, et al.
Veröffentlicht: (2024)
von: Feng, Ruoxuan, et al.
Veröffentlicht: (2024)
Language-Guided Grasp Detection with Coarse-to-Fine Learning for Robotic Manipulation
von: Jiang, Zebin, et al.
Veröffentlicht: (2025)
von: Jiang, Zebin, et al.
Veröffentlicht: (2025)
Benchmarking Vision, Language, & Action Models on Robotic Learning Tasks
von: Guruprasad, Pranav, et al.
Veröffentlicht: (2024)
von: Guruprasad, Pranav, et al.
Veröffentlicht: (2024)
Towards Generalizable Robotic Manipulation in Dynamic Environments
von: Fang, Heng, et al.
Veröffentlicht: (2026)
von: Fang, Heng, et al.
Veröffentlicht: (2026)
Improving Generalization of Language-Conditioned Robot Manipulation
von: Cui, Chenglin, et al.
Veröffentlicht: (2025)
von: Cui, Chenglin, et al.
Veröffentlicht: (2025)
ST-$π$: Structured SpatioTemporal VLA for Robotic Manipulation
von: Ma, Chuanhao, et al.
Veröffentlicht: (2026)
von: Ma, Chuanhao, et al.
Veröffentlicht: (2026)
Breathless: An 8-hour Performance Contrasting Human and Robot Expressiveness
von: Cuan, Catie, et al.
Veröffentlicht: (2024)
von: Cuan, Catie, et al.
Veröffentlicht: (2024)
Prioritized Semantic Learning for Zero-shot Instance Navigation
von: Sun, Xinyu, et al.
Veröffentlicht: (2024)
von: Sun, Xinyu, et al.
Veröffentlicht: (2024)
TapSampling: Inference-Time Sampling with a Task-Progress-Understanding Verifier for Robotic Manipulation
von: Zhao, Sizhe, et al.
Veröffentlicht: (2026)
von: Zhao, Sizhe, et al.
Veröffentlicht: (2026)
Hierarchical Diffusion Policy for Kinematics-Aware Multi-Task Robotic Manipulation
von: Ma, Xiao, et al.
Veröffentlicht: (2024)
von: Ma, Xiao, et al.
Veröffentlicht: (2024)
SaPaVe: Towards Active Perception and Manipulation in Vision-Language-Action Models for Robotics
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
von: Liu, Mengzhen, et al.
Veröffentlicht: (2026)
An Examination of the Compositionality of Large Generative Vision-Language Models
von: Ma, Teli, et al.
Veröffentlicht: (2023)
von: Ma, Teli, et al.
Veröffentlicht: (2023)
Think Proprioceptively: Embodied Visual Reasoning for VLA Manipulation
von: Wang, Fangyuan, et al.
Veröffentlicht: (2026)
von: Wang, Fangyuan, et al.
Veröffentlicht: (2026)
BitVLA: 1-bit Vision-Language-Action Models for Robotics Manipulation
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
von: Wang, Hongyu, et al.
Veröffentlicht: (2025)
ManiVID-3D: Generalizable View-Invariant Reinforcement Learning for Robotic Manipulation via Disentangled 3D Representations
von: Li, Zheng, et al.
Veröffentlicht: (2025)
von: Li, Zheng, et al.
Veröffentlicht: (2025)
RoboPearls: Editable Video Simulation for Robot Manipulation
von: Tang, Tao, et al.
Veröffentlicht: (2025)
von: Tang, Tao, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mitigating the Human-Robot Domain Discrepancy in Visual Pre-training for Robotic Manipulation
von: Zhou, Jiaming, et al.
Veröffentlicht: (2024) -
GLOVER++: Unleashing the Potential of Affordance Learning from Human Behaviors for Robotic Manipulation
von: Ma, Teli, et al.
Veröffentlicht: (2025) -
Exploring the Limits of Vision-Language-Action Manipulations in Cross-task Generalization
von: Zhou, Jiaming, et al.
Veröffentlicht: (2025) -
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
von: Ma, Teli, et al.
Veröffentlicht: (2024) -
SD-OVON: A Semantics-aware Dataset and Benchmark Generation Pipeline for Open-Vocabulary Object Navigation in Dynamic Scenes
von: Qiu, Dicong, et al.
Veröffentlicht: (2025)