IMRL: Integrating Visual, Physical, Temporal, and Geometric Representations for Enhanced Food Acquisition
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Rui, Mahammad, Zahiruddin, Bhaskar, Amisha, Tokekar, Pratap |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sketch-to-Skill: Bootstrapping Robot Learning with Human Drawn Trajectory Sketches
by: Yu, Peihong, et al.
Published: (2025)
by: Yu, Peihong, et al.
Published: (2025)
Adaptive Visual Imitation Learning for Robotic Assisted Feeding Across Varied Bowl Configurations and Food Types
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
PLANRL: A Motion Planning and Imitation Learning Framework to Bootstrap Reinforcement Learning
by: Bhaskar, Amisha, et al.
Published: (2024)
by: Bhaskar, Amisha, et al.
Published: (2024)
LAVA: Long-horizon Visual Action based Food Acquisition
by: Bhaskar, Amisha, et al.
Published: (2024)
by: Bhaskar, Amisha, et al.
Published: (2024)
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
PRISM: Performer RS-IMLE for Single-pass Multisensory Imitation Learning
by: Bhaskar, Amisha, et al.
Published: (2026)
by: Bhaskar, Amisha, et al.
Published: (2026)
CAML: Collaborative Auxiliary Modality Learning for Multi-Agent Systems
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
When to Localize? A Risk-Constrained Reinforcement Learning Approach
by: Shek, Chak Lam, et al.
Published: (2024)
by: Shek, Chak Lam, et al.
Published: (2024)
MMCD: Multi-Modal Collaborative Decision-Making for Connected Autonomy with Knowledge Distillation
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
PEnGUiN: Partially Equivariant Graph NeUral Networks for Sample Efficient MARL
by: McClellan, Joshua, et al.
Published: (2025)
by: McClellan, Joshua, et al.
Published: (2025)
AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation
by: Singh, Anukriti, et al.
Published: (2025)
by: Singh, Anukriti, et al.
Published: (2025)
AG-CVG: Coverage Planning with a Mobile Recharging UGV and an Energy-Constrained UAV
by: Karapetyan, Nare, et al.
Published: (2023)
by: Karapetyan, Nare, et al.
Published: (2023)
REBEL: Reward Regularization-Based Approach for Robotic Reinforcement Learning from Human Feedback
by: Chakraborty, Souradip, et al.
Published: (2023)
by: Chakraborty, Souradip, et al.
Published: (2023)
Learning Multi-Robot Coordination through Locality-Based Factorized Multi-Agent Actor-Critic Algorithm
by: Shek, Chak Lam, et al.
Published: (2025)
by: Shek, Chak Lam, et al.
Published: (2025)
Semantic-Geometric-Physical-Driven Robot Manipulation Skill Transfer via Skill Library and Tactile Representation
by: Qi, Mingchao, et al.
Published: (2024)
by: Qi, Mingchao, et al.
Published: (2024)
Data-Driven Distributionally Robust Optimal Control with State-Dependent Noise
by: Liu, Rui, et al.
Published: (2023)
by: Liu, Rui, et al.
Published: (2023)
Beyond Joint Demonstrations: Personalized Expert Guidance for Efficient Multi-Agent Reinforcement Learning
by: Yu, Peihong, et al.
Published: (2024)
by: Yu, Peihong, et al.
Published: (2024)
Active Asymmetric Multi-Agent Multimodal Learning under Uncertainty
by: Liu, Rui, et al.
Published: (2026)
by: Liu, Rui, et al.
Published: (2026)
TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
by: Zheng, Ruijie, et al.
Published: (2024)
by: Zheng, Ruijie, et al.
Published: (2024)
Wonderful Team: Zero-Shot Physical Task Planning with Visual LLMs
by: Wang, Zidan, et al.
Published: (2024)
by: Wang, Zidan, et al.
Published: (2024)
Exploring Temporal Representation in Neural Processes for Multimodal Action Prediction
by: Fedozzi, Marco Gabriele, et al.
Published: (2026)
by: Fedozzi, Marco Gabriele, et al.
Published: (2026)
Real-Time Verification of Embodied Reasoning for Generative Skill Acquisition
by: Yue, Bo, et al.
Published: (2025)
by: Yue, Bo, et al.
Published: (2025)
Visual Sculpting: Visually-Aligned Planning Representations for Long-Horizon Robot Clay Sculpting
by: Schaldenbrand, Peter, et al.
Published: (2026)
by: Schaldenbrand, Peter, et al.
Published: (2026)
Temporal Action Representation Learning for Tactical Resource Control and Subsequent Maneuver Generation
by: Jung, Hoseong, et al.
Published: (2026)
by: Jung, Hoseong, et al.
Published: (2026)
VLMPlanner: Integrating Visual Language Models with Motion Planning
by: Tang, Zhipeng, et al.
Published: (2025)
by: Tang, Zhipeng, et al.
Published: (2025)
M2R2: MultiModal Robotic Representation for Temporal Action Segmentation
by: Sliwowski, Daniel, et al.
Published: (2025)
by: Sliwowski, Daniel, et al.
Published: (2025)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
Learning Structural Latent Points for Efficient Visual Representations in Robotic Manipulation
by: Jiang, Yicheng, et al.
Published: (2026)
by: Jiang, Yicheng, et al.
Published: (2026)
Visual Forecasting as a Mid-level Representation for Avoidance
by: Yang, Hsuan-Kung, et al.
Published: (2023)
by: Yang, Hsuan-Kung, et al.
Published: (2023)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
by: Shek, Chak Lam, et al.
Published: (2025)
by: Shek, Chak Lam, et al.
Published: (2025)
UniJEPA: Enhancing Robot Policy via Unified Continuous and Discrete Representation Learning
by: Zhang, Jianke, et al.
Published: (2025)
by: Zhang, Jianke, et al.
Published: (2025)
Reconciling Spatial and Temporal Abstractions for Goal Representation
by: Zadem, Mehdi, et al.
Published: (2024)
by: Zadem, Mehdi, et al.
Published: (2024)
Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis
by: Liu, Rui, et al.
Published: (2024)
by: Liu, Rui, et al.
Published: (2024)
ATLASv2: LLM-Guided Adaptive Landmark Acquisition and Navigation on the Edge
by: Walczak, Mikolaj, et al.
Published: (2025)
by: Walczak, Mikolaj, et al.
Published: (2025)
Action with Visual Primitives
by: Guo, Weilong, et al.
Published: (2026)
by: Guo, Weilong, et al.
Published: (2026)
Visualizing Latent Phase Structures in Locomotion Policies: A Multi-Environment Study with Temporal Feature Extension
by: Yasui, Daisuke, et al.
Published: (2026)
by: Yasui, Daisuke, et al.
Published: (2026)
Screw Geometry Meets Bandits: Incremental Acquisition of Demonstrations to Generate Manipulation Plans
by: Das, Dibyendu, et al.
Published: (2024)
by: Das, Dibyendu, et al.
Published: (2024)
Exploring Spatial Representation to Enhance LLM Reasoning in Aerial Vision-Language Navigation
by: Gao, Yunpeng, et al.
Published: (2024)
by: Gao, Yunpeng, et al.
Published: (2024)
Shape Completion and Real-Time Visualization in Robotic Ultrasound Spine Acquisitions
by: Gafencu, Miruna-Alexandra, et al.
Published: (2025)
by: Gafencu, Miruna-Alexandra, et al.
Published: (2025)
DTRT: Enhancing Human Intent Estimation and Role Allocation for Physical Human-Robot Collaboration
by: Liu, Haotian, et al.
Published: (2025)
by: Liu, Haotian, et al.
Published: (2025)
Similar Items
-
Sketch-to-Skill: Bootstrapping Robot Learning with Human Drawn Trajectory Sketches
by: Yu, Peihong, et al.
Published: (2025) -
Adaptive Visual Imitation Learning for Robotic Assisted Feeding Across Varied Bowl Configurations and Food Types
by: Liu, Rui, et al.
Published: (2024) -
PLANRL: A Motion Planning and Imitation Learning Framework to Bootstrap Reinforcement Learning
by: Bhaskar, Amisha, et al.
Published: (2024) -
LAVA: Long-horizon Visual Action based Food Acquisition
by: Bhaskar, Amisha, et al.
Published: (2024) -
VARP: Reinforcement Learning from Vision-Language Model Feedback with Agent Regularized Preferences
by: Singh, Anukriti, et al.
Published: (2025)