Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yang, Jonathan, Fu, Chuyuan Kelly, Shah, Dhruv, Sadigh, Dorsa, Xia, Fei, Zhang, Tingnan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Invariance Co-training for Robot Visual Generalization
von: Yang, Jonathan, et al.
Veröffentlicht: (2025)
von: Yang, Jonathan, et al.
Veröffentlicht: (2025)
Data Analogies Enable Efficient Cross-Embodiment Transfer
von: Yang, Jonathan, et al.
Veröffentlicht: (2026)
von: Yang, Jonathan, et al.
Veröffentlicht: (2026)
Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation
von: Yang, Jonathan, et al.
Veröffentlicht: (2024)
von: Yang, Jonathan, et al.
Veröffentlicht: (2024)
CLAW: A Vision-Language-Action Framework for Weight-Aware Robotic Grasping
von: An, Zijian, et al.
Veröffentlicht: (2025)
von: An, Zijian, et al.
Veröffentlicht: (2025)
Grounding Robot Generalization in Training Data via Retrieval-Augmented VLMs
von: Gao, Jensen, et al.
Veröffentlicht: (2026)
von: Gao, Jensen, et al.
Veröffentlicht: (2026)
Automated Robot Recovery from Assumption Violations of High-Level Specifications
von: Meng, Qian, et al.
Veröffentlicht: (2024)
von: Meng, Qian, et al.
Veröffentlicht: (2024)
Context Representation via Action-Free Transformer encoder-decoder for Meta Reinforcement Learning
von: Enayati, Amir M. Soufi, et al.
Veröffentlicht: (2025)
von: Enayati, Amir M. Soufi, et al.
Veröffentlicht: (2025)
SeqVLA: Sequential Task Execution for Long-Horizon Manipulation with Completion-Aware Vision-Language-Action Model
von: Yang, Ran, et al.
Veröffentlicht: (2025)
von: Yang, Ran, et al.
Veröffentlicht: (2025)
CAVER: Curious Audiovisual Exploring Robot
von: Macesanu, Luca, et al.
Veröffentlicht: (2025)
von: Macesanu, Luca, et al.
Veröffentlicht: (2025)
Toward a Better Understanding of Robot Energy Consumption in Agroecological Applications
von: Bras, Alexis, et al.
Veröffentlicht: (2024)
von: Bras, Alexis, et al.
Veröffentlicht: (2024)
Action-Aware Pro-Active Safe Exploration for Mobile Robot Mapping
von: İşleyen, Aykut, et al.
Veröffentlicht: (2025)
von: İşleyen, Aykut, et al.
Veröffentlicht: (2025)
Autonomous Robotic System with Optical Coherence Tomography Guidance for Vascular Anastomosis
von: Haworth, Jesse, et al.
Veröffentlicht: (2024)
von: Haworth, Jesse, et al.
Veröffentlicht: (2024)
Development of Compositionality and Generalization through Interactive Learning of Language and Action of Robots
von: Vijayaraghavan, Prasanna, et al.
Veröffentlicht: (2024)
von: Vijayaraghavan, Prasanna, et al.
Veröffentlicht: (2024)
Action Agent: Agentic Video Generation Meets Flow-Constrained Diffusion
von: Sam, Jeffrin, et al.
Veröffentlicht: (2026)
von: Sam, Jeffrin, et al.
Veröffentlicht: (2026)
Enhancing Robustness in Language-Driven Robotics: A Modular Approach to Failure Reduction
von: Garrabé, Émiland, et al.
Veröffentlicht: (2024)
von: Garrabé, Émiland, et al.
Veröffentlicht: (2024)
Recent Advances of Deep Robotic Affordance Learning: A Reinforcement Learning Perspective
von: Yang, Xintong, et al.
Veröffentlicht: (2023)
von: Yang, Xintong, et al.
Veröffentlicht: (2023)
Teacher Motion Priors: Enhancing Robot Locomotion over Challenging Terrain
von: Jin, Fangcheng, et al.
Veröffentlicht: (2025)
von: Jin, Fangcheng, et al.
Veröffentlicht: (2025)
MaP-AVR: A Meta-Action Planner for Agents Leveraging Vision Language Models and Retrieval-Augmented Generation
von: Guo, Zhenglong, et al.
Veröffentlicht: (2025)
von: Guo, Zhenglong, et al.
Veröffentlicht: (2025)
Dimension-variable Mapless Navigation with Deep Reinforcement Learning
von: Zhang, Wei, et al.
Veröffentlicht: (2020)
von: Zhang, Wei, et al.
Veröffentlicht: (2020)
A Time-dependent Risk-aware distributed Multi-Agent Path Finder based on A*
von: Nordström, S, et al.
Veröffentlicht: (2025)
von: Nordström, S, et al.
Veröffentlicht: (2025)
Self-Supervised Depth Correction of Lidar Measurements from Map Consistency Loss
von: Agishev, Ruslan, et al.
Veröffentlicht: (2023)
von: Agishev, Ruslan, et al.
Veröffentlicht: (2023)
UniLab: A Heterogeneous Architecture for Robot RL Beyond GPU-Dominant Paradigms
von: Jia, Yufei, et al.
Veröffentlicht: (2026)
von: Jia, Yufei, et al.
Veröffentlicht: (2026)
DISCOVERSE: Efficient Robot Simulation in Complex High-Fidelity Environments
von: Jia, Yufei, et al.
Veröffentlicht: (2025)
von: Jia, Yufei, et al.
Veröffentlicht: (2025)
FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR
von: Wu, Junzhe, et al.
Veröffentlicht: (2025)
von: Wu, Junzhe, et al.
Veröffentlicht: (2025)
RLPP: A Residual Method for Zero-Shot Real-World Autonomous Racing on Scaled Platforms
von: Ghignone, Edoardo, et al.
Veröffentlicht: (2025)
von: Ghignone, Edoardo, et al.
Veröffentlicht: (2025)
GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning
von: Jia, Yufei, et al.
Veröffentlicht: (2026)
von: Jia, Yufei, et al.
Veröffentlicht: (2026)
The Impact of Class Uncertainty Propagation in Perception-Based Motion Planning
von: Shah, Jibran Iqbal, et al.
Veröffentlicht: (2026)
von: Shah, Jibran Iqbal, et al.
Veröffentlicht: (2026)
MonoForce: Self-supervised Learning of Physics-informed Model for Predicting Robot-terrain Interaction
von: Agishev, Ruslan, et al.
Veröffentlicht: (2023)
von: Agishev, Ruslan, et al.
Veröffentlicht: (2023)
Two-Steps Diffusion Policy for Robotic Manipulation via Genetic Denoising
von: Clemente, Mateo, et al.
Veröffentlicht: (2025)
von: Clemente, Mateo, et al.
Veröffentlicht: (2025)
Automated Feature Selection for Inverse Reinforcement Learning
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
von: Baimukashev, Daulet, et al.
Veröffentlicht: (2024)
Learning Affordances at Inference-Time for Vision-Language-Action Models
von: Shah, Ameesh, et al.
Veröffentlicht: (2025)
von: Shah, Ameesh, et al.
Veröffentlicht: (2025)
A Fast Online Omnidirectional Quadrupedal Jumping Framework Via Virtual-Model Control and Minimum Jerk Trajectory Generation
von: Yue, Linzhu, et al.
Veröffentlicht: (2024)
von: Yue, Linzhu, et al.
Veröffentlicht: (2024)
Multi Object Tracking for Predictive Collision Avoidance
von: Gebregziabher, Bruk, et al.
Veröffentlicht: (2023)
von: Gebregziabher, Bruk, et al.
Veröffentlicht: (2023)
PyroTrack: Belief-Based Deep Reinforcement Learning Path Planning for Aerial Wildfire Monitoring in Partially Observable Environments
von: Khoshdel, Sahand, et al.
Veröffentlicht: (2024)
von: Khoshdel, Sahand, et al.
Veröffentlicht: (2024)
Heuristic Search for Path Finding with Refuelling
von: Zhao, Shizhe, et al.
Veröffentlicht: (2023)
von: Zhao, Shizhe, et al.
Veröffentlicht: (2023)
Learning Barrier-Certified Polynomial Dynamical Systems for Obstacle Avoidance with Robots
von: Schonger, Martin, et al.
Veröffentlicht: (2024)
von: Schonger, Martin, et al.
Veröffentlicht: (2024)
Learning to Communicate Functional States with Nonverbal Expressions for Improved Human-Robot Collaboration
von: Roy, Liam, et al.
Veröffentlicht: (2024)
von: Roy, Liam, et al.
Veröffentlicht: (2024)
Stinger Robot: A Self-Bracing Robotic Platform for Autonomous Drilling in Confined Underground Environments
von: Liu, H., et al.
Veröffentlicht: (2025)
von: Liu, H., et al.
Veröffentlicht: (2025)
Simulation-based Scenario Generation for Robust Hybrid AI for Autonomy
von: Keno, Hambisa, et al.
Veröffentlicht: (2024)
von: Keno, Hambisa, et al.
Veröffentlicht: (2024)
Open, Reproducible and Trustworthy Robot-Based Experiments with Virtual Labs and Digital-Twin-Based Execution Tracing
von: Alt, Benjamin, et al.
Veröffentlicht: (2025)
von: Alt, Benjamin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Invariance Co-training for Robot Visual Generalization
von: Yang, Jonathan, et al.
Veröffentlicht: (2025) -
Data Analogies Enable Efficient Cross-Embodiment Transfer
von: Yang, Jonathan, et al.
Veröffentlicht: (2026) -
Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation
von: Yang, Jonathan, et al.
Veröffentlicht: (2024) -
CLAW: A Vision-Language-Action Framework for Weight-Aware Robotic Grasping
von: An, Zijian, et al.
Veröffentlicht: (2025) -
Grounding Robot Generalization in Training Data via Retrieval-Augmented VLMs
von: Gao, Jensen, et al.
Veröffentlicht: (2026)