Augmented Reality for RObots (ARRO): Pointing Visuomotor Policies Towards Visual Robustness
Fuente:
arXiv
Saved in:
| Main Authors: | Mirjalili, Reihaneh, Jülg, Tobias, Walter, Florian, Burgard, Wolfram |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025)
by: Jülg, Tobias, et al.
Published: (2025)
VLM-Vac: Enhancing Smart Vacuums through VLM Knowledge Distillation and Language-Guided Experience Replay
by: Mirjalili, Reihaneh, et al.
Published: (2024)
by: Mirjalili, Reihaneh, et al.
Published: (2024)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
by: Krack, Pierre, et al.
Published: (2026)
by: Krack, Pierre, et al.
Published: (2026)
Lan-grasp: Using Large Language Models for Semantic Object Grasping and Placement
by: Mirjalili, Reihaneh, et al.
Published: (2023)
by: Mirjalili, Reihaneh, et al.
Published: (2023)
LLM-Pack: Intuitive Grocery Handling for Logistics Applications
by: Blei, Yannik, et al.
Published: (2025)
by: Blei, Yannik, et al.
Published: (2025)
VLAgents: A Policy Server for Efficient VLA Inference
by: Jülg, Tobias, et al.
Published: (2026)
by: Jülg, Tobias, et al.
Published: (2026)
Leveraging Foundation Models for Enhancing Robot Perception and Action
by: Mirjalili, Reihaneh
Published: (2025)
by: Mirjalili, Reihaneh
Published: (2025)
FlowTouch: View-Invariant Visuo-Tactile Prediction
by: Bien, Seongjin, et al.
Published: (2026)
by: Bien, Seongjin, et al.
Published: (2026)
Robot Control Stack: A Lean Ecosystem for Robot Learning at Scale
by: Jülg, Tobias, et al.
Published: (2025)
by: Jülg, Tobias, et al.
Published: (2025)
uPLAM: Robust Panoptic Localization and Mapping Leveraging Perception Uncertainties
by: Sirohi, Kshitij, et al.
Published: (2024)
by: Sirohi, Kshitij, et al.
Published: (2024)
MEMROC: Multi-Eye to Mobile RObot Calibration
by: Allegro, Davide, et al.
Published: (2024)
by: Allegro, Davide, et al.
Published: (2024)
Feature Extractor or Decision Maker: Rethinking the Role of Visual Encoders in Visuomotor Policies
by: Wang, Ruiyu, et al.
Published: (2024)
by: Wang, Ruiyu, et al.
Published: (2024)
History-Aware Visuomotor Policy Learning via Point Tracking
by: Chen, Jingjing, et al.
Published: (2025)
by: Chen, Jingjing, et al.
Published: (2025)
Bayesian Optimization for Sample-Efficient Policy Improvement in Robotic Manipulation
by: Röfer, Adrian, et al.
Published: (2024)
by: Röfer, Adrian, et al.
Published: (2024)
LiDAR Registration with Visual Foundation Models
by: Vödisch, Niclas, et al.
Published: (2025)
by: Vödisch, Niclas, et al.
Published: (2025)
Latent Policy Barrier: Learning Robust Visuomotor Policies by Staying In-Distribution
by: Sun, Zhanyi, et al.
Published: (2025)
by: Sun, Zhanyi, et al.
Published: (2025)
ConPoSe: LLM-Guided Contact Point Selection for Scalable Cooperative Object Pushing
by: Steinkrüger, Noah, et al.
Published: (2025)
by: Steinkrüger, Noah, et al.
Published: (2025)
Trajectory-Consistent Flow Matching for Robust Visuomotor Policy Learning
by: Ahmed, Riad, et al.
Published: (2026)
by: Ahmed, Riad, et al.
Published: (2026)
PALM: Enhanced Generalizability for Local Visuomotor Policies via Perception Alignment
by: Wang, Ruiyu, et al.
Published: (2026)
by: Wang, Ruiyu, et al.
Published: (2026)
Fast and Robust Visuomotor Riemannian Flow Matching Policy
by: Ding, Haoran, et al.
Published: (2024)
by: Ding, Haoran, et al.
Published: (2024)
CDP: Towards Robust Autoregressive Visuomotor Policy Learning via Causal Diffusion
by: Ma, Jiahua, et al.
Published: (2025)
by: Ma, Jiahua, et al.
Published: (2025)
A Study on Enhancing the Generalization Ability of Visuomotor Policies via Data Augmentation
by: Wang, Hanwen
Published: (2025)
by: Wang, Hanwen
Published: (2025)
Imagination at Inference: Synthesizing In-Hand Views for Robust Visuomotor Policy Inference
by: Ding, Haoran, et al.
Published: (2025)
by: Ding, Haoran, et al.
Published: (2025)
Agent-Agnostic Centralized Training for Decentralized Multi-Agent Cooperative Driving
by: Yan, Shengchao, et al.
Published: (2024)
by: Yan, Shengchao, et al.
Published: (2024)
Automatic Target-Less Camera-LiDAR Calibration From Motion and Deep Point Correspondences
by: Petek, Kürsat, et al.
Published: (2024)
by: Petek, Kürsat, et al.
Published: (2024)
MemoAct: Atkinson-Shiffrin-Inspired Memory-Augmented Visuomotor Policy for Robotic Manipulation
by: Tan, Liufan, et al.
Published: (2026)
by: Tan, Liufan, et al.
Published: (2026)
Parse-Augment-Distill: Learning Generalizable Bimanual Visuomotor Policies from Single Human Video
by: Tziafas, Georgios, et al.
Published: (2025)
by: Tziafas, Georgios, et al.
Published: (2025)
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation
by: Zhao, Yinuo, et al.
Published: (2024)
by: Zhao, Yinuo, et al.
Published: (2024)
MetricNet: Recovering Metric Scale in Generative Navigation Policies
by: Nayak, Abhijeet, et al.
Published: (2025)
by: Nayak, Abhijeet, et al.
Published: (2025)
Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
by: Chi, Cheng, et al.
Published: (2023)
by: Chi, Cheng, et al.
Published: (2023)
Enabling Dynamic Tracking in Vision-Language-Action Models via Time-Discrete and Time-Continuous Velocity Feedforward
by: Hechtl, Johannes, et al.
Published: (2026)
by: Hechtl, Johannes, et al.
Published: (2026)
AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies
by: Dai, Yinpei, et al.
Published: (2025)
by: Dai, Yinpei, et al.
Published: (2025)
Falcon: Fast Visuomotor Policies via Partial Denoising
by: Chen, Haojun, et al.
Published: (2025)
by: Chen, Haojun, et al.
Published: (2025)
Spatial-Temporal Aware Visuomotor Diffusion Policy Learning
by: Liu, Zhenyang, et al.
Published: (2025)
by: Liu, Zhenyang, et al.
Published: (2025)
Learning Continuous Control with Geometric Regularity from Robot Intrinsic Symmetry
by: Yan, Shengchao, et al.
Published: (2023)
by: Yan, Shengchao, et al.
Published: (2023)
Collaborative Dynamic 3D Scene Graphs for Automated Driving
by: Greve, Elias, et al.
Published: (2023)
by: Greve, Elias, et al.
Published: (2023)
DIPOLE: Fusing Vision and Geometry for Robust Visuomotor Generalization
by: Tang, Yikai, et al.
Published: (2025)
by: Tang, Yikai, et al.
Published: (2025)
Consistency Policy: Accelerated Visuomotor Policies via Consistency Distillation
by: Prasad, Aaditya, et al.
Published: (2024)
by: Prasad, Aaditya, et al.
Published: (2024)
BYE: Build Your Encoder with One Sequence of Exploration Data for Long-Term Dynamic Scene Understanding
by: Huang, Chenguang, et al.
Published: (2024)
by: Huang, Chenguang, et al.
Published: (2024)
Normalizing Flows are Capable Models for Bi-manual Visuomotor Policy
by: Li, Jialong, et al.
Published: (2025)
by: Li, Jialong, et al.
Published: (2025)
Similar Items
-
Refined Policy Distillation: From VLA Generalists to RL Experts
by: Jülg, Tobias, et al.
Published: (2025) -
VLM-Vac: Enhancing Smart Vacuums through VLM Knowledge Distillation and Language-Guided Experience Replay
by: Mirjalili, Reihaneh, et al.
Published: (2024) -
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
by: Krack, Pierre, et al.
Published: (2026) -
Lan-grasp: Using Large Language Models for Semantic Object Grasping and Placement
by: Mirjalili, Reihaneh, et al.
Published: (2023) -
LLM-Pack: Intuitive Grocery Handling for Logistics Applications
by: Blei, Yannik, et al.
Published: (2025)