Saved in:
| Main Authors: | Rasouli, Amir, Wu, Yangzheng, Li, Zhiyuan, Yang, Rui Heng, Zhao, Xuan, Eret, Charles, Pakdamansavoji, Sajjad |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.21192 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Robotic Manipulation Robustness via NICE Scene Surgery
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
by: Rasouli, Amir, et al.
Published: (2025)
by: Rasouli, Amir, et al.
Published: (2025)
Do World Action Models Generalize Better than VLAs? A Robustness Study
by: Zhang, Zhanguang, et al.
Published: (2026)
by: Zhang, Zhanguang, et al.
Published: (2026)
WALDO: Where Unseen Model-based 6D Pose Estimation Meets Occlusion
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
by: Pakdamansavoji, Sajjad, et al.
Published: (2025)
Box6D : Zero-shot Category-level 6D Pose Estimation of Warehouse Boxes
by: Ma, Yintao, et al.
Published: (2025)
by: Ma, Yintao, et al.
Published: (2025)
How Do VLAs Effectively Inherit from VLMs?
by: Zhang, Chuheng, et al.
Published: (2025)
by: Zhang, Chuheng, et al.
Published: (2025)
CAPE: Context-Aware Diffusion Policy Via Proximal Mode Expansion for Collision Avoidance
by: Yang, Rui Heng, et al.
Published: (2025)
by: Yang, Rui Heng, et al.
Published: (2025)
VLA-0: Building State-of-the-Art VLAs with Zero Modification
by: Goyal, Ankit, et al.
Published: (2025)
by: Goyal, Ankit, et al.
Published: (2025)
Getting SMARTER for Motion Planning in Autonomous Driving Systems
by: Alban, Montgomery, et al.
Published: (2025)
by: Alban, Montgomery, et al.
Published: (2025)
Retrieve-then-Steer: Online Success Memory for Test-Time Adaptation of Generative VLAs
by: Zhao, Jianchao, et al.
Published: (2026)
by: Zhao, Jianchao, et al.
Published: (2026)
Two-Steps Diffusion Policy for Robotic Manipulation via Genetic Denoising
by: Clemente, Mateo, et al.
Published: (2025)
by: Clemente, Mateo, et al.
Published: (2025)
Scaling Sim-to-Real Reinforcement Learning for Robot VLAs with Generative 3D Worlds
by: Choi, Andrew, et al.
Published: (2026)
by: Choi, Andrew, et al.
Published: (2026)
Lost in Fog: Sensor Perturbations Expose Reasoning Fragility in Driving VLAs
by: Priyadershi, Abhinaw, et al.
Published: (2026)
by: Priyadershi, Abhinaw, et al.
Published: (2026)
VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference
by: Tang, Jiaming, et al.
Published: (2025)
by: Tang, Jiaming, et al.
Published: (2025)
WorldGym: World Model as An Environment for Policy Evaluation
by: Quevedo, Julian, et al.
Published: (2025)
by: Quevedo, Julian, et al.
Published: (2025)
Reinforcing VLAs in Task-Agnostic World Models
by: Wang, Yucen, et al.
Published: (2026)
by: Wang, Yucen, et al.
Published: (2026)
Cog-GA: A Large Language Models-based Generative Agent for Vision-Language Navigation in Continuous Environments
by: Li, Zhiyuan, et al.
Published: (2024)
by: Li, Zhiyuan, et al.
Published: (2024)
Robot Task Planning and Situation Handling in Open Worlds
by: Ding, Yan, et al.
Published: (2022)
by: Ding, Yan, et al.
Published: (2022)
VTAM: Video-Tactile-Action Models for Complex Physical Interaction Beyond VLAs
by: Yuan, Haoran, et al.
Published: (2026)
by: Yuan, Haoran, et al.
Published: (2026)
Scenarios Engineering driven Autonomous Transportation in Open-Pit Mines
by: Teng, Siyu, et al.
Published: (2024)
by: Teng, Siyu, et al.
Published: (2024)
OpenNav: Open-World Navigation with Multimodal Large Language Models
by: Yuan, Mingfeng, et al.
Published: (2025)
by: Yuan, Mingfeng, et al.
Published: (2025)
EMMOE: A Comprehensive Benchmark for Embodied Mobile Manipulation in Open Environments
by: Li, Dongping, et al.
Published: (2025)
by: Li, Dongping, et al.
Published: (2025)
Humanoid World Models: Open World Foundation Models for Humanoid Robotics
by: Ali, Muhammad Qasim, et al.
Published: (2025)
by: Ali, Muhammad Qasim, et al.
Published: (2025)
mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs
by: Pai, Jonas, et al.
Published: (2025)
by: Pai, Jonas, et al.
Published: (2025)
Scenario Engineering for Autonomous Transportation: A New Stage in Open-Pit Mines
by: Teng, Siyu, et al.
Published: (2024)
by: Teng, Siyu, et al.
Published: (2024)
URDFormer: A Pipeline for Constructing Articulated Simulation Environments from Real-World Images
by: Chen, Zoey, et al.
Published: (2024)
by: Chen, Zoey, et al.
Published: (2024)
DKPROMPT: Domain Knowledge Prompting Vision-Language Models for Open-World Planning
by: Zhang, Xiaohan, et al.
Published: (2024)
by: Zhang, Xiaohan, et al.
Published: (2024)
10 Open Challenges Steering the Future of Vision-Language-Action Models
by: Poria, Soujanya, et al.
Published: (2025)
by: Poria, Soujanya, et al.
Published: (2025)
Creating and Repairing Robot Programs in Open-World Domains
by: Schlesinger, Claire, et al.
Published: (2024)
by: Schlesinger, Claire, et al.
Published: (2024)
Contextual Safety Reasoning and Grounding for Open-World Robots
by: Ravichandran, Zachary, et al.
Published: (2026)
by: Ravichandran, Zachary, et al.
Published: (2026)
Hybrid Framework for Robotic Manipulation: Integrating Reinforcement Learning and Large Language Models
by: Saad, Md, et al.
Published: (2026)
by: Saad, Md, et al.
Published: (2026)
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making
by: Wang, Yucen, et al.
Published: (2025)
by: Wang, Yucen, et al.
Published: (2025)
RA-DP: Rapid Adaptive Diffusion Policy for Training-Free High-frequency Robotics Replanning
by: Ye, Xi, et al.
Published: (2025)
by: Ye, Xi, et al.
Published: (2025)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
by: Zhou, Zhiyuan, et al.
Published: (2025)
by: Zhou, Zhiyuan, et al.
Published: (2025)
Cooperative-Competitive Team Play of Real-World Craft Robots
by: Zhao, Rui, et al.
Published: (2026)
by: Zhao, Rui, et al.
Published: (2026)
Open-World Drone Active Tracking with Goal-Centered Rewards
by: Sun, Haowei, et al.
Published: (2024)
by: Sun, Haowei, et al.
Published: (2024)
OpenGraph: Open-Vocabulary Hierarchical 3D Graph Representation in Large-Scale Outdoor Environments
by: Deng, Yinan, et al.
Published: (2024)
by: Deng, Yinan, et al.
Published: (2024)
How Secure Are Large Language Models (LLMs) for Navigation in Urban Environments?
by: Wen, Congcong, et al.
Published: (2024)
by: Wen, Congcong, et al.
Published: (2024)
Control Synthesis in Partially Observable Environments for Complex Perception-Related Objectives
by: Xuan, Zetong, et al.
Published: (2025)
by: Xuan, Zetong, et al.
Published: (2025)
HA-VLN 2.0: An Open Benchmark and Leaderboard for Human-Aware Navigation in Discrete and Continuous Environments with Dynamic Multi-Human Interactions
by: Dong, Yifei, et al.
Published: (2025)
by: Dong, Yifei, et al.
Published: (2025)
Similar Items
-
Improving Robotic Manipulation Robustness via NICE Scene Surgery
by: Pakdamansavoji, Sajjad, et al.
Published: (2025) -
Distracted Robot: How Visual Clutter Undermine Robotic Manipulation
by: Rasouli, Amir, et al.
Published: (2025) -
Do World Action Models Generalize Better than VLAs? A Robustness Study
by: Zhang, Zhanguang, et al.
Published: (2026) -
WALDO: Where Unseen Model-based 6D Pose Estimation Meets Occlusion
by: Pakdamansavoji, Sajjad, et al.
Published: (2025) -
Box6D : Zero-shot Category-level 6D Pose Estimation of Warehouse Boxes
by: Ma, Yintao, et al.
Published: (2025)