ForeAct: Steering Your VLA with Efficient Visual Foresight Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Zhuoyang, Yang, Shang, Hu, Qinghao, Huang, Luke J., Hou, James, Sun, Yufei, Lu, Yao, Han, Song |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference
by: Tang, Jiaming, et al.
Published: (2025)
by: Tang, Jiaming, et al.
Published: (2025)
ForeDiffusion: Foresight-Conditioned Diffusion Policy via Future View Construction for Robot Manipulation
by: Xie, Weize, et al.
Published: (2026)
by: Xie, Weize, et al.
Published: (2026)
Unified Noise Steering for Efficient Human-Guided VLA Adaptation
by: Lu, Junjie, et al.
Published: (2026)
by: Lu, Junjie, et al.
Published: (2026)
HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
by: Lin, Minghui, et al.
Published: (2025)
by: Lin, Minghui, et al.
Published: (2025)
SteerVLA: Steering Vision-Language-Action Models in Long-Tail Driving Scenarios
by: Gao, Tian, et al.
Published: (2026)
by: Gao, Tian, et al.
Published: (2026)
Make Your VLA More Robust Without More Data By Interleaving Motion Planning
by: Choe, Dan BW, et al.
Published: (2026)
by: Choe, Dan BW, et al.
Published: (2026)
Towards Unified World Models for Visual Navigation via Memory-Augmented Planning and Foresight
by: Dong, Yifei, et al.
Published: (2025)
by: Dong, Yifei, et al.
Published: (2025)
StreamVLA: Breaking the Reason-Act Cycle via Completion-State Gating
by: Chen, Tongqing, et al.
Published: (2026)
by: Chen, Tongqing, et al.
Published: (2026)
Stable Asynchrony: Variance-Controlled Off-Policy RL for LLMs
by: Huang, Luke J., et al.
Published: (2026)
by: Huang, Luke J., et al.
Published: (2026)
From Foresight to Forethought: VLM-In-the-Loop Policy Steering via Latent Alignment
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
ManualVLA: A Unified VLA Model for Chain-of-Thought Manual Generation and Robotic Manipulation
by: Gu, Chenyang, et al.
Published: (2025)
by: Gu, Chenyang, et al.
Published: (2025)
CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models
by: Zhao, Qingqing, et al.
Published: (2025)
by: Zhao, Qingqing, et al.
Published: (2025)
Afford-VLA: Action-Aligned Visual Planning via Internalized Affordance
by: Wang, Runze, et al.
Published: (2026)
by: Wang, Runze, et al.
Published: (2026)
Don't Blind Your VLA: Aligning Visual Representations for OOD Generalization
by: Kachaev, Nikita, et al.
Published: (2025)
by: Kachaev, Nikita, et al.
Published: (2025)
ForesightNav: Learning Scene Imagination for Efficient Exploration
by: Shah, Hardik, et al.
Published: (2025)
by: Shah, Hardik, et al.
Published: (2025)
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions
by: Bu, Qingwen, et al.
Published: (2025)
by: Bu, Qingwen, et al.
Published: (2025)
CLaD: Planning with Grounded Foresight via Cross-Modal Latent Dynamics
by: Jeong, Andrew, et al.
Published: (2026)
by: Jeong, Andrew, et al.
Published: (2026)
VLA-GSE: Boosting Parameter-Efficient Fine-Tuning in VLA with Generalized and Specialized Experts
by: Jiang, Yuhua, et al.
Published: (2026)
by: Jiang, Yuhua, et al.
Published: (2026)
Learning to Feel the Future: DreamTacVLA for Contact-Rich Manipulation
by: Ye, Guo, et al.
Published: (2025)
by: Ye, Guo, et al.
Published: (2025)
SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
by: Qu, Delin, et al.
Published: (2025)
by: Qu, Delin, et al.
Published: (2025)
BEV-ODOM2: Enhanced BEV-based Monocular Visual Odometry with PV-BEV Fusion and Dense Flow Supervision for Ground Robots
by: Wei, Yufei, et al.
Published: (2025)
by: Wei, Yufei, et al.
Published: (2025)
When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering
by: Yuan, Jessie, et al.
Published: (2026)
by: Yuan, Jessie, et al.
Published: (2026)
Generative Visual Foresight Meets Task-Agnostic Pose Estimation in Robotic Table-Top Manipulation
by: Zhang, Chuye, et al.
Published: (2025)
by: Zhang, Chuye, et al.
Published: (2025)
Efficient Vision-Language-Action Models for Embodied Manipulation: A Systematic Survey
by: Guan, Weifan, et al.
Published: (2025)
by: Guan, Weifan, et al.
Published: (2025)
VP-VLA: Visual Prompting as an Interface for Vision-Language-Action Models
by: Wang, Zixuan, et al.
Published: (2026)
by: Wang, Zixuan, et al.
Published: (2026)
BEV-ODOM: Reducing Scale Drift in Monocular Visual Odometry with BEV Representation
by: Wei, Yufei, et al.
Published: (2024)
by: Wei, Yufei, et al.
Published: (2024)
PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation
by: Li, Yutai, et al.
Published: (2026)
by: Li, Yutai, et al.
Published: (2026)
Towards Deploying VLA without Fine-Tuning: Plug-and-Play Inference-Time VLA Policy Steering via Embodied Evolutionary Diffusion
by: Li, Zhuo, et al.
Published: (2025)
by: Li, Zhuo, et al.
Published: (2025)
CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding
by: Song, Wenxuan, et al.
Published: (2025)
by: Song, Wenxuan, et al.
Published: (2025)
GeoAware-VLA: Implicit Geometry Aware Vision-Language-Action Model
by: Abouzeid, Ali, et al.
Published: (2025)
by: Abouzeid, Ali, et al.
Published: (2025)
Sample-Efficient Learning with Online Expert Correction for Autonomous Catheter Steering in Endovascular Bifurcation Navigation
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid
by: Xu, Xinyu, et al.
Published: (2024)
by: Xu, Xinyu, et al.
Published: (2024)
Long-Horizon Manipulation via Trace-Conditioned VLA Planning
by: Liu, Isabella, et al.
Published: (2026)
by: Liu, Isabella, et al.
Published: (2026)
Spatial-Aware VLA Pretraining through Visual-Physical Alignment from Human Videos
by: Feng, Yicheng, et al.
Published: (2025)
by: Feng, Yicheng, et al.
Published: (2025)
AnchorVLA: Anchored Diffusion for Efficient End-to-End Mobile Manipulation
by: Lim, Jia Syuen, et al.
Published: (2026)
by: Lim, Jia Syuen, et al.
Published: (2026)
FoAM: Foresight-Augmented Multi-Task Imitation Policy for Robotic Manipulation
by: Liu, Litao, et al.
Published: (2024)
by: Liu, Litao, et al.
Published: (2024)
ProAct: A Benchmark and Multimodal Framework for Structure-Aware Proactive Response
by: Zhu, Xiaomeng, et al.
Published: (2026)
by: Zhu, Xiaomeng, et al.
Published: (2026)
DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
by: Yue, Yang, et al.
Published: (2024)
by: Yue, Yang, et al.
Published: (2024)
AeroPlace-Flow: Language-Grounded Object Placement for Aerial Manipulators via Visual Foresight and Object Flow
by: Mishra, Sarthak, et al.
Published: (2026)
by: Mishra, Sarthak, et al.
Published: (2026)
VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
by: Wang, Yihao, et al.
Published: (2025)
by: Wang, Yihao, et al.
Published: (2025)
Similar Items
-
VLASH: Real-Time VLAs via Future-State-Aware Asynchronous Inference
by: Tang, Jiaming, et al.
Published: (2025) -
ForeDiffusion: Foresight-Conditioned Diffusion Policy via Future View Construction for Robot Manipulation
by: Xie, Weize, et al.
Published: (2026) -
Unified Noise Steering for Efficient Human-Guided VLA Adaptation
by: Lu, Junjie, et al.
Published: (2026) -
HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models
by: Lin, Minghui, et al.
Published: (2025) -
SteerVLA: Steering Vision-Language-Action Models in Long-Tail Driving Scenarios
by: Gao, Tian, et al.
Published: (2026)