SeqVLA: Sequential Task Execution for Long-Horizon Manipulation with Completion-Aware Vision-Language-Action Model
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Ran, An, Zijian, ZHou, Lifeng, Feng, Yiming |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLAW: A Vision-Language-Action Framework for Weight-Aware Robotic Grasping
by: An, Zijian, et al.
Published: (2025)
by: An, Zijian, et al.
Published: (2025)
MaP-AVR: A Meta-Action Planner for Agents Leveraging Vision Language Models and Retrieval-Augmented Generation
by: Guo, Zhenglong, et al.
Published: (2025)
by: Guo, Zhenglong, et al.
Published: (2025)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
by: Syed, Shahram Najam, et al.
Published: (2025)
by: Syed, Shahram Najam, et al.
Published: (2025)
Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization
by: Yang, Jonathan, et al.
Published: (2025)
by: Yang, Jonathan, et al.
Published: (2025)
Action-Aware Pro-Active Safe Exploration for Mobile Robot Mapping
by: İşleyen, Aykut, et al.
Published: (2025)
by: İşleyen, Aykut, et al.
Published: (2025)
Altered Thoughts, Altered Actions: Probing Chain-of-Thought Vulnerabilities in VLA Robotic Manipulation
by: Trinh, Tuan Duong, et al.
Published: (2026)
by: Trinh, Tuan Duong, et al.
Published: (2026)
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
by: Chen, Kewei, et al.
Published: (2026)
by: Chen, Kewei, et al.
Published: (2026)
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
by: Chen, Kewei, et al.
Published: (2025)
by: Chen, Kewei, et al.
Published: (2025)
Long-VLA: Unleashing Long-Horizon Capability of Vision Language Action Model for Robot Manipulation
by: Fan, Yiguo, et al.
Published: (2025)
by: Fan, Yiguo, et al.
Published: (2025)
Context Representation via Action-Free Transformer encoder-decoder for Meta Reinforcement Learning
by: Enayati, Amir M. Soufi, et al.
Published: (2025)
by: Enayati, Amir M. Soufi, et al.
Published: (2025)
Learning Affordances at Inference-Time for Vision-Language-Action Models
by: Shah, Ameesh, et al.
Published: (2025)
by: Shah, Ameesh, et al.
Published: (2025)
Pushing the Limits of Cross-Embodiment Learning for Manipulation and Navigation
by: Yang, Jonathan, et al.
Published: (2024)
by: Yang, Jonathan, et al.
Published: (2024)
SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models
by: Li, Meng, et al.
Published: (2025)
by: Li, Meng, et al.
Published: (2025)
Enhancing Robustness in Language-Driven Robotics: A Modular Approach to Failure Reduction
by: Garrabé, Émiland, et al.
Published: (2024)
by: Garrabé, Émiland, et al.
Published: (2024)
Dimension-variable Mapless Navigation with Deep Reinforcement Learning
by: Zhang, Wei, et al.
Published: (2020)
by: Zhang, Wei, et al.
Published: (2020)
A Time-dependent Risk-aware distributed Multi-Agent Path Finder based on A*
by: Nordström, S, et al.
Published: (2025)
by: Nordström, S, et al.
Published: (2025)
CAVER: Curious Audiovisual Exploring Robot
by: Macesanu, Luca, et al.
Published: (2025)
by: Macesanu, Luca, et al.
Published: (2025)
Self-Supervised Depth Correction of Lidar Measurements from Map Consistency Loss
by: Agishev, Ruslan, et al.
Published: (2023)
by: Agishev, Ruslan, et al.
Published: (2023)
Toward a Better Understanding of Robot Energy Consumption in Agroecological Applications
by: Bras, Alexis, et al.
Published: (2024)
by: Bras, Alexis, et al.
Published: (2024)
Automated Robot Recovery from Assumption Violations of High-Level Specifications
by: Meng, Qian, et al.
Published: (2024)
by: Meng, Qian, et al.
Published: (2024)
BagelVLA: Enhancing Long-Horizon Manipulation via Interleaved Vision-Language-Action Generation
by: Hu, Yucheng, et al.
Published: (2026)
by: Hu, Yucheng, et al.
Published: (2026)
MVB-Grasp: Minimum-Volume-Box Filtering of Diffusion-based Grasps for Frontal Manipulation
by: Poudel, Bibek, et al.
Published: (2026)
by: Poudel, Bibek, et al.
Published: (2026)
Action Agent: Agentic Video Generation Meets Flow-Constrained Diffusion
by: Sam, Jeffrin, et al.
Published: (2026)
by: Sam, Jeffrin, et al.
Published: (2026)
Two-Steps Diffusion Policy for Robotic Manipulation via Genetic Denoising
by: Clemente, Mateo, et al.
Published: (2025)
by: Clemente, Mateo, et al.
Published: (2025)
Development of Compositionality and Generalization through Interactive Learning of Language and Action of Robots
by: Vijayaraghavan, Prasanna, et al.
Published: (2024)
by: Vijayaraghavan, Prasanna, et al.
Published: (2024)
Autonomous Robotic System with Optical Coherence Tomography Guidance for Vascular Anastomosis
by: Haworth, Jesse, et al.
Published: (2024)
by: Haworth, Jesse, et al.
Published: (2024)
GS-Playground: A High-Throughput Photorealistic Simulator for Vision-Informed Robot Learning
by: Jia, Yufei, et al.
Published: (2026)
by: Jia, Yufei, et al.
Published: (2026)
FILIC: Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks
by: Ge, Haizhou, et al.
Published: (2025)
by: Ge, Haizhou, et al.
Published: (2025)
FGGS-LiDAR: Ultra-Fast, GPU-Accelerated Simulation from General 3DGS Models to LiDAR
by: Wu, Junzhe, et al.
Published: (2025)
by: Wu, Junzhe, et al.
Published: (2025)
ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation
by: Chen, Kewei, et al.
Published: (2026)
by: Chen, Kewei, et al.
Published: (2026)
Exploiting Policy Idling for Dexterous Manipulation
by: Chen, Annie S., et al.
Published: (2025)
by: Chen, Annie S., et al.
Published: (2025)
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
by: Koh, Hyunseo, et al.
Published: (2026)
by: Koh, Hyunseo, et al.
Published: (2026)
LoHoVLA: A Unified Vision-Language-Action Model for Long-Horizon Embodied Tasks
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
A Survey on Vision-Language-Action Models for Embodied AI
by: Ma, Yueen, et al.
Published: (2024)
by: Ma, Yueen, et al.
Published: (2024)
Recent Advances of Deep Robotic Affordance Learning: A Reinforcement Learning Perspective
by: Yang, Xintong, et al.
Published: (2023)
by: Yang, Xintong, et al.
Published: (2023)
Open, Reproducible and Trustworthy Robot-Based Experiments with Virtual Labs and Digital-Twin-Based Execution Tracing
by: Alt, Benjamin, et al.
Published: (2025)
by: Alt, Benjamin, et al.
Published: (2025)
Teacher Motion Priors: Enhancing Robot Locomotion over Challenging Terrain
by: Jin, Fangcheng, et al.
Published: (2025)
by: Jin, Fangcheng, et al.
Published: (2025)
Offline Goal-Conditioned Reinforcement Learning for Safety-Critical Tasks with Recovery Policy
by: Cao, Chenyang, et al.
Published: (2024)
by: Cao, Chenyang, et al.
Published: (2024)
Automated Feature Selection for Inverse Reinforcement Learning
by: Baimukashev, Daulet, et al.
Published: (2024)
by: Baimukashev, Daulet, et al.
Published: (2024)
Drift is a Sampling Error: SNR-Aware Power Distributions for Long-Horizon Robotic Planning
by: Chen, Kewei, et al.
Published: (2026)
by: Chen, Kewei, et al.
Published: (2026)
Similar Items
-
CLAW: A Vision-Language-Action Framework for Weight-Aware Robotic Grasping
by: An, Zijian, et al.
Published: (2025) -
MaP-AVR: A Meta-Action Planner for Agents Leveraging Vision Language Models and Retrieval-Augmented Generation
by: Guo, Zhenglong, et al.
Published: (2025) -
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
by: Syed, Shahram Najam, et al.
Published: (2025) -
Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization
by: Yang, Jonathan, et al.
Published: (2025) -
Action-Aware Pro-Active Safe Exploration for Mobile Robot Mapping
by: İşleyen, Aykut, et al.
Published: (2025)