ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yan, Hongyu, Li, Qiwei, Yang, Jiaolong, Mu, Yadong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
von: Shen, Yichao, et al.
Veröffentlicht: (2025)
von: Shen, Yichao, et al.
Veröffentlicht: (2025)
HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies
von: Du, Zhiying, et al.
Veröffentlicht: (2025)
von: Du, Zhiying, et al.
Veröffentlicht: (2025)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
von: Miao, Cui, et al.
Veröffentlicht: (2025)
von: Miao, Cui, et al.
Veröffentlicht: (2025)
SignVLA: A Gloss-Free Vision-Language-Action Framework for Real-Time Sign Language-Guided Robotic Manipulation
von: Tan, Xinyu, et al.
Veröffentlicht: (2026)
von: Tan, Xinyu, et al.
Veröffentlicht: (2026)
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)
Asynchronous Fast-Slow Vision-Language-Action Policies for Whole-Body Robotic Manipulation
von: Zou, Teqiang, et al.
Veröffentlicht: (2025)
von: Zou, Teqiang, et al.
Veröffentlicht: (2025)
AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation
von: Zhao, Ziyan, et al.
Veröffentlicht: (2025)
von: Zhao, Ziyan, et al.
Veröffentlicht: (2025)
ReconVLA: An Uncertainty-Guided and Failure-Aware Vision-Language-Action Framework for Robotic Control
von: Chen, Lingling, et al.
Veröffentlicht: (2026)
von: Chen, Lingling, et al.
Veröffentlicht: (2026)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models
von: Li, Peiyan, et al.
Veröffentlicht: (2025)
von: Li, Peiyan, et al.
Veröffentlicht: (2025)
Adaptive Diffusion Policy Optimization for Robotic Manipulation
von: Jiang, Huiyun, et al.
Veröffentlicht: (2025)
von: Jiang, Huiyun, et al.
Veröffentlicht: (2025)
VacuumVLA: Boosting VLA Capabilities via a Unified Suction and Gripping Tool for Complex Robotic Manipulation
von: Zhou, Hui, et al.
Veröffentlicht: (2025)
von: Zhou, Hui, et al.
Veröffentlicht: (2025)
RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models
von: Liufu, Weijia, et al.
Veröffentlicht: (2026)
von: Liufu, Weijia, et al.
Veröffentlicht: (2026)
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
von: Hu, Xintong, et al.
Veröffentlicht: (2026)
von: Hu, Xintong, et al.
Veröffentlicht: (2026)
DiffVLA: Vision-Language Guided Diffusion Planning for Autonomous Driving
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
von: Jiang, Anqing, et al.
Veröffentlicht: (2025)
DroneVLA: VLA based Aerial Manipulation
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
von: Mehboob, Fawad, et al.
Veröffentlicht: (2026)
RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design
von: Chen, Tianxing, et al.
Veröffentlicht: (2026)
von: Chen, Tianxing, et al.
Veröffentlicht: (2026)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
von: Su, Taiyi, et al.
Veröffentlicht: (2026)
von: Su, Taiyi, et al.
Veröffentlicht: (2026)
Characterizing Vision-Language-Action Models across XPUs: Constraints and Acceleration for On-Robot Deployment
von: Zhou, Kaijun, et al.
Veröffentlicht: (2026)
von: Zhou, Kaijun, et al.
Veröffentlicht: (2026)
AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
von: Gubernatorov, Konstantin, et al.
Veröffentlicht: (2025)
AnoleVLA: Lightweight Vision-Language-Action Model with Deep State Space Models for Mobile Manipulation
von: Takagi, Yusuke, et al.
Veröffentlicht: (2026)
von: Takagi, Yusuke, et al.
Veröffentlicht: (2026)
SDP: Spiking Diffusion Policy for Robotic Manipulation with Learnable Channel-Wise Membrane Thresholds
von: Hou, Zhixing, et al.
Veröffentlicht: (2024)
von: Hou, Zhixing, et al.
Veröffentlicht: (2024)
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
von: Lu, Guanxing, et al.
Veröffentlicht: (2025)
Parallels Between VLA Model Post-Training and Human Motor Learning: Progress, Challenges, and Trends
von: Xiang, Tian-Yu, et al.
Veröffentlicht: (2025)
von: Xiang, Tian-Yu, et al.
Veröffentlicht: (2025)
PaCo-VLA: Passivity-Shielded Compliance Prior for Contact-Rich Vision-Language-Action Manipulation
von: Cao, Haofan, et al.
Veröffentlicht: (2026)
von: Cao, Haofan, et al.
Veröffentlicht: (2026)
ForeDiffusion: Foresight-Conditioned Diffusion Policy via Future View Construction for Robot Manipulation
von: Xie, Weize, et al.
Veröffentlicht: (2026)
von: Xie, Weize, et al.
Veröffentlicht: (2026)
AdaWorldPolicy: World-Model-Driven Diffusion Policy with Online Adaptive Learning for Robotic Manipulation
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
von: Yuan, Ge, et al.
Veröffentlicht: (2026)
Look Before You Leap: Using Serialized State Machine for Language Conditioned Robotic Manipulation
von: Mu, Tong, et al.
Veröffentlicht: (2025)
von: Mu, Tong, et al.
Veröffentlicht: (2025)
World4RL: Diffusion World Models for Policy Refinement with Reinforcement Learning for Robotic Manipulation
von: Jiang, Zhennan, et al.
Veröffentlicht: (2025)
von: Jiang, Zhennan, et al.
Veröffentlicht: (2025)
Towards Deploying VLA without Fine-Tuning: Plug-and-Play Inference-Time VLA Policy Steering via Embodied Evolutionary Diffusion
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
Surfer: Progressive Reasoning with World Models for Robotic Manipulation
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
von: Ren, Pengzhen, et al.
Veröffentlicht: (2023)
Build on Priors: Vision--Language--Guided Neuro-Symbolic Imitation Learning for Data-Efficient Real-World Robot Manipulation
von: Lorang, Pierrick, et al.
Veröffentlicht: (2026)
von: Lorang, Pierrick, et al.
Veröffentlicht: (2026)
TraceVLA: Visual Trace Prompting Enhances Spatial-Temporal Awareness for Generalist Robotic Policies
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
von: Zheng, Ruijie, et al.
Veröffentlicht: (2024)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
von: Zhang, Yihao, et al.
Veröffentlicht: (2025)
von: Zhang, Yihao, et al.
Veröffentlicht: (2025)
End-to-End Dexterous Arm-Hand VLA Policies via Shared Autonomy: VR Teleoperation Augmented by Autonomous Hand VLA Policy for Efficient Data Collection
von: Cui, Yu, et al.
Veröffentlicht: (2025)
von: Cui, Yu, et al.
Veröffentlicht: (2025)
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation
von: Hong, Youngjin, et al.
Veröffentlicht: (2025)
von: Hong, Youngjin, et al.
Veröffentlicht: (2025)
FP3: A 3D Foundation Policy for Robotic Manipulation
von: Yang, Rujia, et al.
Veröffentlicht: (2025)
von: Yang, Rujia, et al.
Veröffentlicht: (2025)
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
von: Wang, Qiuyue, et al.
Veröffentlicht: (2026)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
von: Wang, Feiyang, et al.
Veröffentlicht: (2025)
von: Wang, Feiyang, et al.
Veröffentlicht: (2025)
Vision-Language-Policy Model for Dynamic Robot Task Planning
von: Wang, Jin, et al.
Veröffentlicht: (2025)
von: Wang, Jin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
von: Shen, Yichao, et al.
Veröffentlicht: (2025) -
HiMoE-VLA: Hierarchical Mixture-of-Experts for Generalist Vision-Language-Action Policies
von: Du, Zhiying, et al.
Veröffentlicht: (2025) -
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
von: Miao, Cui, et al.
Veröffentlicht: (2025) -
SignVLA: A Gloss-Free Vision-Language-Action Framework for Real-Time Sign Language-Guided Robotic Manipulation
von: Tan, Xinyu, et al.
Veröffentlicht: (2026) -
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy
von: Chen, Xinyi, et al.
Veröffentlicht: (2025)