Instruct2Act: From Human Instruction to Actions Sequencing and Execution via Robot Action Network for Robotic Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sharma, Archit, Sharma, Dharmendra, Rebeiro, John, Thakur, Peeyush, Dhar, Narendra, Behera, Laxmidhar |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RoboSubtaskNet: Temporal Sub-task Segmentation for Human-to-Robot Skill Transfer in Real-World Environments
von: Sharma, Dharmendra, et al.
Veröffentlicht: (2026)
von: Sharma, Dharmendra, et al.
Veröffentlicht: (2026)
Dynamic Hand Gesture Recognition for Robot Manipulator Tasks
von: Sharma, Dharmendra, et al.
Veröffentlicht: (2026)
von: Sharma, Dharmendra, et al.
Veröffentlicht: (2026)
Autoregressive Action Sequence Learning for Robotic Manipulation
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024)
PartInstruct: Part-level Instruction Following for Fine-grained Robot Manipulation
von: Yin, Yifan, et al.
Veröffentlicht: (2025)
von: Yin, Yifan, et al.
Veröffentlicht: (2025)
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
von: Li, Yajie, et al.
Veröffentlicht: (2026)
von: Li, Yajie, et al.
Veröffentlicht: (2026)
EVA: Aligning Video World Models with Executable Robot Actions via Inverse Dynamics Rewards
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
von: Wang, Ruixiang, et al.
Veröffentlicht: (2026)
Human-assisted Robotic Policy Refinement via Action Preference Optimization
von: Xia, Wenke, et al.
Veröffentlicht: (2025)
von: Xia, Wenke, et al.
Veröffentlicht: (2025)
CLASS: Contrastive Learning via Action Sequence Supervision for Robot Manipulation
von: Lee, Sung-Wook, et al.
Veröffentlicht: (2025)
von: Lee, Sung-Wook, et al.
Veröffentlicht: (2025)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
von: Zhang, Yihao, et al.
Veröffentlicht: (2025)
von: Zhang, Yihao, et al.
Veröffentlicht: (2025)
OASIS: Observation-Action Space Alignment via SE(3) Trajectory Prediction for Robotic Manipulation
von: Chen, Xinzhe, et al.
Veröffentlicht: (2026)
von: Chen, Xinzhe, et al.
Veröffentlicht: (2026)
Emergence of Human to Robot Transfer in Vision-Language-Action Models
von: Kareer, Simar, et al.
Veröffentlicht: (2025)
von: Kareer, Simar, et al.
Veröffentlicht: (2025)
Thrust Microstepping via Acceleration Feedback in Quadrotor Control for Aerial Grasping of Dynamic Payload
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
RoboAct-CLIP: Video-Driven Pre-training of Atomic Action Understanding for Robotics
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Zhiyuan, et al.
Veröffentlicht: (2025)
Action-Sketcher: From Reasoning to Action via Visual Sketches for Long-Horizon Robotic Manipulation
von: Tan, Huajie, et al.
Veröffentlicht: (2026)
von: Tan, Huajie, et al.
Veröffentlicht: (2026)
ManipDreamer: Boosting Robotic Manipulation World Model with Action Tree and Visual Guidance
von: Li, Ying, et al.
Veröffentlicht: (2025)
von: Li, Ying, et al.
Veröffentlicht: (2025)
Asynchronous Fast-Slow Vision-Language-Action Policies for Whole-Body Robotic Manipulation
von: Zou, Teqiang, et al.
Veröffentlicht: (2025)
von: Zou, Teqiang, et al.
Veröffentlicht: (2025)
EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow
von: Chen, Yixiang, et al.
Veröffentlicht: (2025)
von: Chen, Yixiang, et al.
Veröffentlicht: (2025)
HAMSTER: Hierarchical Action Models For Open-World Robot Manipulation
von: Li, Yi, et al.
Veröffentlicht: (2025)
von: Li, Yi, et al.
Veröffentlicht: (2025)
Recognizing Actions from Robotic View for Natural Human-Robot Interaction
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
von: Wang, Ziyi, et al.
Veröffentlicht: (2025)
Observe Then Act: Asynchronous Active Vision-Action Model for Robotic Manipulation
von: Wang, Guokang, et al.
Veröffentlicht: (2024)
von: Wang, Guokang, et al.
Veröffentlicht: (2024)
From LLMs to Actions: Latent Codes as Bridges in Hierarchical Robot Control
von: Shentu, Yide, et al.
Veröffentlicht: (2024)
von: Shentu, Yide, et al.
Veröffentlicht: (2024)
Open-Ended Goal Inference through Actions and Language for Human-Robot Collaboration
von: Ghose, Debasmita, et al.
Veröffentlicht: (2025)
von: Ghose, Debasmita, et al.
Veröffentlicht: (2025)
Action Flow Matching for Continual Robot Learning
von: Murillo-Gonzalez, Alejandro, et al.
Veröffentlicht: (2025)
von: Murillo-Gonzalez, Alejandro, et al.
Veröffentlicht: (2025)
Toward Accurate Long-Horizon Robotic Manipulation: Language-to-Action with Foundation Models via Scene Graphs
von: Dinesh, Sushil Samuel, et al.
Veröffentlicht: (2025)
von: Dinesh, Sushil Samuel, et al.
Veröffentlicht: (2025)
VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
von: Zhao, Wei, et al.
Veröffentlicht: (2025)
Gaze2Act: Gaze-Conditioned Vision-Language-Action Policies for Interactive Robot Manipulation
von: Zuo, Kuangji, et al.
Veröffentlicht: (2026)
von: Zuo, Kuangji, et al.
Veröffentlicht: (2026)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
von: Miao, Cui, et al.
Veröffentlicht: (2025)
von: Miao, Cui, et al.
Veröffentlicht: (2025)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
von: Wang, Zhijie, et al.
Veröffentlicht: (2024)
von: Wang, Zhijie, et al.
Veröffentlicht: (2024)
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation
von: Hong, Youngjin, et al.
Veröffentlicht: (2025)
von: Hong, Youngjin, et al.
Veröffentlicht: (2025)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
Hi Robot: Open-Ended Instruction Following with Hierarchical Vision-Language-Action Models
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
von: Shi, Lucy Xiaoyang, et al.
Veröffentlicht: (2025)
Design, Localization, Perception, and Control for GPS-Denied Autonomous Aerial Grasping and Harvesting
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
von: Kumar, Ashish, et al.
Veröffentlicht: (2024)
When to Trust Imagination: Adaptive Action Execution for World Action Models
von: Wang, Rui, et al.
Veröffentlicht: (2026)
von: Wang, Rui, et al.
Veröffentlicht: (2026)
Toward Aligning Human and Robot Actions via Multi-Modal Demonstration Learning
von: Zahid, Azizul, et al.
Veröffentlicht: (2025)
von: Zahid, Azizul, et al.
Veröffentlicht: (2025)
Unsupervised Learning of Effective Actions in Robotics
von: Zaric, Marko, et al.
Veröffentlicht: (2024)
von: Zaric, Marko, et al.
Veröffentlicht: (2024)
InstructVLA: Vision-Language-Action Instruction Tuning from Understanding to Manipulation
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
von: Yang, Shuai, et al.
Veröffentlicht: (2025)
Real-Time Robot Execution with Masked Action Chunking
von: Wang, Haoxuan, et al.
Veröffentlicht: (2026)
von: Wang, Haoxuan, et al.
Veröffentlicht: (2026)
Adversarial Attacks on Robotic Vision Language Action Models
von: Jones, Eliot Krzysztof, et al.
Veröffentlicht: (2025)
von: Jones, Eliot Krzysztof, et al.
Veröffentlicht: (2025)
Leveraging Foundation Models for Enhancing Robot Perception and Action
von: Mirjalili, Reihaneh
Veröffentlicht: (2025)
von: Mirjalili, Reihaneh
Veröffentlicht: (2025)
Ähnliche Einträge
-
RoboSubtaskNet: Temporal Sub-task Segmentation for Human-to-Robot Skill Transfer in Real-World Environments
von: Sharma, Dharmendra, et al.
Veröffentlicht: (2026) -
Dynamic Hand Gesture Recognition for Robot Manipulator Tasks
von: Sharma, Dharmendra, et al.
Veröffentlicht: (2026) -
Autoregressive Action Sequence Learning for Robotic Manipulation
von: Zhang, Xinyu, et al.
Veröffentlicht: (2024) -
PartInstruct: Part-level Instruction Following for Fine-grained Robot Manipulation
von: Yin, Yifan, et al.
Veröffentlicht: (2025) -
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation
von: Li, Yajie, et al.
Veröffentlicht: (2026)