CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Feiyang, Yu, Xiaomin, Wu, Wangyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
von: Gao, Jensen, et al.
Veröffentlicht: (2023)
STEER: Flexible Robotic Manipulation via Dense Language Grounding
von: Smith, Laura, et al.
Veröffentlicht: (2024)
von: Smith, Laura, et al.
Veröffentlicht: (2024)
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation
von: Hong, Youngjin, et al.
Veröffentlicht: (2025)
von: Hong, Youngjin, et al.
Veröffentlicht: (2025)
Hierarchical Language Models for Semantic Navigation and Manipulation in an Aerial-Ground Robotic System
von: Liu, Haokun, et al.
Veröffentlicht: (2025)
von: Liu, Haokun, et al.
Veröffentlicht: (2025)
Grounding Sim-to-Real Generalization in Dexterous Manipulation: An Empirical Study with Vision-Language-Action Models
von: Jin, Ruixing, et al.
Veröffentlicht: (2026)
von: Jin, Ruixing, et al.
Veröffentlicht: (2026)
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation
von: Chen, Shizhe, et al.
Veröffentlicht: (2025)
von: Chen, Shizhe, et al.
Veröffentlicht: (2025)
ManipBench: Benchmarking Vision-Language Models for Low-Level Robot Manipulation
von: Zhao, Enyu, et al.
Veröffentlicht: (2025)
von: Zhao, Enyu, et al.
Veröffentlicht: (2025)
Experiences from Benchmarking Vision-Language-Action Models for Robotic Manipulation
von: Zhang, Yihao, et al.
Veröffentlicht: (2025)
von: Zhang, Yihao, et al.
Veröffentlicht: (2025)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
von: Wang, Zhijie, et al.
Veröffentlicht: (2024)
von: Wang, Zhijie, et al.
Veröffentlicht: (2024)
Universality in Collective Intelligence on the Rubik's Cube
von: Krakauer, David, et al.
Veröffentlicht: (2025)
von: Krakauer, David, et al.
Veröffentlicht: (2025)
Solving Rubik's Cube Without Tricky Sampling
von: Lin, Yicheng, et al.
Veröffentlicht: (2024)
von: Lin, Yicheng, et al.
Veröffentlicht: (2024)
Mechanical Automation with Vision: A Design for Rubik's Cube Solver
von: Chalise, Abhinav, et al.
Veröffentlicht: (2025)
von: Chalise, Abhinav, et al.
Veröffentlicht: (2025)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
von: Liu, Jinkun, et al.
Veröffentlicht: (2026)
von: Liu, Jinkun, et al.
Veröffentlicht: (2026)
FedVLA: Federated Vision-Language-Action Learning with Dual Gating Mixture-of-Experts for Robotic Manipulation
von: Miao, Cui, et al.
Veröffentlicht: (2025)
von: Miao, Cui, et al.
Veröffentlicht: (2025)
ProgressVLA: Progress-Guided Diffusion Policy for Vision-Language Robotic Manipulation
von: Yan, Hongyu, et al.
Veröffentlicht: (2026)
von: Yan, Hongyu, et al.
Veröffentlicht: (2026)
MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
von: Zhang, Rongyu, et al.
Veröffentlicht: (2025)
Wormhole Memory: A Rubik's Cube for Cross-Dialogue Retrieval
von: Wang, Libo
Veröffentlicht: (2025)
von: Wang, Libo
Veröffentlicht: (2025)
Survey of Vision-Language-Action Models for Embodied Manipulation
von: Li, Haoran, et al.
Veröffentlicht: (2025)
von: Li, Haoran, et al.
Veröffentlicht: (2025)
Lifelong Language-Conditioned Robotic Manipulation Learning
von: Wang, Xudong, et al.
Veröffentlicht: (2026)
von: Wang, Xudong, et al.
Veröffentlicht: (2026)
Asynchronous Fast-Slow Vision-Language-Action Policies for Whole-Body Robotic Manipulation
von: Zou, Teqiang, et al.
Veröffentlicht: (2025)
von: Zou, Teqiang, et al.
Veröffentlicht: (2025)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
von: Wu, Yanru, et al.
Veröffentlicht: (2026)
Solving a Rubik's Cube Using its Local Graph Structure
von: Yao, Shunyu, et al.
Veröffentlicht: (2024)
von: Yao, Shunyu, et al.
Veröffentlicht: (2024)
HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models
von: Bhat, Vineet, et al.
Veröffentlicht: (2024)
von: Bhat, Vineet, et al.
Veröffentlicht: (2024)
Grounding Language Models with Semantic Digital Twins for Robotic Planning
von: Naeem, Mehreen, et al.
Veröffentlicht: (2025)
von: Naeem, Mehreen, et al.
Veröffentlicht: (2025)
BridgeVLA: Input-Output Alignment for Efficient 3D Manipulation Learning with Vision-Language Models
von: Li, Peiyan, et al.
Veröffentlicht: (2025)
von: Li, Peiyan, et al.
Veröffentlicht: (2025)
Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics
von: Wang, Taowen, et al.
Veröffentlicht: (2024)
von: Wang, Taowen, et al.
Veröffentlicht: (2024)
BUMBLE: Unifying Reasoning and Acting with Vision-Language Models for Building-wide Mobile Manipulation
von: Shah, Rutav, et al.
Veröffentlicht: (2024)
von: Shah, Rutav, et al.
Veröffentlicht: (2024)
Vision-Language-Policy Model for Dynamic Robot Task Planning
von: Wang, Jin, et al.
Veröffentlicht: (2025)
von: Wang, Jin, et al.
Veröffentlicht: (2025)
Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation
von: Feng, Yunhai, et al.
Veröffentlicht: (2025)
von: Feng, Yunhai, et al.
Veröffentlicht: (2025)
MORE: Mobile Manipulation Rearrangement Through Grounded Language Reasoning
von: Mohammadi, Mohammad, et al.
Veröffentlicht: (2025)
von: Mohammadi, Mohammad, et al.
Veröffentlicht: (2025)
Adversarial Attacks on Robotic Vision Language Action Models
von: Jones, Eliot Krzysztof, et al.
Veröffentlicht: (2025)
von: Jones, Eliot Krzysztof, et al.
Veröffentlicht: (2025)
Grounded Vision-Language Navigation for UAVs with Open-Vocabulary Goal Understanding
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
von: Zhang, Yuhang, et al.
Veröffentlicht: (2025)
Explainable Adversarial-Robust Vision-Language-Action Model for Robotic Manipulation
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
von: Kim, Ju-Young, et al.
Veröffentlicht: (2025)
Vision-Language Foundation Models as Effective Robot Imitators
von: Li, Xinghang, et al.
Veröffentlicht: (2023)
von: Li, Xinghang, et al.
Veröffentlicht: (2023)
GraspCorrect: Robotic Grasp Correction via Vision-Language Model-Guided Feedback
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
von: Lee, Sungjae, et al.
Veröffentlicht: (2025)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
von: Tang, Grace, et al.
Veröffentlicht: (2024)
von: Tang, Grace, et al.
Veröffentlicht: (2024)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
von: Singh, Harsh, et al.
Veröffentlicht: (2024)
von: Singh, Harsh, et al.
Veröffentlicht: (2024)
Hybrid Framework for Robotic Manipulation: Integrating Reinforcement Learning and Large Language Models
von: Saad, Md, et al.
Veröffentlicht: (2026)
von: Saad, Md, et al.
Veröffentlicht: (2026)
Vision-Based Hand Shadowing for Robotic Manipulation via Inverse Kinematics
von: Chiche, Hendrik, et al.
Veröffentlicht: (2026)
von: Chiche, Hendrik, et al.
Veröffentlicht: (2026)
FrankenBot: Brain-Morphic Modular Orchestration for Robotic Manipulation with Vision-Language Models
von: Wang, Shiyi, et al.
Veröffentlicht: (2025)
von: Wang, Shiyi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Physically Grounded Vision-Language Models for Robotic Manipulation
von: Gao, Jensen, et al.
Veröffentlicht: (2023) -
STEER: Flexible Robotic Manipulation via Dense Language Grounding
von: Smith, Laura, et al.
Veröffentlicht: (2024) -
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation
von: Hong, Youngjin, et al.
Veröffentlicht: (2025) -
Hierarchical Language Models for Semantic Navigation and Manipulation in an Aerial-Ground Robotic System
von: Liu, Haokun, et al.
Veröffentlicht: (2025) -
Grounding Sim-to-Real Generalization in Dexterous Manipulation: An Empirical Study with Vision-Language-Action Models
von: Jin, Ruixing, et al.
Veröffentlicht: (2026)