Empowering Large Language Models on Robotic Manipulation with Affordance Prompting
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cheng, Guangran, Zhang, Chuheng, Cai, Wenzhe, Zhao, Li, Sun, Changyin, Bian, Jiang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Affordance-based Robot Manipulation with Flow Matching
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
von: Zhang, Fan, et al.
Veröffentlicht: (2024)
ARO: Large Language Model Supervised Robotics Text2Skill Autonomous Learning
von: Chen, Yiwen, et al.
Veröffentlicht: (2024)
von: Chen, Yiwen, et al.
Veröffentlicht: (2024)
ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
von: Wang, Teng, et al.
Veröffentlicht: (2025)
von: Wang, Teng, et al.
Veröffentlicht: (2025)
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
von: Tang, Yihe, et al.
Veröffentlicht: (2025)
von: Tang, Yihe, et al.
Veröffentlicht: (2025)
Policy Filtration for RLHF to Mitigate Noise in Reward Models
von: Zhang, Chuheng, et al.
Veröffentlicht: (2024)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2024)
Hard Prompts Made Interpretable: Sparse Entropy Regularization for Prompt Tuning with RL
von: Choi, Yunseon, et al.
Veröffentlicht: (2024)
von: Choi, Yunseon, et al.
Veröffentlicht: (2024)
Affordance Benchmark for MLLMs
von: Wang, Junying, et al.
Veröffentlicht: (2025)
von: Wang, Junying, et al.
Veröffentlicht: (2025)
AffordSim: A Scalable Data Generator and Benchmark for Affordance-Aware Robotic Manipulation
von: Li, Mingyang, et al.
Veröffentlicht: (2026)
von: Li, Mingyang, et al.
Veröffentlicht: (2026)
Effects of Prompt Length on Domain-specific Tasks for Large Language Models
von: Liu, Qibang, et al.
Veröffentlicht: (2025)
von: Liu, Qibang, et al.
Veröffentlicht: (2025)
AnchorDP3: 3D Affordance Guided Sparse Diffusion Policy for Robotic Manipulation
von: Zhao, Ziyan, et al.
Veröffentlicht: (2025)
von: Zhao, Ziyan, et al.
Veröffentlicht: (2025)
RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
von: Nasiriany, Soroush, et al.
Veröffentlicht: (2024)
What Do Latent Action Models Actually Learn?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
von: Yang, Rushuai, et al.
Veröffentlicht: (2025)
How Do VLAs Effectively Inherit from VLMs?
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
von: Zhang, Chuheng, et al.
Veröffentlicht: (2025)
IGOR: Image-GOal Representations are the Atomic Control Units for Foundation Models in Embodied AI
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2024)
Unveiling and Manipulating Prompt Influence in Large Language Models
von: Feng, Zijian, et al.
Veröffentlicht: (2024)
von: Feng, Zijian, et al.
Veröffentlicht: (2024)
Model Tuning or Prompt Tuning? A Study of Large Language Models for Clinical Concept and Relation Extraction
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
von: Peng, Cheng, et al.
Veröffentlicht: (2023)
Information-driven Affordance Discovery for Efficient Robotic Manipulation
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
von: Mazzaglia, Pietro, et al.
Veröffentlicht: (2024)
FlexCAD: Unified and Versatile Controllable CAD Generation with Fine-tuned Large Language Models
von: Zhang, Zhanwei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhanwei, et al.
Veröffentlicht: (2024)
villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoyu, et al.
Veröffentlicht: (2025)
WizardCoder: Empowering Code Large Language Models with Evol-Instruct
von: Luo, Ziyang, et al.
Veröffentlicht: (2023)
von: Luo, Ziyang, et al.
Veröffentlicht: (2023)
On the Worst Prompt Performance of Large Language Models
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
von: Cao, Bowen, et al.
Veröffentlicht: (2024)
AFFORD2ACT: Affordance-Guided Automatic Keypoint Selection for Generalizable and Lightweight Robotic Manipulation
von: Singh, Anukriti, et al.
Veröffentlicht: (2025)
von: Singh, Anukriti, et al.
Veröffentlicht: (2025)
Peer-aided Repairer: Empowering Large Language Models to Repair Advanced Student Assignments
von: Zhao, Qianhui, et al.
Veröffentlicht: (2024)
von: Zhao, Qianhui, et al.
Veröffentlicht: (2024)
Large Language Models Empowered Personalized Web Agents
von: Cai, Hongru, et al.
Veröffentlicht: (2024)
von: Cai, Hongru, et al.
Veröffentlicht: (2024)
Logic-of-Thought: Empowering Large Language Models with Logic Programs for Solving Puzzles in Natural Language
von: Li, Naiqi, et al.
Veröffentlicht: (2025)
von: Li, Naiqi, et al.
Veröffentlicht: (2025)
Empowering Denoising Sequential Recommendation with Large Language Model Embeddings
von: Wu, Tongzhou, et al.
Veröffentlicht: (2025)
von: Wu, Tongzhou, et al.
Veröffentlicht: (2025)
Empowering Working Memory for Large Language Model Agents
von: Guo, Jing, et al.
Veröffentlicht: (2023)
von: Guo, Jing, et al.
Veröffentlicht: (2023)
Incentivizing Multimodal Reasoning in Large Models for Direct Robot Manipulation
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
von: Tang, Weiliang, et al.
Veröffentlicht: (2025)
General-Purpose Aerial Intelligent Agents Empowered by Large Language Models
von: Zhao, Ji, et al.
Veröffentlicht: (2025)
von: Zhao, Ji, et al.
Veröffentlicht: (2025)
LUK: Empowering Log Understanding with Expert Knowledge from Large Language Models
von: Ma, Lipeng, et al.
Veröffentlicht: (2024)
von: Ma, Lipeng, et al.
Veröffentlicht: (2024)
Pre-train, Align, and Disentangle: Empowering Sequential Recommendation with Large Language Models
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
von: Wang, Yuhao, et al.
Veröffentlicht: (2024)
Safe Semantics, Unsafe Interpretations: Tackling Implicit Reasoning Safety in Large Vision-Language Models
von: Cai, Wei, et al.
Veröffentlicht: (2025)
von: Cai, Wei, et al.
Veröffentlicht: (2025)
Hierarchical Language Models for Semantic Navigation and Manipulation in an Aerial-Ground Robotic System
von: Liu, Haokun, et al.
Veröffentlicht: (2025)
von: Liu, Haokun, et al.
Veröffentlicht: (2025)
KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation
von: Liu, Zixian, et al.
Veröffentlicht: (2025)
von: Liu, Zixian, et al.
Veröffentlicht: (2025)
A3D: Adaptive Affordance Assembly with Dual-Arm Manipulation
von: Liang, Jiaqi, et al.
Veröffentlicht: (2026)
von: Liang, Jiaqi, et al.
Veröffentlicht: (2026)
Evolving in Tasks: Empowering the Multi-modality Large Language Model as the Computer Use Agent
von: Cheng, Yuhao, et al.
Veröffentlicht: (2025)
von: Cheng, Yuhao, et al.
Veröffentlicht: (2025)
SGFormer: Simplifying and Empowering Transformers for Large-Graph Representations
von: Wu, Qitian, et al.
Veröffentlicht: (2023)
von: Wu, Qitian, et al.
Veröffentlicht: (2023)
ChatAD: Reasoning-Enhanced Time-Series Anomaly Detection with Multi-Turn Instruction Evolution
von: Sun, Hui, et al.
Veröffentlicht: (2026)
von: Sun, Hui, et al.
Veröffentlicht: (2026)
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation
von: Singh, Harsh, et al.
Veröffentlicht: (2024)
von: Singh, Harsh, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Affordance-based Robot Manipulation with Flow Matching
von: Zhang, Fan, et al.
Veröffentlicht: (2024) -
ARO: Large Language Model Supervised Robotics Text2Skill Autonomous Learning
von: Chen, Yiwen, et al.
Veröffentlicht: (2024) -
ImagineNav++: Prompting Vision-Language Models as Embodied Navigator through Scene Imagination
von: Wang, Teng, et al.
Veröffentlicht: (2025) -
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
von: Tang, Yihe, et al.
Veröffentlicht: (2025) -
Policy Filtration for RLHF to Mitigate Noise in Reward Models
von: Zhang, Chuheng, et al.
Veröffentlicht: (2024)