AffordanceLLM: Grounding Affordance from Vision Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Qian, Shengyi, Chen, Weifeng, Bai, Min, Zhou, Xiong, Tu, Zhuowen, Li, Li Erran |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
por: Chu, Hengshuo, et al.
Publicado: (2025)
por: Chu, Hengshuo, et al.
Publicado: (2025)
Panoramic Affordance Prediction
por: Zhang, Zixin, et al.
Publicado: (2026)
por: Zhang, Zixin, et al.
Publicado: (2026)
Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model
por: Wang, Hanqing, et al.
Publicado: (2025)
por: Wang, Hanqing, et al.
Publicado: (2025)
PanoAffordanceNet: Towards Holistic Affordance Grounding in 360° Indoor Environments
por: Zhu, Guoliang, et al.
Publicado: (2026)
por: Zhu, Guoliang, et al.
Publicado: (2026)
PAVLM: Advancing Point Cloud based Affordance Understanding Via Vision-Language Model
por: Liu, Shang-Ching, et al.
Publicado: (2024)
por: Liu, Shang-Ching, et al.
Publicado: (2024)
ViGoR: Improving Visual Grounding of Large Vision Language Models with Fine-Grained Reward Modeling
por: Yan, Siming, et al.
Publicado: (2024)
por: Yan, Siming, et al.
Publicado: (2024)
CompassAD: Intent-Driven 3D Affordance Grounding in Functionally Competing Objects
por: Li, Jingliang, et al.
Publicado: (2026)
por: Li, Jingliang, et al.
Publicado: (2026)
ManipGPT: Is Affordance Segmentation by Large Vision Models Enough for Articulated Object Manipulation?
por: Kim, Taewhan, et al.
Publicado: (2024)
por: Kim, Taewhan, et al.
Publicado: (2024)
Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions
por: Zhu, He, et al.
Publicado: (2025)
por: Zhu, He, et al.
Publicado: (2025)
AffordanceGrasp-R1:Leveraging Reasoning-Based Affordance Segmentation with Reinforcement Learning for Robotic Grasping
por: Zhou, Dingyi, et al.
Publicado: (2026)
por: Zhou, Dingyi, et al.
Publicado: (2026)
Affordance Agent Harness: Verification-Gated Skill Orchestration
por: Huang, Haojian, et al.
Publicado: (2026)
por: Huang, Haojian, et al.
Publicado: (2026)
Resource-Efficient Affordance Grounding with Complementary Depth and Semantic Prompts
por: Huang, Yizhou, et al.
Publicado: (2025)
por: Huang, Yizhou, et al.
Publicado: (2025)
Egocentric Instruction-oriented Affordance Prediction via Large Multimodal Model
por: Ji, Bokai, et al.
Publicado: (2025)
por: Ji, Bokai, et al.
Publicado: (2025)
RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
por: Yuan, Wentao, et al.
Publicado: (2024)
por: Yuan, Wentao, et al.
Publicado: (2024)
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
por: Jia, Wanjun, et al.
Publicado: (2025)
por: Jia, Wanjun, et al.
Publicado: (2025)
Visual Affordance Prediction: Survey and Reproducibility
por: Apicella, Tommaso, et al.
Publicado: (2025)
por: Apicella, Tommaso, et al.
Publicado: (2025)
Affordance RAG: Hierarchical Multimodal Retrieval with Affordance-Aware Embodied Memory for Mobile Manipulation
por: Korekata, Ryosuke, et al.
Publicado: (2025)
por: Korekata, Ryosuke, et al.
Publicado: (2025)
Learning Precise Affordances from Egocentric Videos for Robotic Manipulation
por: Li, Gen, et al.
Publicado: (2024)
por: Li, Gen, et al.
Publicado: (2024)
PreAfford: Universal Affordance-Based Pre-Grasping for Diverse Objects and Environments
por: Ding, Kairui, et al.
Publicado: (2024)
por: Ding, Kairui, et al.
Publicado: (2024)
TRACER: Texture-Robust Affordance Chain-of-Thought for Deformable-Object Refinement
por: Jia, Wanjun, et al.
Publicado: (2026)
por: Jia, Wanjun, et al.
Publicado: (2026)
NaturalVLM: Leveraging Fine-grained Natural Language for Affordance-Guided Visual Manipulation
por: Xu, Ran, et al.
Publicado: (2024)
por: Xu, Ran, et al.
Publicado: (2024)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
HRP: Human Affordances for Robotic Pre-Training
por: Srirama, Mohan Kumar, et al.
Publicado: (2024)
por: Srirama, Mohan Kumar, et al.
Publicado: (2024)
Interpretable Affordance Detection on 3D Point Clouds with Probabilistic Prototypes
por: Li, Maximilian Xiling, et al.
Publicado: (2025)
por: Li, Maximilian Xiling, et al.
Publicado: (2025)
BEACON: Language-Conditioned Navigation Affordance Prediction under Occlusion
por: Gao, Xinyu, et al.
Publicado: (2026)
por: Gao, Xinyu, et al.
Publicado: (2026)
GLOVER++: Unleashing the Potential of Affordance Learning from Human Behaviors for Robotic Manipulation
por: Ma, Teli, et al.
Publicado: (2025)
por: Ma, Teli, et al.
Publicado: (2025)
GLOVER: Generalizable Open-Vocabulary Affordance Reasoning for Task-Oriented Grasping
por: Ma, Teli, et al.
Publicado: (2024)
por: Ma, Teli, et al.
Publicado: (2024)
Simultaneous Localization and Affordance Prediction of Tasks from Egocentric Video
por: Chavis, Zachary, et al.
Publicado: (2024)
por: Chavis, Zachary, et al.
Publicado: (2024)
A4-Agent: An Agentic Framework for Zero-Shot Affordance Reasoning
por: Zhang, Zixin, et al.
Publicado: (2025)
por: Zhang, Zixin, et al.
Publicado: (2025)
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation
por: Tang, Yihe, et al.
Publicado: (2025)
por: Tang, Yihe, et al.
Publicado: (2025)
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation
por: Tian, Tongxuan, et al.
Publicado: (2025)
por: Tian, Tongxuan, et al.
Publicado: (2025)
RoboPCA: Pose-centered Affordance Learning from Human Demonstrations for Robot Manipulation
por: Xiao, Zhanqi, et al.
Publicado: (2026)
por: Xiao, Zhanqi, et al.
Publicado: (2026)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
por: Jiang, Dengyang, et al.
Publicado: (2025)
por: Jiang, Dengyang, et al.
Publicado: (2025)
DAP: Diffusion-based Affordance Prediction for Multi-modality Storage
por: Chang, Haonan, et al.
Publicado: (2024)
por: Chang, Haonan, et al.
Publicado: (2024)
RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation
por: Nasiriany, Soroush, et al.
Publicado: (2024)
por: Nasiriany, Soroush, et al.
Publicado: (2024)
VoxAfford: Multi-Scale Voxel-Token Fusion for Open-Vocabulary 3D Affordance Detection
por: Sun, Haowen, et al.
Publicado: (2026)
por: Sun, Haowen, et al.
Publicado: (2026)
Learning 6-DoF Fine-grained Grasp Detection Based on Part Affordance Grounding
por: Song, Yaoxian, et al.
Publicado: (2023)
por: Song, Yaoxian, et al.
Publicado: (2023)
Agentic Scene Policies: Unifying Space, Semantics, and Affordances for Robot Action
por: Morin, Sacha, et al.
Publicado: (2025)
por: Morin, Sacha, et al.
Publicado: (2025)
AffordGrasp: Cross-Modal Diffusion for Affordance-Aware Grasp Synthesis
por: Wu, Xiaofei, et al.
Publicado: (2026)
por: Wu, Xiaofei, et al.
Publicado: (2026)
Multi-Keypoint Affordance Representation for Functional Dexterous Grasping
por: Yang, Fan, et al.
Publicado: (2025)
por: Yang, Fan, et al.
Publicado: (2025)
Ejemplares similares
-
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
por: Chu, Hengshuo, et al.
Publicado: (2025) -
Panoramic Affordance Prediction
por: Zhang, Zixin, et al.
Publicado: (2026) -
Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model
por: Wang, Hanqing, et al.
Publicado: (2025) -
PanoAffordanceNet: Towards Holistic Affordance Grounding in 360° Indoor Environments
por: Zhu, Guoliang, et al.
Publicado: (2026) -
PAVLM: Advancing Point Cloud based Affordance Understanding Via Vision-Language Model
por: Liu, Shang-Ching, et al.
Publicado: (2024)