Guardado en:
| Autores principales: | Mao, Aihua, Huang, Kaihang, Liu, Yong-Jin, Chan, Chee Seng, He, Ying |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.20608 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VideoAfford: Grounding 3D Affordance from Human-Object-Interaction Videos via Multimodal Large Language Model
por: Wang, Hanqing, et al.
Publicado: (2026)
por: Wang, Hanqing, et al.
Publicado: (2026)
OneHOI: Unifying Human-Object Interaction Generation and Editing
por: Hoe, Jiun Tian, et al.
Publicado: (2026)
por: Hoe, Jiun Tian, et al.
Publicado: (2026)
Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions
por: Zhu, He, et al.
Publicado: (2025)
por: Zhu, He, et al.
Publicado: (2025)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
por: Hoe, Jiun Tian, et al.
Publicado: (2025)
por: Hoe, Jiun Tian, et al.
Publicado: (2025)
VAGNet: Vision-based Accident Anticipation with Global Features
por: Vipulananthan, Vipooshan, et al.
Publicado: (2026)
por: Vipulananthan, Vipooshan, et al.
Publicado: (2026)
Grounding 3D Scene Affordance From Egocentric Interactions
por: Liu, Cuiyu, et al.
Publicado: (2024)
por: Liu, Cuiyu, et al.
Publicado: (2024)
DMF-Net: Image-Guided Point Cloud Completion with Dual-Channel Modality Fusion and Shape-Aware Upsampling Transformer
por: Mao, Aihua, et al.
Publicado: (2024)
por: Mao, Aihua, et al.
Publicado: (2024)
InteractAnything: Zero-shot Human Object Interaction Synthesis via LLM Feedback and Object Affordance Parsing
por: Zhang, Jinlu, et al.
Publicado: (2025)
por: Zhang, Jinlu, et al.
Publicado: (2025)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
por: Gao, Xianqiang, et al.
Publicado: (2024)
por: Gao, Xianqiang, et al.
Publicado: (2024)
H2OFlow: Grounding Human-Object Affordances with 3D Generative Models and Dense Diffused Flows
por: Zhang, Harry, et al.
Publicado: (2025)
por: Zhang, Harry, et al.
Publicado: (2025)
IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments
por: Zhang, Can, et al.
Publicado: (2025)
por: Zhang, Can, et al.
Publicado: (2025)
Gorgeous: Create Your Desired Character Facial Makeup from Any Ideas
por: Sii, Jia Wei, et al.
Publicado: (2024)
por: Sii, Jia Wei, et al.
Publicado: (2024)
Protégé: Learn and Generate Basic Makeup Styles with Generative Adversarial Networks (GANs)
por: Sii, Jia Wei, et al.
Publicado: (2024)
por: Sii, Jia Wei, et al.
Publicado: (2024)
Yuan: Yielding Unblemished Aesthetics Through A Unified Network for Visual Imperfections Removal in Generated Images
por: Yu, Zhenyu, et al.
Publicado: (2025)
por: Yu, Zhenyu, et al.
Publicado: (2025)
Object Affordance Recognition and Grounding via Multi-scale Cross-modal Representation Learning
por: Wan, Xinhang, et al.
Publicado: (2025)
por: Wan, Xinhang, et al.
Publicado: (2025)
CompassAD: Intent-Driven 3D Affordance Grounding in Functionally Competing Objects
por: Li, Jingliang, et al.
Publicado: (2026)
por: Li, Jingliang, et al.
Publicado: (2026)
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation
por: Tian, Tongxuan, et al.
Publicado: (2025)
por: Tian, Tongxuan, et al.
Publicado: (2025)
Affostruction: 3D Affordance Grounding with Generative Reconstruction
por: Park, Chunghyun, et al.
Publicado: (2026)
por: Park, Chunghyun, et al.
Publicado: (2026)
Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances
por: Wang, Qirui, et al.
Publicado: (2026)
por: Wang, Qirui, et al.
Publicado: (2026)
Interacted Object Grounding in Spatio-Temporal Human-Object Interactions
por: Liu, Xiaoyang, et al.
Publicado: (2024)
por: Liu, Xiaoyang, et al.
Publicado: (2024)
HAMMER: Harnessing MLLM via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding
por: Yao, Lei, et al.
Publicado: (2026)
por: Yao, Lei, et al.
Publicado: (2026)
InteractDiffusion: Interaction Control in Text-to-Image Diffusion Models
por: Hoe, Jiun Tian, et al.
Publicado: (2023)
por: Hoe, Jiun Tian, et al.
Publicado: (2023)
Part-Aware Open-Vocabulary 3D Affordance Grounding via Prototypical Semantic and Geometric Alignment
por: Gou, Dongqiang, et al.
Publicado: (2026)
por: Gou, Dongqiang, et al.
Publicado: (2026)
Populate-A-Scene: Affordance-Aware Human Video Generation
por: Shan, Mengyi, et al.
Publicado: (2025)
por: Shan, Mengyi, et al.
Publicado: (2025)
INTRA: Interaction Relationship-aware Weakly Supervised Affordance Grounding
por: Jang, Ji Ha, et al.
Publicado: (2024)
por: Jang, Ji Ha, et al.
Publicado: (2024)
Chirpy3D: Part-Aware Multi-View Diffusion for Creative Fine-Grained Object Generation
por: Ng, Kam Woh, et al.
Publicado: (2025)
por: Ng, Kam Woh, et al.
Publicado: (2025)
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
por: Kim, Hyeonwoo, et al.
Publicado: (2025)
por: Kim, Hyeonwoo, et al.
Publicado: (2025)
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
por: Shao, Yawen, et al.
Publicado: (2024)
por: Shao, Yawen, et al.
Publicado: (2024)
HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance
por: Li, Lei, et al.
Publicado: (2025)
por: Li, Lei, et al.
Publicado: (2025)
DAG: Unleash the Potential of Diffusion Model for Open-Vocabulary 3D Affordance Grounding
por: Wang, Hanqing, et al.
Publicado: (2025)
por: Wang, Hanqing, et al.
Publicado: (2025)
InteractMove: Text-Controlled Human-Object Interaction Generation in 3D Scenes with Movable Objects
por: Cai, Xinhao, et al.
Publicado: (2025)
por: Cai, Xinhao, et al.
Publicado: (2025)
DO3D: Self-supervised Learning of Decomposed Object-aware 3D Motion and Depth from Monocular Videos
por: Wu, Xiuzhe, et al.
Publicado: (2024)
por: Wu, Xiuzhe, et al.
Publicado: (2024)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
por: Jiang, Dengyang, et al.
Publicado: (2025)
por: Jiang, Dengyang, et al.
Publicado: (2025)
SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
por: Maillard, Léopold, et al.
Publicado: (2026)
por: Maillard, Léopold, et al.
Publicado: (2026)
Leverage Task Context for Object Affordance Ranking
por: Huang, Haojie, et al.
Publicado: (2024)
por: Huang, Haojie, et al.
Publicado: (2024)
Integrating Affordances and Attention models for Short-Term Object Interaction Anticipation
por: Labadia, Lorenzo Mur, et al.
Publicado: (2026)
por: Labadia, Lorenzo Mur, et al.
Publicado: (2026)
ArtHOI: Articulated Human-Object Interaction Synthesis by 4D Reconstruction from Video Priors
por: Huang, Zihao, et al.
Publicado: (2026)
por: Huang, Zihao, et al.
Publicado: (2026)
AffordanceLLM: Grounding Affordance from Vision Language Models
por: Qian, Shengyi, et al.
Publicado: (2024)
por: Qian, Shengyi, et al.
Publicado: (2024)
Towards Affordance-Aware Articulation Synthesis for Rigged Objects
por: Yu, Yu-Chu, et al.
Publicado: (2025)
por: Yu, Yu-Chu, et al.
Publicado: (2025)
Harnessing Object Grounding for Time-Sensitive Video Understanding
por: Wu, Tz-Ying, et al.
Publicado: (2025)
por: Wu, Tz-Ying, et al.
Publicado: (2025)
Ejemplares similares
-
VideoAfford: Grounding 3D Affordance from Human-Object-Interaction Videos via Multimodal Large Language Model
por: Wang, Hanqing, et al.
Publicado: (2026) -
OneHOI: Unifying Human-Object Interaction Generation and Editing
por: Hoe, Jiun Tian, et al.
Publicado: (2026) -
Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions
por: Zhu, He, et al.
Publicado: (2025) -
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
por: Hoe, Jiun Tian, et al.
Publicado: (2025) -
VAGNet: Vision-based Accident Anticipation with Global Features
por: Vipulananthan, Vipooshan, et al.
Publicado: (2026)