Grounding 3D Scene Affordance From Egocentric Interactions
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Cuiyu, Zhai, Wei, Yang, Yuhang, Luo, Hongchen, Liang, Sen, Cao, Yang, Zha, Zheng-Jun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
di: Shao, Yawen, et al.
Pubblicazione: (2024)
di: Shao, Yawen, et al.
Pubblicazione: (2024)
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
di: Yang, Yuhang, et al.
Pubblicazione: (2023)
di: Yang, Yuhang, et al.
Pubblicazione: (2023)
Visual-Geometric Collaborative Guidance for Affordance Learning
di: Luo, Hongchen, et al.
Pubblicazione: (2024)
di: Luo, Hongchen, et al.
Pubblicazione: (2024)
Leverage Task Context for Object Affordance Ranking
di: Huang, Haojie, et al.
Pubblicazione: (2024)
di: Huang, Haojie, et al.
Pubblicazione: (2024)
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
di: Yang, Yuhang, et al.
Pubblicazione: (2024)
di: Yang, Yuhang, et al.
Pubblicazione: (2024)
GRACE: Estimating Geometry-level 3D Human-Scene Contact from 2D Images
di: Wang, Chengfeng, et al.
Pubblicazione: (2025)
di: Wang, Chengfeng, et al.
Pubblicazione: (2025)
TOUCH: Text-guided Controllable Generation of Free-Form Hand-Object Interactions
di: Han, Guangyi, et al.
Pubblicazione: (2025)
di: Han, Guangyi, et al.
Pubblicazione: (2025)
End-to-End Spatial-Temporal Transformer for Real-time 4D HOI Reconstruction
di: Zhang, Haoyu, et al.
Pubblicazione: (2026)
di: Zhang, Haoyu, et al.
Pubblicazione: (2026)
EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning
di: Yu, Chengjun, et al.
Pubblicazione: (2026)
di: Yu, Chengjun, et al.
Pubblicazione: (2026)
One-Shot Affordance Grounding of Deformable Objects in Egocentric Organizing Scenes
di: Jia, Wanjun, et al.
Pubblicazione: (2025)
di: Jia, Wanjun, et al.
Pubblicazione: (2025)
Bidirectional Progressive Transformer for Interaction Intention Anticipation
di: Zhang, Zichen, et al.
Pubblicazione: (2024)
di: Zhang, Zichen, et al.
Pubblicazione: (2024)
Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances
di: Wang, Qirui, et al.
Pubblicazione: (2026)
di: Wang, Qirui, et al.
Pubblicazione: (2026)
PEAR: Phrase-Based Hand-Object Interaction Anticipation
di: Zhang, Zichen, et al.
Pubblicazione: (2024)
di: Zhang, Zichen, et al.
Pubblicazione: (2024)
HERO: Human Reaction Generation from Videos
di: Yu, Chengjun, et al.
Pubblicazione: (2025)
di: Yu, Chengjun, et al.
Pubblicazione: (2025)
SceneTeract: Agentic Functional Affordances and VLM Grounding in 3D Scenes
di: Maillard, Léopold, et al.
Pubblicazione: (2026)
di: Maillard, Léopold, et al.
Pubblicazione: (2026)
Intention-driven Ego-to-Exo Video Generation
di: Luo, Hongchen, et al.
Pubblicazione: (2024)
di: Luo, Hongchen, et al.
Pubblicazione: (2024)
EMoTive: Event-guided Trajectory Modeling for 3D Motion Estimation
di: Wan, Zengyu, et al.
Pubblicazione: (2025)
di: Wan, Zengyu, et al.
Pubblicazione: (2025)
VMAD: Visual-enhanced Multimodal Large Language Model for Zero-Shot Anomaly Detection
di: Deng, Huilin, et al.
Pubblicazione: (2024)
di: Deng, Huilin, et al.
Pubblicazione: (2024)
Closed-Loop Transfer for Weakly-supervised Affordance Grounding
di: Tang, Jiajin, et al.
Pubblicazione: (2025)
di: Tang, Jiajin, et al.
Pubblicazione: (2025)
3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians
di: Wei, Zeming, et al.
Pubblicazione: (2025)
di: Wei, Zeming, et al.
Pubblicazione: (2025)
SIGMAN:Scaling 3D Human Gaussian Generation with Millions of Assets
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
di: Mao, Aihua, et al.
Pubblicazione: (2026)
di: Mao, Aihua, et al.
Pubblicazione: (2026)
EF-3DGS: Event-Aided Free-Trajectory 3D Gaussian Splatting
di: Liao, Bohao, et al.
Pubblicazione: (2024)
di: Liao, Bohao, et al.
Pubblicazione: (2024)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
di: Gao, Xianqiang, et al.
Pubblicazione: (2024)
di: Gao, Xianqiang, et al.
Pubblicazione: (2024)
Benchmarking Large Vision-Language Models via Directed Scene Graph for Comprehensive Image Captioning
di: Lu, Fan, et al.
Pubblicazione: (2024)
di: Lu, Fan, et al.
Pubblicazione: (2024)
Gloria: Consistent Character Video Generation via Content Anchors
di: Yang, Yuhang, et al.
Pubblicazione: (2026)
di: Yang, Yuhang, et al.
Pubblicazione: (2026)
Event Stream Filtering via Probability Flux Estimation
di: Chen, Jinze, et al.
Pubblicazione: (2025)
di: Chen, Jinze, et al.
Pubblicazione: (2025)
ANNEXE: Unified Analyzing, Answering, and Pixel Grounding for Egocentric Interaction
di: Su, Yuejiao, et al.
Pubblicazione: (2025)
di: Su, Yuejiao, et al.
Pubblicazione: (2025)
R2G: Reasoning to Ground in 3D Scenes
di: Li, Yixuan, et al.
Pubblicazione: (2024)
di: Li, Yixuan, et al.
Pubblicazione: (2024)
FunRec: Reconstructing Functional 3D Scenes from Egocentric Interaction Videos
di: Delitzas, Alexandros, et al.
Pubblicazione: (2026)
di: Delitzas, Alexandros, et al.
Pubblicazione: (2026)
Affostruction: 3D Affordance Grounding with Generative Reconstruction
di: Park, Chunghyun, et al.
Pubblicazione: (2026)
di: Park, Chunghyun, et al.
Pubblicazione: (2026)
AffordanceSAM: Segment Anything Once More in Affordance Grounding
di: Jiang, Dengyang, et al.
Pubblicazione: (2025)
di: Jiang, Dengyang, et al.
Pubblicazione: (2025)
VideoGen-Eval: Agent-based System for Video Generation Evaluation
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
di: Yang, Yuhang, et al.
Pubblicazione: (2025)
PhyScene3D: Physically Consistent Interactive 3D Tabletop Scene Generation
di: Chen, Weixing, et al.
Pubblicazione: (2026)
di: Chen, Weixing, et al.
Pubblicazione: (2026)
Move as You Say, Interact as You Can: Language-guided Human Motion Generation with Scene Affordance
di: Wang, Zan, et al.
Pubblicazione: (2024)
di: Wang, Zan, et al.
Pubblicazione: (2024)
Event-based Visual Deformation Measurement
di: Wu, Yuliang, et al.
Pubblicazione: (2026)
di: Wu, Yuliang, et al.
Pubblicazione: (2026)
ViewPoint: Panoramic Video Generation with Pretrained Diffusion Models
di: Fang, Zixun, et al.
Pubblicazione: (2025)
di: Fang, Zixun, et al.
Pubblicazione: (2025)
MMAR: Towards Lossless Multi-Modal Auto-Regressive Probabilistic Modeling
di: Yang, Jian, et al.
Pubblicazione: (2024)
di: Yang, Jian, et al.
Pubblicazione: (2024)
MATE: Motion-Augmented Temporal Consistency for Event-based Point Tracking
di: Han, Han, et al.
Pubblicazione: (2024)
di: Han, Han, et al.
Pubblicazione: (2024)
Unlocking 3D Affordance Segmentation with 2D Semantic Knowledge
di: Huang, Yu, et al.
Pubblicazione: (2025)
di: Huang, Yu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
di: Shao, Yawen, et al.
Pubblicazione: (2024) -
LEMON: Learning 3D Human-Object Interaction Relation from 2D Images
di: Yang, Yuhang, et al.
Pubblicazione: (2023) -
Visual-Geometric Collaborative Guidance for Affordance Learning
di: Luo, Hongchen, et al.
Pubblicazione: (2024) -
Leverage Task Context for Object Affordance Ranking
di: Huang, Haojie, et al.
Pubblicazione: (2024) -
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
di: Yang, Yuhang, et al.
Pubblicazione: (2024)