DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Hyeonwoo, Baik, Sangwon, Joo, Hanbyul |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
di: Baik, Sangwon, et al.
Pubblicazione: (2025)
di: Baik, Sangwon, et al.
Pubblicazione: (2025)
Beyond the Contact: Discovering Comprehensive Affordance for 3D Objects from Pre-trained 2D Diffusion Models
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2024)
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2024)
Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents
di: Baik, Sangwon, et al.
Pubblicazione: (2026)
di: Baik, Sangwon, et al.
Pubblicazione: (2026)
DAViD: Data-efficient and Accurate Vision Models from Synthetic Data
di: Saleh, Fatemeh, et al.
Pubblicazione: (2025)
di: Saleh, Fatemeh, et al.
Pubblicazione: (2025)
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2026)
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2026)
Target-Aware Video Diffusion Models
di: Kim, Taeksoo, et al.
Pubblicazione: (2025)
di: Kim, Taeksoo, et al.
Pubblicazione: (2025)
Learning to Generate Human-Human-Object Interactions from Textual Descriptions
di: Na, Jeonghyeon, et al.
Pubblicazione: (2025)
di: Na, Jeonghyeon, et al.
Pubblicazione: (2025)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
di: Kim, Jeonghwan, et al.
Pubblicazione: (2024)
di: Kim, Jeonghwan, et al.
Pubblicazione: (2024)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
di: Lee, Inhee, et al.
Pubblicazione: (2024)
di: Lee, Inhee, et al.
Pubblicazione: (2024)
PEGASUS: Personalized Generative 3D Avatars with Composable Attributes
di: Cha, Hyunsoo, et al.
Pubblicazione: (2024)
di: Cha, Hyunsoo, et al.
Pubblicazione: (2024)
GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction
di: Kwon, Patrick, et al.
Pubblicazione: (2024)
di: Kwon, Patrick, et al.
Pubblicazione: (2024)
PERSE: Personalized 3D Generative Avatars from A Single Portrait
di: Cha, Hyunsoo, et al.
Pubblicazione: (2024)
di: Cha, Hyunsoo, et al.
Pubblicazione: (2024)
Dexterous World Models
di: Kim, Byungjun, et al.
Pubblicazione: (2025)
di: Kim, Byungjun, et al.
Pubblicazione: (2025)
Exploring Pre-trained Text-to-Video Diffusion Models for Referring Video Object Segmentation
di: Zhu, Zixin, et al.
Pubblicazione: (2024)
di: Zhu, Zixin, et al.
Pubblicazione: (2024)
OmniEgoCap: Camera-Agnostic Sequence-Level Egocentric Motion Reconstruction
di: Cho, Kyungwon, et al.
Pubblicazione: (2025)
di: Cho, Kyungwon, et al.
Pubblicazione: (2025)
Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer
di: Cha, Hyunsoo, et al.
Pubblicazione: (2025)
di: Cha, Hyunsoo, et al.
Pubblicazione: (2025)
VAGNet: Grounding 3D Affordance from Human-Object Interactions in Videos
di: Mao, Aihua, et al.
Pubblicazione: (2026)
di: Mao, Aihua, et al.
Pubblicazione: (2026)
VideoAfford: Grounding 3D Affordance from Human-Object-Interaction Videos via Multimodal Large Language Model
di: Wang, Hanqing, et al.
Pubblicazione: (2026)
di: Wang, Hanqing, et al.
Pubblicazione: (2026)
H2OFlow: Grounding Human-Object Affordances with 3D Generative Models and Dense Diffused Flows
di: Zhang, Harry, et al.
Pubblicazione: (2025)
di: Zhang, Harry, et al.
Pubblicazione: (2025)
O$^2$-Recon: Completing 3D Reconstruction of Occluded Objects in the Scene with a Pre-trained 2D Diffusion Model
di: Hu, Yubin, et al.
Pubblicazione: (2023)
di: Hu, Yubin, et al.
Pubblicazione: (2023)
Mocap Everyone Everywhere: Lightweight Motion Capture With Smartwatches and a Head-Mounted Camera
di: Lee, Jiye, et al.
Pubblicazione: (2024)
di: Lee, Jiye, et al.
Pubblicazione: (2024)
HairCUP: Hair Compositional Universal Prior for 3D Gaussian Avatars
di: Kim, Byungjun, et al.
Pubblicazione: (2025)
di: Kim, Byungjun, et al.
Pubblicazione: (2025)
GALA: Generating Animatable Layered Assets from a Single Scan
di: Kim, Taeksoo, et al.
Pubblicazione: (2024)
di: Kim, Taeksoo, et al.
Pubblicazione: (2024)
Temporal-Consistent Video Restoration with Pre-trained Diffusion Models
di: Wang, Hengkang, et al.
Pubblicazione: (2025)
di: Wang, Hengkang, et al.
Pubblicazione: (2025)
DAG: Unleash the Potential of Diffusion Model for Open-Vocabulary 3D Affordance Grounding
di: Wang, Hanqing, et al.
Pubblicazione: (2025)
di: Wang, Hanqing, et al.
Pubblicazione: (2025)
Repurposing Pre-trained Video Diffusion Models for Event-based Video Interpolation
di: Chen, Jingxi, et al.
Pubblicazione: (2024)
di: Chen, Jingxi, et al.
Pubblicazione: (2024)
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models
di: Cho, In, et al.
Pubblicazione: (2025)
di: Cho, In, et al.
Pubblicazione: (2025)
Few-Shot Class-Incremental Model Attribution Using Learnable Representation From CLIP-ViT Features
di: Lee, Hanbyul, et al.
Pubblicazione: (2025)
di: Lee, Hanbyul, et al.
Pubblicazione: (2025)
Hybrid Video Diffusion Models with 2D Triplane and 3D Wavelet Representation
di: Kim, Kihong, et al.
Pubblicazione: (2024)
di: Kim, Kihong, et al.
Pubblicazione: (2024)
PropFly: Learning to Propagate via On-the-Fly Supervision from Pre-trained Video Diffusion Models
di: Seo, Wonyong, et al.
Pubblicazione: (2026)
di: Seo, Wonyong, et al.
Pubblicazione: (2026)
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision
di: Cha, Hyunsoo, et al.
Pubblicazione: (2026)
di: Cha, Hyunsoo, et al.
Pubblicazione: (2026)
IAAO: Interactive Affordance Learning for Articulated Objects in 3D Environments
di: Zhang, Can, et al.
Pubblicazione: (2025)
di: Zhang, Can, et al.
Pubblicazione: (2025)
Affordance-Guided Diffusion Prior for 3D Hand Reconstruction
di: Suzuki, Naru, et al.
Pubblicazione: (2025)
di: Suzuki, Naru, et al.
Pubblicazione: (2025)
Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
di: Gao, Xianqiang, et al.
Pubblicazione: (2024)
di: Gao, Xianqiang, et al.
Pubblicazione: (2024)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
di: Chu, Hengshuo, et al.
Pubblicazione: (2025)
di: Chu, Hengshuo, et al.
Pubblicazione: (2025)
PatchContrast: Self-Supervised Pre-training for 3D Object Detection
di: Shrout, Oren, et al.
Pubblicazione: (2023)
di: Shrout, Oren, et al.
Pubblicazione: (2023)
Training-Free Semantic Video Composition via Pre-trained Diffusion Model
di: Guo, Jiaqi, et al.
Pubblicazione: (2024)
di: Guo, Jiaqi, et al.
Pubblicazione: (2024)
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
di: Oh, Youngmin, et al.
Pubblicazione: (2024)
Affordance-Aware Object Insertion via Mask-Aware Dual Diffusion
di: He, Jixuan, et al.
Pubblicazione: (2024)
di: He, Jixuan, et al.
Pubblicazione: (2024)
GO-Renderer: Generative Object Rendering with 3D-aware Controllable Video Diffusion Models
di: Gu, Zekai, et al.
Pubblicazione: (2026)
di: Gu, Zekai, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
di: Baik, Sangwon, et al.
Pubblicazione: (2025) -
Beyond the Contact: Discovering Comprehensive Affordance for 3D Objects from Pre-trained 2D Diffusion Models
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2024) -
Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents
di: Baik, Sangwon, et al.
Pubblicazione: (2026) -
DAViD: Data-efficient and Accurate Vision Models from Synthetic Data
di: Saleh, Fatemeh, et al.
Pubblicazione: (2025) -
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
di: Kim, Hyeonwoo, et al.
Pubblicazione: (2026)