Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Baik, Sangwon, Kim, Hyeonwoo, Joo, Hanbyul |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2025)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2025)
Beyond the Contact: Discovering Comprehensive Affordance for 3D Objects from Pre-trained 2D Diffusion Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024)
Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents
von: Baik, Sangwon, et al.
Veröffentlicht: (2026)
von: Baik, Sangwon, et al.
Veröffentlicht: (2026)
Learning to Generate Human-Human-Object Interactions from Textual Descriptions
von: Na, Jeonghyeon, et al.
Veröffentlicht: (2025)
von: Na, Jeonghyeon, et al.
Veröffentlicht: (2025)
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2026)
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2026)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
von: Lee, Inhee, et al.
Veröffentlicht: (2024)
von: Lee, Inhee, et al.
Veröffentlicht: (2024)
Target-Aware Video Diffusion Models
von: Kim, Taeksoo, et al.
Veröffentlicht: (2025)
von: Kim, Taeksoo, et al.
Veröffentlicht: (2025)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2024)
von: Kim, Jeonghwan, et al.
Veröffentlicht: (2024)
PEGASUS: Personalized Generative 3D Avatars with Composable Attributes
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2024)
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2024)
GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction
von: Kwon, Patrick, et al.
Veröffentlicht: (2024)
von: Kwon, Patrick, et al.
Veröffentlicht: (2024)
PERSE: Personalized 3D Generative Avatars from A Single Portrait
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2024)
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2024)
Dexterous World Models
von: Kim, Byungjun, et al.
Veröffentlicht: (2025)
von: Kim, Byungjun, et al.
Veröffentlicht: (2025)
OmniEgoCap: Camera-Agnostic Sequence-Level Egocentric Motion Reconstruction
von: Cho, Kyungwon, et al.
Veröffentlicht: (2025)
von: Cho, Kyungwon, et al.
Veröffentlicht: (2025)
Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2025)
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2025)
O$^2$-Recon: Completing 3D Reconstruction of Occluded Objects in the Scene with a Pre-trained 2D Diffusion Model
von: Hu, Yubin, et al.
Veröffentlicht: (2023)
von: Hu, Yubin, et al.
Veröffentlicht: (2023)
GALA: Generating Animatable Layered Assets from a Single Scan
von: Kim, Taeksoo, et al.
Veröffentlicht: (2024)
von: Kim, Taeksoo, et al.
Veröffentlicht: (2024)
Mocap Everyone Everywhere: Lightweight Motion Capture With Smartwatches and a Head-Mounted Camera
von: Lee, Jiye, et al.
Veröffentlicht: (2024)
von: Lee, Jiye, et al.
Veröffentlicht: (2024)
HairCUP: Hair Compositional Universal Prior for 3D Gaussian Avatars
von: Kim, Byungjun, et al.
Veröffentlicht: (2025)
von: Kim, Byungjun, et al.
Veröffentlicht: (2025)
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2026)
von: Cha, Hyunsoo, et al.
Veröffentlicht: (2026)
PatchContrast: Self-Supervised Pre-training for 3D Object Detection
von: Shrout, Oren, et al.
Veröffentlicht: (2023)
von: Shrout, Oren, et al.
Veröffentlicht: (2023)
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models
von: Cho, In, et al.
Veröffentlicht: (2025)
von: Cho, In, et al.
Veröffentlicht: (2025)
E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training
von: Zhao, Qitao, et al.
Veröffentlicht: (2025)
von: Zhao, Qitao, et al.
Veröffentlicht: (2025)
Unsupervised Monocular 3D Keypoint Discovery from Multi-View Diffusion Priors
von: Jeon, Subin, et al.
Veröffentlicht: (2025)
von: Jeon, Subin, et al.
Veröffentlicht: (2025)
PropFly: Learning to Propagate via On-the-Fly Supervision from Pre-trained Video Diffusion Models
von: Seo, Wonyong, et al.
Veröffentlicht: (2026)
von: Seo, Wonyong, et al.
Veröffentlicht: (2026)
Solving 3D Inverse Problems using Pre-trained 2D Diffusion Models
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
von: Chung, Hyungjin, et al.
Veröffentlicht: (2022)
GAP-MLLM: Geometry-Aligned Pre-training for Activating 3D Spatial Perception in Multimodal Large Language Models
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
von: Zhang, Jiaxin, et al.
Veröffentlicht: (2026)
B2N3D: Progressive Learning from Binary to N-ary Relationships for 3D Object Grounding
von: Xiao, Feng, et al.
Veröffentlicht: (2025)
von: Xiao, Feng, et al.
Veröffentlicht: (2025)
4D Visual Pre-training for Robot Learning
von: Hou, Chengkai, et al.
Veröffentlicht: (2025)
von: Hou, Chengkai, et al.
Veröffentlicht: (2025)
Exploring Pre-trained Text-to-Video Diffusion Models for Referring Video Object Segmentation
von: Zhu, Zixin, et al.
Veröffentlicht: (2024)
von: Zhu, Zixin, et al.
Veröffentlicht: (2024)
MonoWAD: Weather-Adaptive Diffusion Model for Robust Monocular 3D Object Detection
von: Oh, Youngmin, et al.
Veröffentlicht: (2024)
von: Oh, Youngmin, et al.
Veröffentlicht: (2024)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
von: Xu, Wan, et al.
Veröffentlicht: (2023)
von: Xu, Wan, et al.
Veröffentlicht: (2023)
MaskHOI: Robust 3D Hand-Object Interaction Estimation via Masked Pre-training
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
von: Xie, Yuechen, et al.
Veröffentlicht: (2025)
Pre-training with 3D Synthetic Data: Learning 3D Point Cloud Instance Segmentation from 3D Synthetic Scenes
von: Otsuka, Daichi, et al.
Veröffentlicht: (2025)
von: Otsuka, Daichi, et al.
Veröffentlicht: (2025)
Boosting Zero-Shot 3D Style Transfer with 2D Pre-trained Priors
von: Dong, Xin, et al.
Veröffentlicht: (2026)
von: Dong, Xin, et al.
Veröffentlicht: (2026)
SPOT: Scalable 3D Pre-training via Occupancy Prediction for Learning Transferable 3D Representations
von: Yan, Xiangchao, et al.
Veröffentlicht: (2023)
von: Yan, Xiangchao, et al.
Veröffentlicht: (2023)
ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
von: Kim, Minsu, et al.
Veröffentlicht: (2025)
Let 2D Diffusion Model Know 3D-Consistency for Robust Text-to-3D Generation
von: Seo, Junyoung, et al.
Veröffentlicht: (2023)
von: Seo, Junyoung, et al.
Veröffentlicht: (2023)
Robust Fine-tuning for Pre-trained 3D Point Cloud Models
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024)
von: Zhang, Zhibo, et al.
Veröffentlicht: (2024)
Improving 2D-3D Dense Correspondences with Diffusion Models for 6D Object Pose Estimation
von: Hönig, Peter, et al.
Veröffentlicht: (2024)
von: Hönig, Peter, et al.
Veröffentlicht: (2024)
ULIP-2: Towards Scalable Multimodal Pre-training for 3D Understanding
von: Xue, Le, et al.
Veröffentlicht: (2023)
von: Xue, Le, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2025) -
Beyond the Contact: Discovering Comprehensive Affordance for 3D Objects from Pre-trained 2D Diffusion Models
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2024) -
Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents
von: Baik, Sangwon, et al.
Veröffentlicht: (2026) -
Learning to Generate Human-Human-Object Interactions from Textual Descriptions
von: Na, Jeonghyeon, et al.
Veröffentlicht: (2025) -
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
von: Kim, Hyeonwoo, et al.
Veröffentlicht: (2026)