Saved in:
| Main Authors: | Baik, Sangwon, Kim, Gunhee, Choi, Mingi, Joo, Hanbyul |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.09781 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
by: Baik, Sangwon, et al.
Published: (2025)
by: Baik, Sangwon, et al.
Published: (2025)
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
by: Kim, Hyeonwoo, et al.
Published: (2025)
by: Kim, Hyeonwoo, et al.
Published: (2025)
Learning to Generate Human-Human-Object Interactions from Textual Descriptions
by: Na, Jeonghyeon, et al.
Published: (2025)
by: Na, Jeonghyeon, et al.
Published: (2025)
HRDexDB: A Large-Scale Dataset of Dexterous Human and Robotic Hand Grasps
by: Lim, Jongbin, et al.
Published: (2026)
by: Lim, Jongbin, et al.
Published: (2026)
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
by: Kim, Hyeonwoo, et al.
Published: (2026)
by: Kim, Hyeonwoo, et al.
Published: (2026)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer
by: Cha, Hyunsoo, et al.
Published: (2025)
by: Cha, Hyunsoo, et al.
Published: (2025)
Beyond the Contact: Discovering Comprehensive Affordance for 3D Objects from Pre-trained 2D Diffusion Models
by: Kim, Hyeonwoo, et al.
Published: (2024)
by: Kim, Hyeonwoo, et al.
Published: (2024)
Target-Aware Video Diffusion Models
by: Kim, Taeksoo, et al.
Published: (2025)
by: Kim, Taeksoo, et al.
Published: (2025)
OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction
by: Lee, Junyoung, et al.
Published: (2026)
by: Lee, Junyoung, et al.
Published: (2026)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
by: Lee, Inhee, et al.
Published: (2024)
by: Lee, Inhee, et al.
Published: (2024)
PEGASUS: Personalized Generative 3D Avatars with Composable Attributes
by: Cha, Hyunsoo, et al.
Published: (2024)
by: Cha, Hyunsoo, et al.
Published: (2024)
GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction
by: Kwon, Patrick, et al.
Published: (2024)
by: Kwon, Patrick, et al.
Published: (2024)
OmniEgoCap: Camera-Agnostic Sequence-Level Egocentric Motion Reconstruction
by: Cho, Kyungwon, et al.
Published: (2025)
by: Cho, Kyungwon, et al.
Published: (2025)
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision
by: Cha, Hyunsoo, et al.
Published: (2026)
by: Cha, Hyunsoo, et al.
Published: (2026)
PERSE: Personalized 3D Generative Avatars from A Single Portrait
by: Cha, Hyunsoo, et al.
Published: (2024)
by: Cha, Hyunsoo, et al.
Published: (2024)
Mocap Everyone Everywhere: Lightweight Motion Capture With Smartwatches and a Head-Mounted Camera
by: Lee, Jiye, et al.
Published: (2024)
by: Lee, Jiye, et al.
Published: (2024)
Dexterous World Models
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
GALA: Generating Animatable Layered Assets from a Single Scan
by: Kim, Taeksoo, et al.
Published: (2024)
by: Kim, Taeksoo, et al.
Published: (2024)
VLM6D: VLM based 6Dof Pose Estimation based on RGB-D Images
by: Sarowar, Md Selim, et al.
Published: (2025)
by: Sarowar, Md Selim, et al.
Published: (2025)
MASH-VLM: Mitigating Action-Scene Hallucination in Video-LLMs through Disentangled Spatial-Temporal Representations
by: Bae, Kyungho, et al.
Published: (2025)
by: Bae, Kyungho, et al.
Published: (2025)
LoRA-Loop: Closing the Synthetic Replay Cycle for Continual VLM Learning
by: Wang, Kaihong, et al.
Published: (2025)
by: Wang, Kaihong, et al.
Published: (2025)
TIE-KD: Teacher-Independent and Explainable Knowledge Distillation for Monocular Depth Estimation
by: Choi, Sangwon, et al.
Published: (2024)
by: Choi, Sangwon, et al.
Published: (2024)
HairCUP: Hair Compositional Universal Prior for 3D Gaussian Avatars
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
Mask6D: Masked Pose Priors For 6D Object Pose Estimation
by: Xie, Yuechen, et al.
Published: (2025)
by: Xie, Yuechen, et al.
Published: (2025)
Dynamic VLM-Guided Negative Prompting for Diffusion Models
by: Chang, Hoyeon, et al.
Published: (2025)
by: Chang, Hoyeon, et al.
Published: (2025)
THE-Pose: Topological Prior with Hybrid Graph Fusion for Estimating Category-Level 6D Object Pose
by: Lee, Eunho, et al.
Published: (2025)
by: Lee, Eunho, et al.
Published: (2025)
RefPose: Leveraging Reference Geometric Correspondences for Accurate 6D Pose Estimation of Unseen Objects
by: Kim, Jaeguk, et al.
Published: (2025)
by: Kim, Jaeguk, et al.
Published: (2025)
Event6D: Event-based Novel Object 6D Pose Tracking
by: Kang, Jae-Young, et al.
Published: (2026)
by: Kang, Jae-Young, et al.
Published: (2026)
GMFlow: Global Motion-Guided Recurrent Flow for 6D Object Pose Estimation
by: Liu, Xin, et al.
Published: (2024)
by: Liu, Xin, et al.
Published: (2024)
Pose-Guided Residual Refinement for Interpretable Text-to-Motion Generation and Editing
by: Jeong, Sukhyun, et al.
Published: (2025)
by: Jeong, Sukhyun, et al.
Published: (2025)
Gaussian Blending: Rethinking Alpha Blending in 3D Gaussian Splatting
by: Koo, Junseo, et al.
Published: (2025)
by: Koo, Junseo, et al.
Published: (2025)
SmartAvatar: Text- and Image-Guided Human Avatar Generation with VLM AI Agents
by: Huang-Menders, Alexander, et al.
Published: (2025)
by: Huang-Menders, Alexander, et al.
Published: (2025)
How to Move Your Dragon: Text-to-Motion Synthesis for Large-Vocabulary Objects
by: Lee, Wonkwang, et al.
Published: (2025)
by: Lee, Wonkwang, et al.
Published: (2025)
Leveraging Image Augmentation for Object Manipulation: Towards Interpretable Controllability in Object-Centric Learning
by: Kim, Jinwoo, et al.
Published: (2023)
by: Kim, Jinwoo, et al.
Published: (2023)
Human-Object Interaction via Automatically Designed VLM-Guided Motion Policy
by: Deng, Zekai, et al.
Published: (2025)
by: Deng, Zekai, et al.
Published: (2025)
Parallel qMRI Reconstruction from 4x Accelerated Acquisitions
by: Kang, Mingi
Published: (2025)
by: Kang, Mingi
Published: (2025)
GS2Pose: Two-stage 6D Object Pose Estimation Guided by Gaussian Splatting
by: Mei, Jilan, et al.
Published: (2024)
by: Mei, Jilan, et al.
Published: (2024)
TransPose: 6D Object Pose Estimation with Geometry-Aware Transformer
by: Lin, Xiao, et al.
Published: (2023)
by: Lin, Xiao, et al.
Published: (2023)
AgentSteerTTS: A Multi-Agent Closed-Loop Framework for Composite-Instruction Text-to-Speech
by: Kang, Bin, et al.
Published: (2026)
by: Kang, Bin, et al.
Published: (2026)
Similar Items
-
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
by: Baik, Sangwon, et al.
Published: (2025) -
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
by: Kim, Hyeonwoo, et al.
Published: (2025) -
Learning to Generate Human-Human-Object Interactions from Textual Descriptions
by: Na, Jeonghyeon, et al.
Published: (2025) -
HRDexDB: A Large-Scale Dataset of Dexterous Human and Robotic Hand Grasps
by: Lim, Jongbin, et al.
Published: (2026) -
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
by: Kim, Hyeonwoo, et al.
Published: (2026)