Target-Aware Video Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Kim, Taeksoo, Joo, Hanbyul |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dexterous World Models
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
GALA: Generating Animatable Layered Assets from a Single Scan
by: Kim, Taeksoo, et al.
Published: (2024)
by: Kim, Taeksoo, et al.
Published: (2024)
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
by: Kim, Hyeonwoo, et al.
Published: (2025)
by: Kim, Hyeonwoo, et al.
Published: (2025)
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
by: Baik, Sangwon, et al.
Published: (2025)
by: Baik, Sangwon, et al.
Published: (2025)
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
by: Kim, Hyeonwoo, et al.
Published: (2026)
by: Kim, Hyeonwoo, et al.
Published: (2026)
Beyond the Contact: Discovering Comprehensive Affordance for 3D Objects from Pre-trained 2D Diffusion Models
by: Kim, Hyeonwoo, et al.
Published: (2024)
by: Kim, Hyeonwoo, et al.
Published: (2024)
GraspDiffusion: Synthesizing Realistic Whole-body Hand-Object Interaction
by: Kwon, Patrick, et al.
Published: (2024)
by: Kwon, Patrick, et al.
Published: (2024)
OmniEgoCap: Camera-Agnostic Sequence-Level Egocentric Motion Reconstruction
by: Cho, Kyungwon, et al.
Published: (2025)
by: Cho, Kyungwon, et al.
Published: (2025)
Durian: Dual Reference Image-Guided Portrait Animation with Attribute Transfer
by: Cha, Hyunsoo, et al.
Published: (2025)
by: Cha, Hyunsoo, et al.
Published: (2025)
Guess The Unseen: Dynamic 3D Scene Reconstruction from Partial 2D Glimpses
by: Lee, Inhee, et al.
Published: (2024)
by: Lee, Inhee, et al.
Published: (2024)
PEGASUS: Personalized Generative 3D Avatars with Composable Attributes
by: Cha, Hyunsoo, et al.
Published: (2024)
by: Cha, Hyunsoo, et al.
Published: (2024)
ParaHome: Parameterizing Everyday Home Activities Towards 3D Generative Modeling of Human-Object Interactions
by: Kim, Jeonghwan, et al.
Published: (2024)
by: Kim, Jeonghwan, et al.
Published: (2024)
Mocap Everyone Everywhere: Lightweight Motion Capture With Smartwatches and a Head-Mounted Camera
by: Lee, Jiye, et al.
Published: (2024)
by: Lee, Jiye, et al.
Published: (2024)
Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents
by: Baik, Sangwon, et al.
Published: (2026)
by: Baik, Sangwon, et al.
Published: (2026)
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision
by: Cha, Hyunsoo, et al.
Published: (2026)
by: Cha, Hyunsoo, et al.
Published: (2026)
PERSE: Personalized 3D Generative Avatars from A Single Portrait
by: Cha, Hyunsoo, et al.
Published: (2024)
by: Cha, Hyunsoo, et al.
Published: (2024)
ORBIS: Output-Guided Token Reduction with Distribution-Aware Matching for Video Diffusion Acceleration
by: Lee, Hangyeol, et al.
Published: (2026)
by: Lee, Hangyeol, et al.
Published: (2026)
Rethinking Token Reduction for Diffusion Models via Output-Similarity-Awareness
by: Lee, Hangyeol, et al.
Published: (2026)
by: Lee, Hangyeol, et al.
Published: (2026)
Few-Shot Class-Incremental Model Attribution Using Learnable Representation From CLIP-ViT Features
by: Lee, Hanbyul, et al.
Published: (2025)
by: Lee, Hanbyul, et al.
Published: (2025)
Learning to Generate Human-Human-Object Interactions from Textual Descriptions
by: Na, Jeonghyeon, et al.
Published: (2025)
by: Na, Jeonghyeon, et al.
Published: (2025)
HairCUP: Hair Compositional Universal Prior for 3D Gaussian Avatars
by: Kim, Byungjun, et al.
Published: (2025)
by: Kim, Byungjun, et al.
Published: (2025)
HRDexDB: A Large-Scale Dataset of Dexterous Human and Robotic Hand Grasps
by: Lim, Jongbin, et al.
Published: (2026)
by: Lim, Jongbin, et al.
Published: (2026)
GenVideo: One-shot Target-image and Shape Aware Video Editing using T2I Diffusion Models
by: Harsha, Sai Sree, et al.
Published: (2024)
by: Harsha, Sai Sree, et al.
Published: (2024)
SUPER Decoder Block for Reconstruction-Aware U-Net Variants
by: Joo, Siheon, et al.
Published: (2025)
by: Joo, Siheon, et al.
Published: (2025)
OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction
by: Lee, Junyoung, et al.
Published: (2026)
by: Lee, Junyoung, et al.
Published: (2026)
Latent Space Super-Resolution for Higher-Resolution Image Generation with Diffusion Models
by: Jeong, Jinho, et al.
Published: (2025)
by: Jeong, Jinho, et al.
Published: (2025)
Moaw: Unleashing Motion Awareness for Video Diffusion Models
by: Zhang, Tianqi, et al.
Published: (2026)
by: Zhang, Tianqi, et al.
Published: (2026)
Zero-Shot Video Deraining with Video Diffusion Models
by: Varanka, Tuomas, et al.
Published: (2025)
by: Varanka, Tuomas, et al.
Published: (2025)
Grid Diffusion Models for Text-to-Video Generation
by: Lee, Taegyeong, et al.
Published: (2024)
by: Lee, Taegyeong, et al.
Published: (2024)
Adjusting Initial Noise to Mitigate Memorization in Text-to-Image Diffusion Models
by: Han, Hyeonggeun, et al.
Published: (2025)
by: Han, Hyeonggeun, et al.
Published: (2025)
HieraSurg: Hierarchy-Aware Diffusion Model for Surgical Video Generation
by: Biagini, Diego, et al.
Published: (2025)
by: Biagini, Diego, et al.
Published: (2025)
Open-ended Hierarchical Streaming Video Understanding with Vision Language Models
by: Kang, Hyolim, et al.
Published: (2025)
by: Kang, Hyolim, et al.
Published: (2025)
FastSTAR: Spatiotemporal Token Pruning for Efficient Autoregressive Video Synthesis
by: Yune, Sungwoong, et al.
Published: (2026)
by: Yune, Sungwoong, et al.
Published: (2026)
Image Diffusion Models Exhibit Emergent Temporal Propagation in Videos
by: Kim, Youngseo, et al.
Published: (2025)
by: Kim, Youngseo, et al.
Published: (2025)
RAD: Region-Aware Diffusion Models for Image Inpainting
by: Kim, Sora, et al.
Published: (2024)
by: Kim, Sora, et al.
Published: (2024)
IF-MDM: Implicit Face Motion Diffusion Model for High-Fidelity Realtime Talking Head Generation
by: Yang, Sejong, et al.
Published: (2024)
by: Yang, Sejong, et al.
Published: (2024)
Autoregressive Universal Video Segmentation Model
by: Heo, Miran, et al.
Published: (2025)
by: Heo, Miran, et al.
Published: (2025)
Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers
by: Zhu, Jingyuan, et al.
Published: (2026)
by: Zhu, Jingyuan, et al.
Published: (2026)
Hierarchically Structured Neural Bones for Reconstructing Animatable Objects from Casual Videos
by: Jeon, Subin, et al.
Published: (2024)
by: Jeon, Subin, et al.
Published: (2024)
Spatial Degradation-Aware and Temporal Consistent Diffusion Model for Compressed Video Super-Resolution
by: An, Hongyu, et al.
Published: (2025)
by: An, Hongyu, et al.
Published: (2025)
Similar Items
-
Dexterous World Models
by: Kim, Byungjun, et al.
Published: (2025) -
GALA: Generating Animatable Layered Assets from a Single Scan
by: Kim, Taeksoo, et al.
Published: (2024) -
DAViD: Modeling Dynamic Affordance of 3D Objects Using Pre-trained Video Diffusion Models
by: Kim, Hyeonwoo, et al.
Published: (2025) -
Learning 3D Object Spatial Relationships from Pre-trained 2D Diffusion Models
by: Baik, Sangwon, et al.
Published: (2025) -
DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation
by: Kim, Hyeonwoo, et al.
Published: (2026)