InterDyn: Controllable Interactive Dynamics with Video Diffusion Models
Fuente:
arXiv
Saved in:
| Main Authors: | Akkerman, Rick, Feng, Haiwen, Black, Michael J., Tzionas, Dimitrios, Abrevaya, Victoria Fernández |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GenLit: Reformulating Single-Image Relighting as Video Generation
by: Bharadwaj, Shrisha, et al.
Published: (2024)
by: Bharadwaj, Shrisha, et al.
Published: (2024)
Re-Thinking Inverse Graphics With Large Language Models
by: Kulits, Peter, et al.
Published: (2024)
by: Kulits, Peter, et al.
Published: (2024)
Explorative Inbetweening of Time and Space
by: Feng, Haiwen, et al.
Published: (2024)
by: Feng, Haiwen, et al.
Published: (2024)
OFER: Occluded Face Expression Reconstruction
by: Selvaraju, Pratheba, et al.
Published: (2024)
by: Selvaraju, Pratheba, et al.
Published: (2024)
InteractVLM: 3D Interaction Reasoning from 2D Foundational Models
by: Dwivedi, Sai Kumar, et al.
Published: (2025)
by: Dwivedi, Sai Kumar, et al.
Published: (2025)
GRIP: Generating Interaction Poses Using Spatial Cues and Latent Consistency
by: Taheri, Omid, et al.
Published: (2023)
by: Taheri, Omid, et al.
Published: (2023)
PuzzleAvatar: Assembling 3D Avatars from Personal Albums
by: Xiu, Yuliang, et al.
Published: (2024)
by: Xiu, Yuliang, et al.
Published: (2024)
RHINO: Reconstructing Human Interactions with Novel Objects from Monocular Videos
by: Xue, Lixin, et al.
Published: (2026)
by: Xue, Lixin, et al.
Published: (2026)
Toward Human Understanding with Controllable Synthesis
by: Cuevas-Velasquez, Hanz, et al.
Published: (2024)
by: Cuevas-Velasquez, Hanz, et al.
Published: (2024)
HOI-Dyn: Learning Interaction Dynamics for Human-Object Motion Diffusion
by: Wu, Lin, et al.
Published: (2025)
by: Wu, Lin, et al.
Published: (2025)
Predicting 4D Hand Trajectory from Monocular Videos
by: Ye, Yufei, et al.
Published: (2025)
by: Ye, Yufei, et al.
Published: (2025)
PICO: Reconstructing 3D People In Contact with Objects
by: Cseke, Alpár, et al.
Published: (2025)
by: Cseke, Alpár, et al.
Published: (2025)
3D Whole-body Grasp Synthesis with Directional Controllability
by: Paschalidis, Georgios, et al.
Published: (2024)
by: Paschalidis, Georgios, et al.
Published: (2024)
DynVFX: Augmenting Real Videos with Dynamic Content
by: Yatim, Danah, et al.
Published: (2025)
by: Yatim, Danah, et al.
Published: (2025)
LEXIS: LatEnt ProXimal Interaction Signatures for 3D HOI from an Image
by: Antić, Dimitrije, et al.
Published: (2026)
by: Antić, Dimitrije, et al.
Published: (2026)
DynTok: Dynamic Compression of Visual Tokens for Efficient and Effective Video Understanding
by: Zhang, Hongzhi, et al.
Published: (2025)
by: Zhang, Hongzhi, et al.
Published: (2025)
LEAD: Latent Realignment for Human Motion Diffusion
by: Andreou, Nefeli, et al.
Published: (2024)
by: Andreou, Nefeli, et al.
Published: (2024)
Few-Shot Multi-Human Neural Rendering Using Geometry Constraints
by: li, Qian, et al.
Published: (2025)
by: li, Qian, et al.
Published: (2025)
SDFit: 3D Object Pose and Shape by Fitting a Morphable SDF to a Single Image
by: Antić, Dimitrije, et al.
Published: (2024)
by: Antić, Dimitrije, et al.
Published: (2024)
DynRefer: Delving into Region-level Multimodal Tasks via Dynamic Resolution
by: Zhao, Yuzhong, et al.
Published: (2024)
by: Zhao, Yuzhong, et al.
Published: (2024)
DynCIM: Dynamic Curriculum for Imbalanced Multimodal Learning
by: Qian, Chengxuan, et al.
Published: (2025)
by: Qian, Chengxuan, et al.
Published: (2025)
DynPoint: Dynamic Neural Point For View Synthesis
by: Zhou, Kaichen, et al.
Published: (2023)
by: Zhou, Kaichen, et al.
Published: (2023)
Dyn-HaMR: Recovering 4D Interacting Hand Motion from a Dynamic Camera
by: Yu, Zhengdi, et al.
Published: (2024)
by: Yu, Zhengdi, et al.
Published: (2024)
ChatDyn: Language-Driven Multi-Actor Dynamics Generation in Street Scenes
by: Wei, Yuxi, et al.
Published: (2024)
by: Wei, Yuxi, et al.
Published: (2024)
DynProto: Dynamic Prototype Evolution for Out-of-Distribution Detection
by: Wu, Yanqi, et al.
Published: (2026)
by: Wu, Yanqi, et al.
Published: (2026)
DynFocus: Dynamic Cooperative Network Empowers LLMs with Video Understanding
by: Han, Yudong, et al.
Published: (2024)
by: Han, Yudong, et al.
Published: (2024)
ETCH: Generalizing Body Fitting to Clothed Humans via Equivariant Tightness
by: Li, Boqian, et al.
Published: (2025)
by: Li, Boqian, et al.
Published: (2025)
DynT2I-Eval: A Dynamic Evaluation Framework for Text-to-Image Models
by: Wang, Juntong, et al.
Published: (2026)
by: Wang, Juntong, et al.
Published: (2026)
Half-Physics: Enabling Kinematic 3D Human Model with Physical Interactions
by: Siyao, Li, et al.
Published: (2025)
by: Siyao, Li, et al.
Published: (2025)
MoTrans: Customized Motion Transfer with Text-driven Video Diffusion Models
by: Li, Xiaomin, et al.
Published: (2024)
by: Li, Xiaomin, et al.
Published: (2024)
InterRVOS: Interaction-aware Referring Video Object Segmentation
by: Jin, Woojeong, et al.
Published: (2025)
by: Jin, Woojeong, et al.
Published: (2025)
Dyn-E: Local Appearance Editing of Dynamic Neural Radiance Fields
by: Zhang, Shangzan, et al.
Published: (2023)
by: Zhang, Shangzan, et al.
Published: (2023)
DynAlign: Unsupervised Dynamic Taxonomy Alignment for Cross-Domain Segmentation
by: Sun, Han, et al.
Published: (2025)
by: Sun, Han, et al.
Published: (2025)
DynSUP: Dynamic Gaussian Splatting from An Unposed Image Pair
by: Li, Weihang, et al.
Published: (2024)
by: Li, Weihang, et al.
Published: (2024)
DynRsl-VLM: Enhancing Autonomous Driving Perception with Dynamic Resolution Vision-Language Models
by: Zhou, Xirui, et al.
Published: (2025)
by: Zhou, Xirui, et al.
Published: (2025)
DynFlowDrive: Flow-Based Dynamic World Modeling for Autonomous Driving
by: Liu, Xiaolu, et al.
Published: (2026)
by: Liu, Xiaolu, et al.
Published: (2026)
InterDyad: Interactive Dyadic Speech-to-Video Generation by Querying Intermediate Visual Guidance
by: Pan, Dongwei, et al.
Published: (2026)
by: Pan, Dongwei, et al.
Published: (2026)
DiffLocks: Generating 3D Hair from a Single Image using Diffusion Models
by: Rosu, Radu Alexandru, et al.
Published: (2025)
by: Rosu, Radu Alexandru, et al.
Published: (2025)
St4RTrack: Simultaneous 4D Reconstruction and Tracking in the World
by: Feng, Haiwen, et al.
Published: (2025)
by: Feng, Haiwen, et al.
Published: (2025)
DynFaceRestore: Balancing Fidelity and Quality in Diffusion-Guided Blind Face Restoration with Dynamic Blur-Level Mapping and Guidance
by: Do, Huu-Phu, et al.
Published: (2025)
by: Do, Huu-Phu, et al.
Published: (2025)
Similar Items
-
GenLit: Reformulating Single-Image Relighting as Video Generation
by: Bharadwaj, Shrisha, et al.
Published: (2024) -
Re-Thinking Inverse Graphics With Large Language Models
by: Kulits, Peter, et al.
Published: (2024) -
Explorative Inbetweening of Time and Space
by: Feng, Haiwen, et al.
Published: (2024) -
OFER: Occluded Face Expression Reconstruction
by: Selvaraju, Pratheba, et al.
Published: (2024) -
InteractVLM: 3D Interaction Reasoning from 2D Foundational Models
by: Dwivedi, Sai Kumar, et al.
Published: (2025)