Puppet-Master: Scaling Interactive Video Generation as a Motion Prior for Part-Level Dynamics
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Ruining, Zheng, Chuanxia, Rupprecht, Christian, Vedaldi, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DragAPart: Learning a Part-Level Motion Prior for Articulated Objects
by: Li, Ruining, et al.
Published: (2024)
by: Li, Ruining, et al.
Published: (2024)
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
Particulate: Feed-Forward 3D Object Articulation
by: Li, Ruining, et al.
Published: (2025)
by: Li, Ruining, et al.
Published: (2025)
What Happens Next? Anticipating Future Motion by Generating Point Trajectories
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Free3D: Consistent Novel View Synthesis without 3D Representation
by: Zheng, Chuanxia, et al.
Published: (2023)
by: Zheng, Chuanxia, et al.
Published: (2023)
Learning segmentation from point trajectories
by: Karazija, Laurynas, et al.
Published: (2025)
by: Karazija, Laurynas, et al.
Published: (2025)
VFMF: World Modeling by Forecasting Vision Foundation Model Features
by: Boduljak, Gabrijel, et al.
Published: (2025)
by: Boduljak, Gabrijel, et al.
Published: (2025)
Articraft: An Agentic System for Scalable Articulated 3D Asset Generation
by: Zhou, Matt, et al.
Published: (2026)
by: Zhou, Matt, et al.
Published: (2026)
Farm3D: Learning Articulated 3D Animals by Distilling 2D Diffusion
by: Jakab, Tomas, et al.
Published: (2023)
by: Jakab, Tomas, et al.
Published: (2023)
Robust Motion Generation using Part-level Reliable Data from Videos
by: Li, Boyuan, et al.
Published: (2025)
by: Li, Boyuan, et al.
Published: (2025)
IM-3D: Iterative Multiview Diffusion and Reconstruction for High-Quality 3D Generation
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
by: Melas-Kyriazi, Luke, et al.
Published: (2024)
Flash3D: Feed-Forward Generalisable 3D Scene Reconstruction from a Single Image
by: Szymanowicz, Stanislaw, et al.
Published: (2024)
by: Szymanowicz, Stanislaw, et al.
Published: (2024)
Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular Video
by: Jiang, Zeren, et al.
Published: (2026)
by: Jiang, Zeren, et al.
Published: (2026)
SHIC: Shape-Image Correspondences with no Keypoint Supervision
by: Shtedritski, Aleksandar, et al.
Published: (2024)
by: Shtedritski, Aleksandar, et al.
Published: (2024)
Splatter Image: Ultra-Fast Single-View 3D Reconstruction
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
by: Szymanowicz, Stanislaw, et al.
Published: (2023)
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
by: Jiang, Zeren, et al.
Published: (2025)
by: Jiang, Zeren, et al.
Published: (2025)
Invisible Stitch: Generating Smooth 3D Scenes with Depth Inpainting
by: Engstler, Paul, et al.
Published: (2024)
by: Engstler, Paul, et al.
Published: (2024)
Resource-Efficient Motion Control for Video Generation via Dynamic Mask Guidance
by: Feng, Sicong, et al.
Published: (2025)
by: Feng, Sicong, et al.
Published: (2025)
Motion-Grounded Video Reasoning: Understanding and Perceiving Motion at Pixel Level
by: Deng, Andong, et al.
Published: (2024)
by: Deng, Andong, et al.
Published: (2024)
Motion-Aware Caching for Efficient Autoregressive Video Generation
by: Xu, Jing, et al.
Published: (2026)
by: Xu, Jing, et al.
Published: (2026)
MotionCrafter: Dense Geometry and Motion Reconstruction with a 4D VAE
by: Zhu, Ruijie, et al.
Published: (2026)
by: Zhu, Ruijie, et al.
Published: (2026)
NOVA3R: Non-pixel-aligned Visual Transformer for Amodal 3D Reconstruction
by: Chen, Weirong, et al.
Published: (2026)
by: Chen, Weirong, et al.
Published: (2026)
HaSPeR: An Image Repository for Hand Shadow Puppet Recognition
by: Raiyan, Syed Rifat, et al.
Published: (2024)
by: Raiyan, Syed Rifat, et al.
Published: (2024)
Diffusion Models for Open-Vocabulary Segmentation
by: Karazija, Laurynas, et al.
Published: (2023)
by: Karazija, Laurynas, et al.
Published: (2023)
SynCity: Training-Free Generation of 3D Worlds
by: Engstler, Paul, et al.
Published: (2025)
by: Engstler, Paul, et al.
Published: (2025)
Ada-VE: Training-Free Consistent Video Editing Using Adaptive Motion Prior
by: Mahmud, Tanvir, et al.
Published: (2024)
by: Mahmud, Tanvir, et al.
Published: (2024)
SPATIALALIGN: Aligning Dynamic Spatial Relationships in Video Generation
by: Liu, Fengming, et al.
Published: (2026)
by: Liu, Fengming, et al.
Published: (2026)
Unmasking Puppeteers: Leveraging Biometric Leakage to Expose Impersonation in AI-Based Videoconferencing
by: Vahdati, Danial Samadi, et al.
Published: (2025)
by: Vahdati, Danial Samadi, et al.
Published: (2025)
Temporal Continual Learning with Prior Compensation for Human Motion Prediction
by: Tang, Jianwei, et al.
Published: (2025)
by: Tang, Jianwei, et al.
Published: (2025)
Boximator: Generating Rich and Controllable Motions for Video Synthesis
by: Wang, Jiawei, et al.
Published: (2024)
by: Wang, Jiawei, et al.
Published: (2024)
Seamless Interaction: Dyadic Audiovisual Motion Modeling and Large-Scale Dataset
by: Agrawal, Vasu, et al.
Published: (2025)
by: Agrawal, Vasu, et al.
Published: (2025)
VideoMaMa: Mask-Guided Video Matting via Generative Prior
by: Lim, Sangbeom, et al.
Published: (2026)
by: Lim, Sangbeom, et al.
Published: (2026)
Amodal3R: Amodal 3D Reconstruction from Occluded 2D Images
by: Wu, Tianhao, et al.
Published: (2025)
by: Wu, Tianhao, et al.
Published: (2025)
TokenMotion: Decoupled Motion Control via Token Disentanglement for Human-centric Video Generation
by: Li, Ruineng, et al.
Published: (2025)
by: Li, Ruineng, et al.
Published: (2025)
CoTracker3: Simpler and Better Point Tracking by Pseudo-Labelling Real Videos
by: Karaev, Nikita, et al.
Published: (2024)
by: Karaev, Nikita, et al.
Published: (2024)
LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior
by: Wang, Hanyu, et al.
Published: (2024)
by: Wang, Hanyu, et al.
Published: (2024)
Motion Dreamer: Boundary Conditional Motion Reasoning for Physically Coherent Video Generation
by: Xu, Tianshuo, et al.
Published: (2024)
by: Xu, Tianshuo, et al.
Published: (2024)
MimicParts: Part-aware Style Injection for Speech-Driven 3D Motion Generation
by: Liu, Lianlian, et al.
Published: (2025)
by: Liu, Lianlian, et al.
Published: (2025)
MOFA-Video: Controllable Image Animation via Generative Motion Field Adaptions in Frozen Image-to-Video Diffusion Model
by: Niu, Muyao, et al.
Published: (2024)
by: Niu, Muyao, et al.
Published: (2024)
Learning the 3D Fauna of the Web
by: Li, Zizhang, et al.
Published: (2024)
by: Li, Zizhang, et al.
Published: (2024)
Similar Items
-
DragAPart: Learning a Part-Level Motion Prior for Articulated Objects
by: Li, Ruining, et al.
Published: (2024) -
DSO: Aligning 3D Generators with Simulation Feedback for Physical Soundness
by: Li, Ruining, et al.
Published: (2025) -
Particulate: Feed-Forward 3D Object Articulation
by: Li, Ruining, et al.
Published: (2025) -
What Happens Next? Anticipating Future Motion by Generating Point Trajectories
by: Boduljak, Gabrijel, et al.
Published: (2025) -
Free3D: Consistent Novel View Synthesis without 3D Representation
by: Zheng, Chuanxia, et al.
Published: (2023)