PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Shaowei, Ren, Zhongzheng, Gupta, Saurabh, Wang, Shenlong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PhysGen3D: Crafting a Miniature Interactive World from a Single Image
by: Chen, Boyuan, et al.
Published: (2025)
by: Chen, Boyuan, et al.
Published: (2025)
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025)
by: Liu, Shaowei, et al.
Published: (2025)
MoRight: Motion Control Done Right
by: Liu, Shaowei, et al.
Published: (2026)
by: Liu, Shaowei, et al.
Published: (2026)
PhysGen: Physically Grounded 3D Shape Generation for Industrial Design
by: You, Yingxuan, et al.
Published: (2025)
by: You, Yingxuan, et al.
Published: (2025)
Learning to Generate Rigid Body Interactions with Video Diffusion Models
by: Romero, David, et al.
Published: (2025)
by: Romero, David, et al.
Published: (2025)
Mitigating Perspective Distortion-induced Shape Ambiguity in Image Crops
by: Prakash, Aditya, et al.
Published: (2023)
by: Prakash, Aditya, et al.
Published: (2023)
Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video
by: Yao, David Yifan, et al.
Published: (2025)
by: Yao, David Yifan, et al.
Published: (2025)
Bimanual 3D Hand Motion and Articulation Forecasting in Everyday Images
by: Prakash, Aditya, et al.
Published: (2025)
by: Prakash, Aditya, et al.
Published: (2025)
PhysTwin: Physics-Informed Reconstruction and Simulation of Deformable Objects from Videos
by: Jiang, Hanxiao, et al.
Published: (2025)
by: Jiang, Hanxiao, et al.
Published: (2025)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
PhysGaussian: Physics-Integrated 3D Gaussians for Generative Dynamics
by: Xie, Tianyi, et al.
Published: (2023)
by: Xie, Tianyi, et al.
Published: (2023)
PhysInOne: Visual Physics Learning and Reasoning in One Suite
by: Zhou, Siyuan, et al.
Published: (2026)
by: Zhou, Siyuan, et al.
Published: (2026)
GenMM: Geometrically and Temporally Consistent Multimodal Data Generation for Video and LiDAR
by: Singh, Bharat, et al.
Published: (2024)
by: Singh, Bharat, et al.
Published: (2024)
3D Hand Pose Estimation in Everyday Egocentric Images
by: Prakash, Aditya, et al.
Published: (2023)
by: Prakash, Aditya, et al.
Published: (2023)
DreamPartGen: Semantically Grounded Part-Level 3D Generation via Collaborative Latent Denoising
by: Yu, Tianjiao, et al.
Published: (2026)
by: Yu, Tianjiao, et al.
Published: (2026)
Push Past Green: Learning to Look Behind Plant Foliage by Moving It
by: Zhang, Xiaoyu, et al.
Published: (2023)
by: Zhang, Xiaoyu, et al.
Published: (2023)
Precise Mobile Manipulation of Small Everyday Objects
by: Gupta, Arjun, et al.
Published: (2025)
by: Gupta, Arjun, et al.
Published: (2025)
SnapGen-V: Generating a Five-Second Video within Five Seconds on a Mobile Device
by: Wu, Yushu, et al.
Published: (2024)
by: Wu, Yushu, et al.
Published: (2024)
Cortex-Grounded Diffusion Models for Brain Image Generation
by: Bongratz, Fabian, et al.
Published: (2026)
by: Bongratz, Fabian, et al.
Published: (2026)
Gen-n-Val: Agentic Image Data Generation and Validation
by: Huang, Jing-En, et al.
Published: (2025)
by: Huang, Jing-En, et al.
Published: (2025)
Ctrl-GenAug: Controllable Generative Augmentation for Medical Sequence Classification
by: Zhou, Xinrui, et al.
Published: (2024)
by: Zhou, Xinrui, et al.
Published: (2024)
PhyGround: Benchmarking Physical Reasoning in Generative World Models
by: Lin, Juyi, et al.
Published: (2026)
by: Lin, Juyi, et al.
Published: (2026)
Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models
by: Jeong, Hyeonho, et al.
Published: (2023)
by: Jeong, Hyeonho, et al.
Published: (2023)
VideoPhy: Evaluating Physical Commonsense for Video Generation
by: Bansal, Hritik, et al.
Published: (2024)
by: Bansal, Hritik, et al.
Published: (2024)
GenURL: A General Framework for Unsupervised Representation Learning
by: Li, Siyuan, et al.
Published: (2021)
by: Li, Siyuan, et al.
Published: (2021)
Contextualized Diffusion Models for Text-Guided Image and Video Generation
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
Opening Articulated Structures in the Real World
by: Gupta, Arjun, et al.
Published: (2024)
by: Gupta, Arjun, et al.
Published: (2024)
Pathways on the Image Manifold: Image Editing via Video Generation
by: Rotstein, Noam, et al.
Published: (2024)
by: Rotstein, Noam, et al.
Published: (2024)
FrameBridge: Improving Image-to-Video Generation with Bridge Models
by: Wang, Yuji, et al.
Published: (2024)
by: Wang, Yuji, et al.
Published: (2024)
PhysMoDPO: Physically-Plausible Humanoid Motion with Preference Optimization
by: Zhang, Yangsong, et al.
Published: (2026)
by: Zhang, Yangsong, et al.
Published: (2026)
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
by: Moon, Sungho, et al.
Published: (2026)
by: Moon, Sungho, et al.
Published: (2026)
Bridging Vision Language Models and Symbolic Grounding for Video Question Answering
by: Ma, Haodi, et al.
Published: (2025)
by: Ma, Haodi, et al.
Published: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
SurGen: Text-Guided Diffusion Model for Surgical Video Generation
by: Cho, Joseph, et al.
Published: (2024)
by: Cho, Joseph, et al.
Published: (2024)
Moving Off-the-Grid: Scene-Grounded Video Representations
by: van Steenkiste, Sjoerd, et al.
Published: (2024)
by: van Steenkiste, Sjoerd, et al.
Published: (2024)
GenHMR: Generative Human Mesh Recovery
by: Saleem, Muhammad Usama, et al.
Published: (2024)
by: Saleem, Muhammad Usama, et al.
Published: (2024)
PoM: Efficient Image and Video Generation with the Polynomial Mixer
by: Picard, David, et al.
Published: (2024)
by: Picard, David, et al.
Published: (2024)
3D Reconstruction of Objects in Hands without Real World 3D Supervision
by: Prakash, Aditya, et al.
Published: (2023)
by: Prakash, Aditya, et al.
Published: (2023)
Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer
by: Kienzle, Daniel, et al.
Published: (2025)
by: Kienzle, Daniel, et al.
Published: (2025)
R2GenKG: Hierarchical Multi-modal Knowledge Graph for LLM-based Radiology Report Generation
by: Wang, Futian, et al.
Published: (2025)
by: Wang, Futian, et al.
Published: (2025)
Similar Items
-
PhysGen3D: Crafting a Miniature Interactive World from a Single Image
by: Chen, Boyuan, et al.
Published: (2025) -
Visual Sync: Multi-Camera Synchronization via Cross-View Object Motion
by: Liu, Shaowei, et al.
Published: (2025) -
MoRight: Motion Control Done Right
by: Liu, Shaowei, et al.
Published: (2026) -
PhysGen: Physically Grounded 3D Shape Generation for Industrial Design
by: You, Yingxuan, et al.
Published: (2025) -
Learning to Generate Rigid Body Interactions with Video Diffusion Models
by: Romero, David, et al.
Published: (2025)