DreamPhysics: Learning Physics-Based 3D Dynamics with Video Diffusion Priors
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Tianyu, Zhang, Haoze, Zeng, Yihan, Zhang, Zhilu, Li, Hui, Zuo, Wangmeng, Lau, Rynson W. H. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
di: Huang, Tianyu, et al.
Pubblicazione: (2023)
di: Huang, Tianyu, et al.
Pubblicazione: (2023)
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
di: Huang, Tianyu, et al.
Pubblicazione: (2023)
di: Huang, Tianyu, et al.
Pubblicazione: (2023)
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
di: Yang, Yu, et al.
Pubblicazione: (2025)
di: Yang, Yu, et al.
Pubblicazione: (2025)
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
di: Zhang, Yabo, et al.
Pubblicazione: (2025)
di: Zhang, Yabo, et al.
Pubblicazione: (2025)
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
di: Zhang, Haoze, et al.
Pubblicazione: (2025)
di: Zhang, Haoze, et al.
Pubblicazione: (2025)
Voyager: Long-Range and World-Consistent Video Diffusion for Explorable 3D Scene Generation
di: Huang, Tianyu, et al.
Pubblicazione: (2025)
di: Huang, Tianyu, et al.
Pubblicazione: (2025)
Rethinking Transformer-Based Blind-Spot Network for Self-Supervised Image Denoising
di: Li, Junyi, et al.
Pubblicazione: (2024)
di: Li, Junyi, et al.
Pubblicazione: (2024)
Deblur4DGS: 4D Gaussian Splatting from Blurry Monocular Video
di: Wu, Renlong, et al.
Pubblicazione: (2024)
di: Wu, Renlong, et al.
Pubblicazione: (2024)
Self-Supervised Learning for Real-World Super-Resolution from Dual and Multiple Zoomed Observations
di: Zhang, Zhilu, et al.
Pubblicazione: (2024)
di: Zhang, Zhilu, et al.
Pubblicazione: (2024)
Lie Flow: Video Dynamic Fields Modeling and Predicting with Lie Algebra as Geometric Physics Principle
di: Qiao, Weidong, et al.
Pubblicazione: (2026)
di: Qiao, Weidong, et al.
Pubblicazione: (2026)
ArtHOI: Taming Foundation Models for Monocular 4D Reconstruction of Hand-Articulated-Object Interactions
di: Wang, Zikai, et al.
Pubblicazione: (2026)
di: Wang, Zikai, et al.
Pubblicazione: (2026)
MV-VTON: Multi-View Virtual Try-On with Diffusion Models
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
di: Wang, Haoyu, et al.
Pubblicazione: (2024)
Bridging Geometry-Coherent Text-to-3D Generation with Multi-View Diffusion Priors and Gaussian Splatting
di: Yang, Feng, et al.
Pubblicazione: (2025)
di: Yang, Feng, et al.
Pubblicazione: (2025)
SelfHVD: Self-Supervised Handheld Video Deblurring
di: Xu, Honglei, et al.
Pubblicazione: (2025)
di: Xu, Honglei, et al.
Pubblicazione: (2025)
ScrollScape: Unlocking 32K Image Generation With Video Diffusion Priors
di: Yu, Haodong, et al.
Pubblicazione: (2026)
di: Yu, Haodong, et al.
Pubblicazione: (2026)
Image Demoiréing Using Dual Camera Fusion on Mobile Phones
di: Mei, Yanting, et al.
Pubblicazione: (2025)
di: Mei, Yanting, et al.
Pubblicazione: (2025)
NIR-Assisted Image Denoising: A Selective Fusion Approach and A Real-World Benchmark Dataset
di: Xu, Rongjian, et al.
Pubblicazione: (2024)
di: Xu, Rongjian, et al.
Pubblicazione: (2024)
Dual-Camera Smooth Zoom on Mobile Phones
di: Wu, Renlong, et al.
Pubblicazione: (2024)
di: Wu, Renlong, et al.
Pubblicazione: (2024)
Self-Supervised High Dynamic Range Imaging with Multi-Exposure Images in Dynamic Scenes
di: Zhang, Zhilu, et al.
Pubblicazione: (2023)
di: Zhang, Zhilu, et al.
Pubblicazione: (2023)
FILP-3D: Enhancing 3D Few-shot Class-incremental Learning with Pre-trained Vision-Language Models
di: Xu, Wan, et al.
Pubblicazione: (2023)
di: Xu, Wan, et al.
Pubblicazione: (2023)
Hierarchical Cross-Modal Alignment for Open-Vocabulary 3D Object Detection
di: Zhao, Youjun, et al.
Pubblicazione: (2025)
di: Zhao, Youjun, et al.
Pubblicazione: (2025)
Shape-for-Motion: Precise and Consistent Video Editing with 3D Proxy
di: Liu, Yuhao, et al.
Pubblicazione: (2025)
di: Liu, Yuhao, et al.
Pubblicazione: (2025)
MirrorMamba: Towards Scalable and Robust Mirror Detection in Videos
di: Song, Rui, et al.
Pubblicazione: (2025)
di: Song, Rui, et al.
Pubblicazione: (2025)
Exposure Bracketing Is All You Need For A High-Quality Image
di: Zhang, Zhilu, et al.
Pubblicazione: (2024)
di: Zhang, Zhilu, et al.
Pubblicazione: (2024)
Pseudo-Label Guided Real-World Image De-weathering: A Learning Framework with Imperfect Supervision
di: Xu, Heming, et al.
Pubblicazione: (2025)
di: Xu, Heming, et al.
Pubblicazione: (2025)
Self-Supervised Video Desmoking for Laparoscopic Surgery
di: Wu, Renlong, et al.
Pubblicazione: (2024)
di: Wu, Renlong, et al.
Pubblicazione: (2024)
S2AM3D: Scale-controllable Part Segmentation of 3D Point Clouds
di: Su, Han, et al.
Pubblicazione: (2025)
di: Su, Han, et al.
Pubblicazione: (2025)
Phidias: A Generative Model for Creating 3D Content from Text, Image, and 3D Conditions with Reference-Augmented Diffusion
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
di: Wang, Zhenwei, et al.
Pubblicazione: (2024)
Tool-R1: Sample-Efficient Reinforcement Learning for Agentic Tool Use
di: Zhang, Yabo, et al.
Pubblicazione: (2025)
di: Zhang, Yabo, et al.
Pubblicazione: (2025)
Arbitrary-Scale Video Super-Resolution with Structural and Textural Priors
di: Shang, Wei, et al.
Pubblicazione: (2024)
di: Shang, Wei, et al.
Pubblicazione: (2024)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
di: Wang, Haoyu, et al.
Pubblicazione: (2025)
Unleashing the Potential of Multimodal LLMs for Zero-Shot Spatio-Temporal Video Grounding
di: Yang, Zaiquan, et al.
Pubblicazione: (2025)
di: Yang, Zaiquan, et al.
Pubblicazione: (2025)
DreamShot: Personalized Storyboard Synthesis with Video Diffusion Prior
di: Huang, Junjia, et al.
Pubblicazione: (2026)
di: Huang, Junjia, et al.
Pubblicazione: (2026)
Improving Image Restoration through Removing Degradations in Textual Representations
di: Lin, Jingbo, et al.
Pubblicazione: (2023)
di: Lin, Jingbo, et al.
Pubblicazione: (2023)
Harnessing Scale and Physics: A Multi-Graph Neural Operator Framework for PDEs on Arbitrary Geometries
di: Li, Zhihao, et al.
Pubblicazione: (2024)
di: Li, Zhihao, et al.
Pubblicazione: (2024)
Spatiotemporal Field Generation Based on Hybrid Mamba-Transformer with Physics-informed Fine-tuning
di: Du, Peimian, et al.
Pubblicazione: (2025)
di: Du, Peimian, et al.
Pubblicazione: (2025)
RefSTAR: Blind Facial Image Restoration with Reference Selection, Transfer, and Reconstruction
di: Yin, Zhicun, et al.
Pubblicazione: (2025)
di: Yin, Zhicun, et al.
Pubblicazione: (2025)
RelayAttention for Efficient Large Language Model Serving with Long System Prompts
di: Zhu, Lei, et al.
Pubblicazione: (2024)
di: Zhu, Lei, et al.
Pubblicazione: (2024)
Revisiting the Integration of Convolution and Attention for Vision Backbone
di: Zhu, Lei, et al.
Pubblicazione: (2024)
di: Zhu, Lei, et al.
Pubblicazione: (2024)
Delving into Dark Regions for Robust Shadow Detection
di: Guan, Huankang, et al.
Pubblicazione: (2024)
di: Guan, Huankang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
di: Huang, Tianyu, et al.
Pubblicazione: (2023) -
TextField3D: Towards Enhancing Open-Vocabulary 3D Generation with Noisy Text Fields
di: Huang, Tianyu, et al.
Pubblicazione: (2023) -
PhysWorld: From Real Videos to World Models of Deformable Objects via Physics-Aware Demonstration Synthesis
di: Yang, Yu, et al.
Pubblicazione: (2025) -
FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors
di: Zhang, Yabo, et al.
Pubblicazione: (2025) -
PhysChoreo: Physics-Controllable Video Generation with Part-Aware Semantic Grounding
di: Zhang, Haoze, et al.
Pubblicazione: (2025)