ReconPhys: Reconstruct Appearance and Physical Attributes from Single Video
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wang, Boyuan, Wang, Xiaofeng, Li, Yongkang, Zhu, Zheng, Chang, Yifan, Ye, Angen, Zhao, Guosheng, Ni, Chaojun, Huang, Guan, Ren, Yijie, Duan, Yueqi, Wang, Xingang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
von: Zhao, Guosheng, et al.
Veröffentlicht: (2025)
von: Zhao, Guosheng, et al.
Veröffentlicht: (2025)
HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)
ReconX: Reconstruct Any Scene from Sparse Views with Video Diffusion Model
von: Liu, Fangfu, et al.
Veröffentlicht: (2024)
von: Liu, Fangfu, et al.
Veröffentlicht: (2024)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
von: Wang, Boyuan, et al.
Veröffentlicht: (2025)
SimRecon: SimReady Compositional Scene Reconstruction from Real Videos
von: Xia, Chong, et al.
Veröffentlicht: (2026)
von: Xia, Chong, et al.
Veröffentlicht: (2026)
DriveDreamer4D: World Models Are Effective Data Machines for 4D Driving Scene Representation
von: Zhao, Guosheng, et al.
Veröffentlicht: (2024)
von: Zhao, Guosheng, et al.
Veröffentlicht: (2024)
VLA-R1: Enhancing Reasoning in Vision-Language-Action Models
von: Ye, Angen, et al.
Veröffentlicht: (2025)
von: Ye, Angen, et al.
Veröffentlicht: (2025)
DriveDreamer-2: LLM-Enhanced World Models for Diverse Driving Video Generation
von: Zhao, Guosheng, et al.
Veröffentlicht: (2024)
von: Zhao, Guosheng, et al.
Veröffentlicht: (2024)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
GigaWorld-Policy: An Efficient Action-Centered World--Action Model
von: Ye, Angen, et al.
Veröffentlicht: (2026)
von: Ye, Angen, et al.
Veröffentlicht: (2026)
GigaWorld-0: World Models as Data Engine to Empower Embodied AI
von: GigaWorld Team, et al.
Veröffentlicht: (2025)
von: GigaWorld Team, et al.
Veröffentlicht: (2025)
EgoVid-5M: A Large-Scale Video-Action Dataset for Egocentric Video Generation
von: Wang, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiaofeng, et al.
Veröffentlicht: (2024)
MimicDreamer: Aligning Human and Robot Demonstrations for Scalable VLA Training
von: Li, Haoyun, et al.
Veröffentlicht: (2025)
von: Li, Haoyun, et al.
Veröffentlicht: (2025)
PhyRecon: Physically Plausible Neural Scene Reconstruction
von: Ni, Junfeng, et al.
Veröffentlicht: (2024)
von: Ni, Junfeng, et al.
Veröffentlicht: (2024)
SwiftVLA: Unlocking Spatiotemporal Dynamics for Lightweight VLA Models at Minimal Overhead
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
WonderTurbo: Generating Interactive 3D World in 0.72 Seconds
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
PhysTwin: Physics-Informed Reconstruction and Simulation of Deformable Objects from Videos
von: Jiang, Hanxiao, et al.
Veröffentlicht: (2025)
von: Jiang, Hanxiao, et al.
Veröffentlicht: (2025)
WorldDreamer: Towards General World Models for Video Generation via Predicting Masked Tokens
von: Wang, Xiaofeng, et al.
Veröffentlicht: (2024)
von: Wang, Xiaofeng, et al.
Veröffentlicht: (2024)
GigaBrain-0: A World Model-Powered Vision-Language-Action Model
von: GigaBrain Team, et al.
Veröffentlicht: (2025)
von: GigaBrain Team, et al.
Veröffentlicht: (2025)
UniDriveDreamer: A Single-Stage Multimodal World Model for Autonomous Driving
von: Zhao, Guosheng, et al.
Veröffentlicht: (2026)
von: Zhao, Guosheng, et al.
Veröffentlicht: (2026)
Recon: Reconstruction-Guided Reasoning Synthesis for User Modeling
von: Zhu, Alan, et al.
Veröffentlicht: (2026)
von: Zhu, Alan, et al.
Veröffentlicht: (2026)
Rethinking Lanes and Points in Complex Scenarios for Monocular 3D Lane Detection
von: Chang, Yifan, et al.
Veröffentlicht: (2025)
von: Chang, Yifan, et al.
Veröffentlicht: (2025)
VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step
von: Wang, Hanyang, et al.
Veröffentlicht: (2025)
von: Wang, Hanyang, et al.
Veröffentlicht: (2025)
FedRecon: Missing Modality Reconstruction in Heterogeneous Distributed Environments
von: Liu, Junming, et al.
Veröffentlicht: (2025)
von: Liu, Junming, et al.
Veröffentlicht: (2025)
Scalable Training for Vector-Quantized Networks with 100% Codebook Utilization
von: Chang, Yifan, et al.
Veröffentlicht: (2025)
von: Chang, Yifan, et al.
Veröffentlicht: (2025)
VAG: Dual-Stream Video-Action Generation for Embodied Data Synthesis
von: Lang, Xiaolei, et al.
Veröffentlicht: (2026)
von: Lang, Xiaolei, et al.
Veröffentlicht: (2026)
LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion
von: Liu, Fangfu, et al.
Veröffentlicht: (2025)
von: Liu, Fangfu, et al.
Veröffentlicht: (2025)
LLM-Powered Nuanced Video Attribute Annotation for Enhanced Recommendations
von: Long, Boyuan, et al.
Veröffentlicht: (2025)
von: Long, Boyuan, et al.
Veröffentlicht: (2025)
AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model
von: Chen, Yutian, et al.
Veröffentlicht: (2026)
von: Chen, Yutian, et al.
Veröffentlicht: (2026)
DressRecon: Freeform 4D Human Reconstruction from Monocular Video
von: Tan, Jeff, et al.
Veröffentlicht: (2024)
von: Tan, Jeff, et al.
Veröffentlicht: (2024)
EMMA: Generalizing Real-World Robot Manipulation via Generative Visual Transfer
von: Dong, Zhehao, et al.
Veröffentlicht: (2025)
von: Dong, Zhehao, et al.
Veröffentlicht: (2025)
GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning
von: Bao, Xiaoyi, et al.
Veröffentlicht: (2025)
von: Bao, Xiaoyi, et al.
Veröffentlicht: (2025)
SF-Recon: Simplification-Free Lightweight Building Reconstruction via 3D Gaussian Splatting
von: Li, Zihan, et al.
Veröffentlicht: (2025)
von: Li, Zihan, et al.
Veröffentlicht: (2025)
REACTO: Reconstructing Articulated Objects from a Single Video
von: Song, Chaoyue, et al.
Veröffentlicht: (2024)
von: Song, Chaoyue, et al.
Veröffentlicht: (2024)
PhysGen3D: Crafting a Miniature Interactive World from a Single Image
von: Chen, Boyuan, et al.
Veröffentlicht: (2025)
von: Chen, Boyuan, et al.
Veröffentlicht: (2025)
DriveGen3D: Boosting Feed-Forward Driving Scene Generation with Efficient Video Diffusion
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
von: Wang, Weijie, et al.
Veröffentlicht: (2025)
PhysGen: Rigid-Body Physics-Grounded Image-to-Video Generation
von: Liu, Shaowei, et al.
Veröffentlicht: (2024)
von: Liu, Shaowei, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
ReconDreamer++: Harmonizing Generative and Reconstructive Models for Driving Scene Representation
von: Zhao, Guosheng, et al.
Veröffentlicht: (2025) -
HumanDreamer-X: Photorealistic Single-image Human Avatars Reconstruction via Gaussian Restoration
von: Wang, Boyuan, et al.
Veröffentlicht: (2025) -
EmbodieDreamer: Advancing Real2Sim2Real Transfer for Policy Training via Embodied World Modeling
von: Wang, Boyuan, et al.
Veröffentlicht: (2025) -
ReconDreamer-RL: Enhancing Reinforcement Learning via Diffusion-based Scene Reconstruction
von: Ni, Chaojun, et al.
Veröffentlicht: (2025) -
ReconDreamer: Crafting World Models for Driving Scene Reconstruction via Online Restoration
von: Ni, Chaojun, et al.
Veröffentlicht: (2024)