AnyLift: Scaling Motion Reconstruction from Internet Videos via 2D Diffusion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Hongjie, Yu, Heng, Li, Jiaman, Yu, Hong-Xing, Adeli, Ehsan, Liu, C. Karen, Wu, Jiajun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Lifting Motion to the 3D World via 2D Diffusion
von: Li, Jiaman, et al.
Veröffentlicht: (2024)
von: Li, Jiaman, et al.
Veröffentlicht: (2024)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
von: Li, Hongjie, et al.
Veröffentlicht: (2024)
von: Li, Hongjie, et al.
Veröffentlicht: (2024)
WHOLE: World-Grounded Hand-Object Lifted from Egocentric Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2026)
von: Ye, Yufei, et al.
Veröffentlicht: (2026)
FluidNexus: 3D Fluid Reconstruction and Prediction from a Single Video
von: Gao, Yue, et al.
Veröffentlicht: (2025)
von: Gao, Yue, et al.
Veröffentlicht: (2025)
Reconstruction and Simulation of Elastic Objects with Spring-Mass 3D Gaussians
von: Zhong, Licheng, et al.
Veröffentlicht: (2024)
von: Zhong, Licheng, et al.
Veröffentlicht: (2024)
AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model
von: Chen, Yutian, et al.
Veröffentlicht: (2026)
von: Chen, Yutian, et al.
Veröffentlicht: (2026)
WonderZoom: Multi-Scale 3D World Generation
von: Cao, Jin, et al.
Veröffentlicht: (2025)
von: Cao, Jin, et al.
Veröffentlicht: (2025)
UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and Generation
von: Patel, Chaitanya, et al.
Veröffentlicht: (2025)
von: Patel, Chaitanya, et al.
Veröffentlicht: (2025)
VisualMimic: Visual Humanoid Loco-Manipulation via Motion Tracking and Generation
von: Yin, Shaofeng, et al.
Veröffentlicht: (2025)
von: Yin, Shaofeng, et al.
Veröffentlicht: (2025)
HumanScore: Benchmarking Human Motions in Generated Videos
von: Fang, Yusu, et al.
Veröffentlicht: (2026)
von: Fang, Yusu, et al.
Veröffentlicht: (2026)
AnyMo: Scaling Any-Modality Conditional Motion Generation with Masked Modeling
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
von: Li, Yiheng, et al.
Veröffentlicht: (2026)
LiftImage3D: Lifting Any Single Image to 3D Gaussians with Video Generation Priors
von: Chen, Yabo, et al.
Veröffentlicht: (2024)
von: Chen, Yabo, et al.
Veröffentlicht: (2024)
PerpetualWonder: Long-Horizon Action-Conditioned 4D Scene Generation
von: Zhan, Jiahao, et al.
Veröffentlicht: (2026)
von: Zhan, Jiahao, et al.
Veröffentlicht: (2026)
Repurposing 2D Diffusion Models for 3D Shape Completion
von: He, Yao, et al.
Veröffentlicht: (2025)
von: He, Yao, et al.
Veröffentlicht: (2025)
Controllable Human-Object Interaction Synthesis
von: Li, Jiaman, et al.
Veröffentlicht: (2023)
von: Li, Jiaman, et al.
Veröffentlicht: (2023)
Integrating Anatomical Priors into a Causal Diffusion Model
von: Li, Binxu, et al.
Veröffentlicht: (2025)
von: Li, Binxu, et al.
Veröffentlicht: (2025)
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion
von: Chen, Changan, et al.
Veröffentlicht: (2024)
von: Chen, Changan, et al.
Veröffentlicht: (2024)
Segment Any Motion in Videos
von: Huang, Nan, et al.
Veröffentlicht: (2025)
von: Huang, Nan, et al.
Veröffentlicht: (2025)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2025)
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2025)
Drive Any Mesh: 4D Latent Diffusion for Mesh Deformation from Video
von: Shi, Yahao, et al.
Veröffentlicht: (2025)
von: Shi, Yahao, et al.
Veröffentlicht: (2025)
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D
von: T, Mukund Varma, et al.
Veröffentlicht: (2024)
von: T, Mukund Varma, et al.
Veröffentlicht: (2024)
AdaVid: Adaptive Video-Language Pretraining
von: Patel, Chaitanya, et al.
Veröffentlicht: (2025)
von: Patel, Chaitanya, et al.
Veröffentlicht: (2025)
AVID: Any-Length Video Inpainting with Diffusion Model
von: Zhang, Zhixing, et al.
Veröffentlicht: (2023)
von: Zhang, Zhixing, et al.
Veröffentlicht: (2023)
Repurposing 2D Diffusion Models with Gaussian Atlas for 3D Generation
von: Xiang, Tiange, et al.
Veröffentlicht: (2025)
von: Xiang, Tiange, et al.
Veröffentlicht: (2025)
GenFusion: Feed-forward Human Performance Capture via Progressive Canonical Space Updates
von: Kwon, Youngjoong, et al.
Veröffentlicht: (2026)
von: Kwon, Youngjoong, et al.
Veröffentlicht: (2026)
Web-Scale Collection of Video Data for 4D Animal Reconstruction
von: Zhao, Brian Nlong, et al.
Veröffentlicht: (2025)
von: Zhao, Brian Nlong, et al.
Veröffentlicht: (2025)
LiftVSR: Lifting Image Diffusion to Video Super-Resolution via Hybrid Temporal Modeling with Only 4$\times$RTX 4090s
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
von: Wang, Xijun, et al.
Veröffentlicht: (2025)
OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
von: Sun, Adam, et al.
Veröffentlicht: (2024)
von: Sun, Adam, et al.
Veröffentlicht: (2024)
Animate3D: Animating Any 3D Model with Multi-view Video Diffusion
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
von: Jiang, Yanqin, et al.
Veröffentlicht: (2024)
Ponymation: Learning Articulated 3D Animal Motions from Unlabeled Online Videos
von: Sun, Keqiang, et al.
Veröffentlicht: (2023)
von: Sun, Keqiang, et al.
Veröffentlicht: (2023)
GeoGS3D: Single-view 3D Reconstruction via Geometric-aware Diffusion Model and Gaussian Splatting
von: Feng, Qijun, et al.
Veröffentlicht: (2024)
von: Feng, Qijun, et al.
Veröffentlicht: (2024)
A Tool Bottleneck Framework for Clinically-Informed and Interpretable Medical Image Understanding
von: Liu, Christina, et al.
Veröffentlicht: (2025)
von: Liu, Christina, et al.
Veröffentlicht: (2025)
Motion Anything: Any to Motion Generation
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
von: Zhang, Zeyu, et al.
Veröffentlicht: (2025)
PhysDreamer: Physics-Based Interaction with 3D Objects via Video Generation
von: Zhang, Tianyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyuan, et al.
Veröffentlicht: (2024)
Unsupervised Discovery of Object-Centric Neural Fields
von: Luo, Rundong, et al.
Veröffentlicht: (2024)
von: Luo, Rundong, et al.
Veröffentlicht: (2024)
Towards Robust 3D Pose Transfer with Adversarial Learning
von: Chen, Haoyu, et al.
Veröffentlicht: (2024)
von: Chen, Haoyu, et al.
Veröffentlicht: (2024)
AnyAct: Towards Human Reenactment of Character Motion From Video
von: Chen, Liuhan, et al.
Veröffentlicht: (2026)
von: Chen, Liuhan, et al.
Veröffentlicht: (2026)
Diffusion Models for Computational Neuroimaging: A Survey
von: Zhao, Haokai, et al.
Veröffentlicht: (2025)
von: Zhao, Haokai, et al.
Veröffentlicht: (2025)
Anatomically Guided Latent Diffusion for Brain MRI Progression Modeling
von: Wan, Cheng, et al.
Veröffentlicht: (2026)
von: Wan, Cheng, et al.
Veröffentlicht: (2026)
MVGenMaster: Scaling Multi-View Generation from Any Image via 3D Priors Enhanced Diffusion Model
von: Cao, Chenjie, et al.
Veröffentlicht: (2024)
von: Cao, Chenjie, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Lifting Motion to the 3D World via 2D Diffusion
von: Li, Jiaman, et al.
Veröffentlicht: (2024) -
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
von: Li, Hongjie, et al.
Veröffentlicht: (2024) -
WHOLE: World-Grounded Hand-Object Lifted from Egocentric Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2026) -
FluidNexus: 3D Fluid Reconstruction and Prediction from a Single Video
von: Gao, Yue, et al.
Veröffentlicht: (2025) -
Reconstruction and Simulation of Elastic Objects with Spring-Mass 3D Gaussians
von: Zhong, Licheng, et al.
Veröffentlicht: (2024)