VideoLifter: Lifting Videos to 3D with Fast Hierarchical Stereo Alignment
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cong, Wenyan, Zhu, Hanqing, Wang, Kevin, Lei, Jiahui, Stearns, Colton, Cai, Yuanhao, Guibas, Leonidas, Wang, Zhangyang, Fan, Zhiwen |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
View-Consistent Hierarchical 3D Segmentation Using Ultrametric Feature Fields
par: He, Haodi, et autres
Publié: (2024)
par: He, Haodi, et autres
Publié: (2024)
Dynamic Gaussian Marbles for Novel View Synthesis of Casual Monocular Videos
par: Stearns, Colton, et autres
Publié: (2024)
par: Stearns, Colton, et autres
Publié: (2024)
Dynamic Reflections: Probing Video Representations with Text Alignment
par: Zhu, Tyler, et autres
Publié: (2025)
par: Zhu, Tyler, et autres
Publié: (2025)
MoSca: Dynamic Gaussian Fusion from Casual Videos via 4D Motion Scaffolds
par: Lei, Jiahui, et autres
Publié: (2024)
par: Lei, Jiahui, et autres
Publié: (2024)
Can Test-Time Scaling Improve World Foundation Model?
par: Cong, Wenyan, et autres
Publié: (2025)
par: Cong, Wenyan, et autres
Publié: (2025)
StereoDiff: Stereo-Diffusion Synergy for Video Depth Estimation
par: Li, Haodong, et autres
Publié: (2025)
par: Li, Haodong, et autres
Publié: (2025)
Feature4X: Bridging Any Monocular Video to 4D Agentic AI with Versatile Gaussian Feature Fields
par: Zhou, Shijie, et autres
Publié: (2025)
par: Zhou, Shijie, et autres
Publié: (2025)
CuLifter: Lifting GPU Binaries to Typed IR
par: Zhao, Jisheng, et autres
Publié: (2026)
par: Zhao, Jisheng, et autres
Publié: (2026)
Global Motion Corresponder for 3D Point-Based Scene Interpolation under Large Motion
par: Lin, Junru, et autres
Publié: (2025)
par: Lin, Junru, et autres
Publié: (2025)
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D
par: T, Mukund Varma, et autres
Publié: (2024)
par: T, Mukund Varma, et autres
Publié: (2024)
CurveCloudNet: Processing Point Clouds with 1D Structure
par: Stearns, Colton, et autres
Publié: (2023)
par: Stearns, Colton, et autres
Publié: (2023)
Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions
par: Li, Longfei, et autres
Publié: (2025)
par: Li, Longfei, et autres
Publié: (2025)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
par: Liu, Yu, et autres
Publié: (2024)
par: Liu, Yu, et autres
Publié: (2024)
MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps
par: Lei, Jiahui, et autres
Publié: (2025)
par: Lei, Jiahui, et autres
Publié: (2025)
Mode Seeking meets Mean Seeking for Fast Long Video Generation
par: Cai, Shengqu, et autres
Publié: (2026)
par: Cai, Shengqu, et autres
Publié: (2026)
Expressive Gaussian Human Avatars from Monocular RGB Video
par: Hu, Hezhen, et autres
Publié: (2024)
par: Hu, Hezhen, et autres
Publié: (2024)
Match-Stereo-Videos: Bidirectional Alignment for Consistent Dynamic Stereo Matching
par: Jing, Junpeng, et autres
Publié: (2024)
par: Jing, Junpeng, et autres
Publié: (2024)
Collaborative Video Diffusion: Consistent Multi-video Generation with Camera Control
par: Kuang, Zhengfei, et autres
Publié: (2024)
par: Kuang, Zhengfei, et autres
Publié: (2024)
Match Stereo Videos via Bidirectional Alignment
par: Jing, Junpeng, et autres
Publié: (2024)
par: Jing, Junpeng, et autres
Publié: (2024)
Studentized Tests of Independence: Random-Lifter approach
par: Gao, Zhe, et autres
Publié: (2024)
par: Gao, Zhe, et autres
Publié: (2024)
FSGS: Real-Time Few-shot View Synthesis using Gaussian Splatting
par: Zhu, Zehao, et autres
Publié: (2023)
par: Zhu, Zehao, et autres
Publié: (2023)
LightGaussian: Unbounded 3D Gaussian Compression with 15x Reduction and 200+ FPS
par: Fan, Zhiwen, et autres
Publié: (2023)
par: Fan, Zhiwen, et autres
Publié: (2023)
Forklift: An Extensible Neural Lifter
par: Armengol-Estapé, Jordi, et autres
Publié: (2024)
par: Armengol-Estapé, Jordi, et autres
Publié: (2024)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
par: Deng, Boyang, et autres
Publié: (2024)
par: Deng, Boyang, et autres
Publié: (2024)
InstantSplat: Sparse-view Gaussian Splatting in Seconds
par: Fan, Zhiwen, et autres
Publié: (2024)
par: Fan, Zhiwen, et autres
Publié: (2024)
E3D-Bench: A Benchmark for End-to-End 3D Geometric Foundation Models
par: Cong, Wenyan, et autres
Publié: (2025)
par: Cong, Wenyan, et autres
Publié: (2025)
Video Perception Models for 3D Scene Synthesis
par: Huang, Rui, et autres
Publié: (2025)
par: Huang, Rui, et autres
Publié: (2025)
StereoWorld: Geometry-Aware Monocular-to-Stereo Video Generation
par: Xing, Ke, et autres
Publié: (2025)
par: Xing, Ke, et autres
Publié: (2025)
SlowFast-VGen: Slow-Fast Learning for Action-Driven Long Video Generation
par: Hong, Yining, et autres
Publié: (2024)
par: Hong, Yining, et autres
Publié: (2024)
PASTA: Controllable Part-Aware Shape Generation with Autoregressive Transformers
par: Li, Songlin, et autres
Publié: (2024)
par: Li, Songlin, et autres
Publié: (2024)
BlenderAlchemy: Editing 3D Graphics with Vision-Language Models
par: Huang, Ian, et autres
Publié: (2024)
par: Huang, Ian, et autres
Publié: (2024)
INR-Arch: A Dataflow Architecture and Compiler for Arbitrary-Order Gradient Computations in Implicit Neural Representation Processing
par: Abi-Karam, Stefan, et autres
Publié: (2023)
par: Abi-Karam, Stefan, et autres
Publié: (2023)
DissolveStereo: Coarse Depth Injection for Zero-Shot Stereo Video Generation
par: Shi, Jian, et autres
Publié: (2024)
par: Shi, Jian, et autres
Publié: (2024)
Stereo Any Video: Temporally Consistent Stereo Matching
par: Jing, Junpeng, et autres
Publié: (2025)
par: Jing, Junpeng, et autres
Publié: (2025)
APOLLO: SGD-like Memory, AdamW-level Performance
par: Zhu, Hanqing, et autres
Publié: (2024)
par: Zhu, Hanqing, et autres
Publié: (2024)
ActAnywhere: Subject-Aware Video Background Generation
par: Pan, Boxiao, et autres
Publié: (2024)
par: Pan, Boxiao, et autres
Publié: (2024)
Make a Donut: Hierarchical EMD-Space Planning for Zero-Shot Deformable Manipulation with Tools
par: You, Yang, et autres
Publié: (2023)
par: You, Yang, et autres
Publié: (2023)
Fast Adversarial Training with Weak-to-Strong Spatial-Temporal Consistency in the Frequency Domain on Videos
par: Wang, Songping, et autres
Publié: (2025)
par: Wang, Songping, et autres
Publié: (2025)
Mixture of Contexts for Long Video Generation
par: Cai, Shengqu, et autres
Publié: (2025)
par: Cai, Shengqu, et autres
Publié: (2025)
FFCA-Net: Stereo Image Compression via Fast Cascade Alignment of Side Information
par: Xia, Yichong, et autres
Publié: (2023)
par: Xia, Yichong, et autres
Publié: (2023)
Documents similaires
-
View-Consistent Hierarchical 3D Segmentation Using Ultrametric Feature Fields
par: He, Haodi, et autres
Publié: (2024) -
Dynamic Gaussian Marbles for Novel View Synthesis of Casual Monocular Videos
par: Stearns, Colton, et autres
Publié: (2024) -
Dynamic Reflections: Probing Video Representations with Text Alignment
par: Zhu, Tyler, et autres
Publié: (2025) -
MoSca: Dynamic Gaussian Fusion from Casual Videos via 4D Motion Scaffolds
par: Lei, Jiahui, et autres
Publié: (2024) -
Can Test-Time Scaling Improve World Foundation Model?
par: Cong, Wenyan, et autres
Publié: (2025)