4D-LRM: Large Space-Time Reconstruction Model From and To Any View at Any Time
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Ziqiao, Chen, Xuweiyi, Yu, Shoubin, Bi, Sai, Zhang, Kai, Ziwen, Chen, Xu, Sihan, Yang, Jianing, Xu, Zexiang, Sunkavalli, Kalyan, Bansal, Mohit, Chai, Joyce, Tan, Hao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
MeshLRM: Large Reconstruction Model for High-Quality Meshes
by: Wei, Xinyue, et al.
Published: (2024)
by: Wei, Xinyue, et al.
Published: (2024)
tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
by: Wang, Chen, et al.
Published: (2026)
by: Wang, Chen, et al.
Published: (2026)
Long-LRM: Long-sequence Large Reconstruction Model for Wide-coverage Gaussian Splats
by: Ziwen, Chen, et al.
Published: (2024)
by: Ziwen, Chen, et al.
Published: (2024)
Long-LRM++: Preserving Fine Details in Feed-Forward Wide-Coverage Reconstruction
by: Ziwen, Chen, et al.
Published: (2025)
by: Ziwen, Chen, et al.
Published: (2025)
LRM: Large Reconstruction Model for Single Image to 3D
by: Hong, Yicong, et al.
Published: (2023)
by: Hong, Yicong, et al.
Published: (2023)
SAB3R: Semantic-Augmented Backbone in 3D Reconstruction
by: Chen, Xuweiyi, et al.
Published: (2025)
by: Chen, Xuweiyi, et al.
Published: (2025)
NeuManifold: Neural Watertight Manifold Reconstruction with Efficient and High-Quality Rendering Support
by: Wei, Xinyue, et al.
Published: (2023)
by: Wei, Xinyue, et al.
Published: (2023)
Neural Directional Encoding for Efficient and Accurate View-Dependent Appearance Modeling
by: Wu, Liwen, et al.
Published: (2024)
by: Wu, Liwen, et al.
Published: (2024)
Multi-Object Hallucination in Vision-Language Models
by: Chen, Xuweiyi, et al.
Published: (2024)
by: Chen, Xuweiyi, et al.
Published: (2024)
LRM-Zero: Training Large Reconstruction Models with Synthesized Data
by: Xie, Desai, et al.
Published: (2024)
by: Xie, Desai, et al.
Published: (2024)
RelitLRM: Generative Relightable Radiance for Large Reconstruction Models
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
Next-Embedding Prediction Makes Strong Vision Learners
by: Xu, Sihan, et al.
Published: (2025)
by: Xu, Sihan, et al.
Published: (2025)
Neural Gaffer: Relighting Any Object via Diffusion
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation
by: Yu, Shoubin, et al.
Published: (2025)
by: Yu, Shoubin, et al.
Published: (2025)
Test-Time Training Done Right
by: Zhang, Tianyuan, et al.
Published: (2025)
by: Zhang, Tianyuan, et al.
Published: (2025)
UniCtrl: Improving the Spatiotemporal Consistency of Text-to-Video Diffusion Models via Training-Free Unified Attention Control
by: Xia, Tian, et al.
Published: (2024)
by: Xia, Tian, et al.
Published: (2024)
PointRecon: Online Point-based 3D Reconstruction via Ray-based 2D-3D Matching
by: Ziwen, Chen, et al.
Published: (2024)
by: Ziwen, Chen, et al.
Published: (2024)
E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training
by: Zhao, Qitao, et al.
Published: (2025)
by: Zhao, Qitao, et al.
Published: (2025)
AnyView: Synthesizing Any Novel View in Dynamic Scenes
by: Van Hoorick, Basile, et al.
Published: (2026)
by: Van Hoorick, Basile, et al.
Published: (2026)
CycleNet: Rethinking Cycle Consistency in Text-Guided Diffusion for Image Manipulation
by: Xu, Sihan, et al.
Published: (2023)
by: Xu, Sihan, et al.
Published: (2023)
Structurally Prune Anything: Any Architecture, Any Framework, Any Time
by: Wang, Xun, et al.
Published: (2024)
by: Wang, Xun, et al.
Published: (2024)
RACCooN: A Versatile Instructional Video Editing Framework with Auto-Generated Narratives
by: Yoon, Jaehong, et al.
Published: (2024)
by: Yoon, Jaehong, et al.
Published: (2024)
CREMA: Generalizable and Efficient Video-Language Reasoning via Multimodal Modular Fusion
by: Yu, Shoubin, et al.
Published: (2024)
by: Yu, Shoubin, et al.
Published: (2024)
AnyRecon: Arbitrary-View 3D Reconstruction with Video Diffusion Model
by: Chen, Yutian, et al.
Published: (2026)
by: Chen, Yutian, et al.
Published: (2026)
RoboPaint: From Human Demonstration to Any Robot and Any View
by: Fan, Jiacheng, et al.
Published: (2026)
by: Fan, Jiacheng, et al.
Published: (2026)
RayZer: A Self-supervised Large View Synthesis Model
by: Jiang, Hanwen, et al.
Published: (2025)
by: Jiang, Hanwen, et al.
Published: (2025)
Ctrl-Adapter: An Efficient and Versatile Framework for Adapting Diverse Controls to Any Diffusion Model
by: Lin, Han, et al.
Published: (2024)
by: Lin, Han, et al.
Published: (2024)
Depth Anything 3: Recovering the Visual Space from Any Views
by: Lin, Haotong, et al.
Published: (2025)
by: Lin, Haotong, et al.
Published: (2025)
Any Model, Any Place, Any Time: Get Remote Sensing Foundation Model Embeddings On Demand
by: Ye, Dingqi, et al.
Published: (2026)
by: Ye, Dingqi, et al.
Published: (2026)
FreeTuner: Any Subject in Any Style with Training-free Diffusion
by: Xu, Youcan, et al.
Published: (2024)
by: Xu, Youcan, et al.
Published: (2024)
3D-GRAND: A Million-Scale Dataset for 3D-LLMs with Better Grounding and Less Hallucination
by: Yang, Jianing, et al.
Published: (2024)
by: Yang, Jianing, et al.
Published: (2024)
EmbodiedSAM: Online Segment Any 3D Thing in Real Time
by: Xu, Xiuwei, et al.
Published: (2024)
by: Xu, Xiuwei, et al.
Published: (2024)
OmniControl: Control Any Joint at Any Time for Human Motion Generation
by: Xie, Yiming, et al.
Published: (2023)
by: Xie, Yiming, et al.
Published: (2023)
Any Resolution Any Geometry: From Multi-View To Multi-Patch
by: Cui, Wenqing, et al.
Published: (2026)
by: Cui, Wenqing, et al.
Published: (2026)
Softmax-GS: Generalized Gaussians Learning When to Blend or Bound
by: Ziwen, Chen, et al.
Published: (2026)
by: Ziwen, Chen, et al.
Published: (2026)
MagicDrive3D: Controllable 3D Generation for Any-View Rendering in Street Scenes
by: Gao, Ruiyuan, et al.
Published: (2024)
by: Gao, Ruiyuan, et al.
Published: (2024)
ClipGStream: Clip-Stream Gaussian Splatting for Any Length and Any Motion Multi-View Dynamic Scene Reconstruction
by: Liang, Jie, et al.
Published: (2026)
by: Liang, Jie, et al.
Published: (2026)
Symbolic Representation for Any-to-Any Generative Tasks
by: Chen, Jiaqi, et al.
Published: (2025)
by: Chen, Jiaqi, et al.
Published: (2025)
VisionCoach: Reinforcing Grounded Video Reasoning via Visual-Perception Prompting
by: Lee, Daeun, et al.
Published: (2026)
by: Lee, Daeun, et al.
Published: (2026)
Similar Items
-
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
by: Zhang, Kai, et al.
Published: (2024) -
MeshLRM: Large Reconstruction Model for High-Quality Meshes
by: Wei, Xinyue, et al.
Published: (2024) -
tttLRM: Test-Time Training for Long Context and Autoregressive 3D Reconstruction
by: Wang, Chen, et al.
Published: (2026) -
Long-LRM: Long-sequence Large Reconstruction Model for Wide-coverage Gaussian Splats
by: Ziwen, Chen, et al.
Published: (2024) -
Long-LRM++: Preserving Fine Details in Feed-Forward Wide-Coverage Reconstruction
by: Ziwen, Chen, et al.
Published: (2025)