Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Kuang, Zhengfei, Zhang, Tianyuan, Zhang, Kai, Tan, Hao, Bi, Sai, Hu, Yiwei, Xu, Zexiang, Hasan, Milos, Wetzstein, Gordon, Luan, Fujun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RelitLRM: Generative Relightable Radiance for Large Reconstruction Models
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
Gaussian Mixture Flow Matching Models
by: Chen, Hansheng, et al.
Published: (2025)
by: Chen, Hansheng, et al.
Published: (2025)
PBIR-NIE: Glossy Object Capture under Non-Distant Lighting
by: Cai, Guangyan, et al.
Published: (2024)
by: Cai, Guangyan, et al.
Published: (2024)
LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
by: Chou, Gene, et al.
Published: (2024)
by: Chou, Gene, et al.
Published: (2024)
Turbo3D: Ultra-fast Text-to-3D Generation
by: Hu, Hanzhe, et al.
Published: (2024)
by: Hu, Hanzhe, et al.
Published: (2024)
Position-Normal Manifold for Efficient Glint Rendering on High-Resolution Normal Maps
by: Wu, Liwen, et al.
Published: (2025)
by: Wu, Liwen, et al.
Published: (2025)
Neural BRDF Importance Sampling by Reparameterization
by: Wu, Liwen, et al.
Published: (2025)
by: Wu, Liwen, et al.
Published: (2025)
Long-LRM: Long-sequence Large Reconstruction Model for Wide-coverage Gaussian Splats
by: Ziwen, Chen, et al.
Published: (2024)
by: Ziwen, Chen, et al.
Published: (2024)
EP-CFG: Energy-Preserving Classifier-Free Guidance
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
MeshLRM: Large Reconstruction Model for High-Quality Meshes
by: Wei, Xinyue, et al.
Published: (2024)
by: Wei, Xinyue, et al.
Published: (2024)
Test-Time Training Done Right
by: Zhang, Tianyuan, et al.
Published: (2025)
by: Zhang, Tianyuan, et al.
Published: (2025)
pi-Flow: Policy-Based Few-Step Generation via Imitation Distillation
by: Chen, Hansheng, et al.
Published: (2025)
by: Chen, Hansheng, et al.
Published: (2025)
GeoFlow: Enforcing Implicit Geometric Consistency in Video Generation
by: Ackermann, Jan, et al.
Published: (2026)
by: Ackermann, Jan, et al.
Published: (2026)
Collaborative Video Diffusion: Consistent Multi-video Generation with Camera Control
by: Kuang, Zhengfei, et al.
Published: (2024)
by: Kuang, Zhengfei, et al.
Published: (2024)
Generating 360° Video is What You Need For a 3D Scene
by: Zhang, Zhaoyang, et al.
Published: (2025)
by: Zhang, Zhaoyang, et al.
Published: (2025)
Neural Directional Encoding for Efficient and Accurate View-Dependent Appearance Modeling
by: Wu, Liwen, et al.
Published: (2024)
by: Wu, Liwen, et al.
Published: (2024)
DATENeRF: Depth-Aware Text-based Editing of NeRFs
by: Rojas, Sara, et al.
Published: (2024)
by: Rojas, Sara, et al.
Published: (2024)
Neural Gaffer: Relighting Any Object via Diffusion
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
RNA: Relightable Neural Assets
by: Mullia, Krishna, et al.
Published: (2023)
by: Mullia, Krishna, et al.
Published: (2023)
VULCAN: Tool-Augmented Multi Agents for Iterative 3D Object Arrangement
by: Kuang, Zhengfei, et al.
Published: (2025)
by: Kuang, Zhengfei, et al.
Published: (2025)
MaterialPicker: Multi-Modal DiT-Based Material Generation
by: Ma, Xiaohe, et al.
Published: (2024)
by: Ma, Xiaohe, et al.
Published: (2024)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
by: Cai, Shengqu, et al.
Published: (2024)
by: Cai, Shengqu, et al.
Published: (2024)
LRM-Zero: Training Large Reconstruction Models with Synthesized Data
by: Xie, Desai, et al.
Published: (2024)
by: Xie, Desai, et al.
Published: (2024)
RandAR: Decoder-only Autoregressive Visual Generation in Random Orders
by: Pang, Ziqi, et al.
Published: (2024)
by: Pang, Ziqi, et al.
Published: (2024)
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
BulletTime: Decoupled Control of Time and Camera Pose for Video Generation
by: Wang, Yiming, et al.
Published: (2025)
by: Wang, Yiming, et al.
Published: (2025)
RNG: Relightable Neural Gaussians
by: Fan, Jiahui, et al.
Published: (2024)
by: Fan, Jiahui, et al.
Published: (2024)
Rendering Participating Media Using Path Graphs
by: Hu, Becky, et al.
Published: (2024)
by: Hu, Becky, et al.
Published: (2024)
Neural Product Importance Sampling via Warp Composition
by: Litalien, Joey, et al.
Published: (2024)
by: Litalien, Joey, et al.
Published: (2024)
Free Your Hands: Lightweight Turntable-Based Object Capture Pipeline
by: Fan, Jiahui, et al.
Published: (2025)
by: Fan, Jiahui, et al.
Published: (2025)
Envision: Embodied Visual Planning via Goal-Imagery Video Diffusion
by: Gu, Yuming, et al.
Published: (2025)
by: Gu, Yuming, et al.
Published: (2025)
MegaSynth: Scaling Up 3D Scene Reconstruction with Synthesized Data
by: Jiang, Hanwen, et al.
Published: (2024)
by: Jiang, Hanwen, et al.
Published: (2024)
RGB$\leftrightarrow$X: Image decomposition and synthesis using material- and lighting-aware diffusion models
by: Zeng, Zheng, et al.
Published: (2024)
by: Zeng, Zheng, et al.
Published: (2024)
H-OmniStereo: Zero-Shot Omnidirectional Stereo Matching with Heading-Aligned Normal Priors
by: Jiang, Chenxing, et al.
Published: (2026)
by: Jiang, Chenxing, et al.
Published: (2026)
AlignTok: Aligning Visual Foundation Encoders to Tokenizers for Diffusion Models
by: Chen, Bowei, et al.
Published: (2025)
by: Chen, Bowei, et al.
Published: (2025)
RayZer: A Self-supervised Large View Synthesis Model
by: Jiang, Hanwen, et al.
Published: (2025)
by: Jiang, Hanwen, et al.
Published: (2025)
Uncertainty for SVBRDF Acquisition using Frequency Analysis
by: Wiersma, Ruben, et al.
Published: (2024)
by: Wiersma, Ruben, et al.
Published: (2024)
SplatPainter: Interactive Authoring of 3D Gaussians from 2D Edits via Test-Time Training
by: Zheng, Yang, et al.
Published: (2025)
by: Zheng, Yang, et al.
Published: (2025)
Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models
by: Zhang, Lvmin, et al.
Published: (2025)
by: Zhang, Lvmin, et al.
Published: (2025)
Similar Items
-
RelitLRM: Generative Relightable Radiance for Large Reconstruction Models
by: Zhang, Tianyuan, et al.
Published: (2024) -
Gaussian Mixture Flow Matching Models
by: Chen, Hansheng, et al.
Published: (2025) -
PBIR-NIE: Glossy Object Capture under Non-Distant Lighting
by: Cai, Guangyan, et al.
Published: (2024) -
LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
by: Jin, Haian, et al.
Published: (2024) -
KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
by: Chou, Gene, et al.
Published: (2024)