KFC-W: Generating 3D-Consistent Videos from Unposed Internet Photos
Fuente:
arXiv
Saved in:
| Main Authors: | Chou, Gene, Zhang, Kai, Bi, Sai, Tan, Hao, Xu, Zexiang, Luan, Fujun, Hariharan, Bharath, Snavely, Noah |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
Turbo3D: Ultra-fast Text-to-3D Generation
by: Hu, Hanzhe, et al.
Published: (2024)
by: Hu, Hanzhe, et al.
Published: (2024)
Neural Gaffer: Relighting Any Object via Diffusion
by: Jin, Haian, et al.
Published: (2024)
by: Jin, Haian, et al.
Published: (2024)
FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution
by: Chou, Gene, et al.
Published: (2025)
by: Chou, Gene, et al.
Published: (2025)
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation
by: Chou, Gene, et al.
Published: (2026)
by: Chou, Gene, et al.
Published: (2026)
MegaScenes: Scene-Level View Synthesis at Scale
by: Tung, Joseph, et al.
Published: (2024)
by: Tung, Joseph, et al.
Published: (2024)
Learning Feature Descriptors using Camera Pose Supervision
by: Wang, Qianqian, et al.
Published: (2020)
by: Wang, Qianqian, et al.
Published: (2020)
C3Po: Cross-View Cross-Modality Correspondence by Pointmap Prediction
by: Huang, Kuan Wei, et al.
Published: (2025)
by: Huang, Kuan Wei, et al.
Published: (2025)
Gaussian Mixture Flow Matching Models
by: Chen, Hansheng, et al.
Published: (2025)
by: Chen, Hansheng, et al.
Published: (2025)
Neural BRDF Importance Sampling by Reparameterization
by: Wu, Liwen, et al.
Published: (2025)
by: Wu, Liwen, et al.
Published: (2025)
Long-LRM: Long-sequence Large Reconstruction Model for Wide-coverage Gaussian Splats
by: Ziwen, Chen, et al.
Published: (2024)
by: Ziwen, Chen, et al.
Published: (2024)
Buffer Anytime: Zero-Shot Video Depth and Normal from Image Priors
by: Kuang, Zhengfei, et al.
Published: (2024)
by: Kuang, Zhengfei, et al.
Published: (2024)
MeshLRM: Large Reconstruction Model for High-Quality Meshes
by: Wei, Xinyue, et al.
Published: (2024)
by: Wei, Xinyue, et al.
Published: (2024)
G3T Up! Gravity Aligned Coordinate Frames Simplify Pointmap Processing
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026)
by: Kani, Bharath Raj Nagoor, et al.
Published: (2026)
ObjectCarver: Semi-automatic segmentation, reconstruction and separation of 3D objects
by: Hassena, Gemmechu, et al.
Published: (2024)
by: Hassena, Gemmechu, et al.
Published: (2024)
Neural Directional Encoding for Efficient and Accurate View-Dependent Appearance Modeling
by: Wu, Liwen, et al.
Published: (2024)
by: Wu, Liwen, et al.
Published: (2024)
PBIR-NIE: Glossy Object Capture under Non-Distant Lighting
by: Cai, Guangyan, et al.
Published: (2024)
by: Cai, Guangyan, et al.
Published: (2024)
EP-CFG: Energy-Preserving Classifier-Free Guidance
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
GS-LRM: Large Reconstruction Model for 3D Gaussian Splatting
by: Zhang, Kai, et al.
Published: (2024)
by: Zhang, Kai, et al.
Published: (2024)
RelitLRM: Generative Relightable Radiance for Large Reconstruction Models
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
MegaSynth: Scaling Up 3D Scene Reconstruction with Synthesized Data
by: Jiang, Hanwen, et al.
Published: (2024)
by: Jiang, Hanwen, et al.
Published: (2024)
Stereo4D: Learning How Things Move in 3D from Internet Stereo Videos
by: Jin, Linyi, et al.
Published: (2024)
by: Jin, Linyi, et al.
Published: (2024)
MOD-UV: Learning Mobile Object Detectors from Unlabeled Videos
by: Sun, Yihong, et al.
Published: (2024)
by: Sun, Yihong, et al.
Published: (2024)
Flat-Pack Bench: Evaluating Spatio-Temporal Understanding in Large Vision-Language Models through Furniture Assembly
by: Chetan, Aditya, et al.
Published: (2026)
by: Chetan, Aditya, et al.
Published: (2026)
Streetscapes: Large-scale Consistent Street View Generation Using Autoregressive Video Diffusion
by: Deng, Boyang, et al.
Published: (2024)
by: Deng, Boyang, et al.
Published: (2024)
Test-Time Training Done Right
by: Zhang, Tianyuan, et al.
Published: (2025)
by: Zhang, Tianyuan, et al.
Published: (2025)
Beyond the Frame: Generating 360 Panoramic Videos from Perspective Videos
by: Luo, Rundong, et al.
Published: (2025)
by: Luo, Rundong, et al.
Published: (2025)
LucidFusion: Reconstructing 3D Gaussians with Arbitrary Unposed Images
by: He, Hao, et al.
Published: (2024)
by: He, Hao, et al.
Published: (2024)
DATENeRF: Depth-Aware Text-based Editing of NeRFs
by: Rojas, Sara, et al.
Published: (2024)
by: Rojas, Sara, et al.
Published: (2024)
ArchSym: Detecting 3D-Grounded Architectural Symmetries in the Wild
by: Chen, Hanyu, et al.
Published: (2026)
by: Chen, Hanyu, et al.
Published: (2026)
Long-tail Internet photo reconstruction
by: Li, Yuan, et al.
Published: (2026)
by: Li, Yuan, et al.
Published: (2026)
Generating 360° Video is What You Need For a 3D Scene
by: Zhang, Zhaoyang, et al.
Published: (2025)
by: Zhang, Zhaoyang, et al.
Published: (2025)
PhysDreamer: Physics-Based Interaction with 3D Objects via Video Generation
by: Zhang, Tianyuan, et al.
Published: (2024)
by: Zhang, Tianyuan, et al.
Published: (2024)
UVRM: A Scalable 3D Reconstruction Model from Unposed Videos
by: Kao, Shiu-hong, et al.
Published: (2025)
by: Kao, Shiu-hong, et al.
Published: (2025)
Counter-Current Learning: A Biologically Plausible Dual Network Approach for Deep Learning
by: Kao, Chia-Hsiang, et al.
Published: (2024)
by: Kao, Chia-Hsiang, et al.
Published: (2024)
Generative Image Dynamics
by: Li, Zhengqi, et al.
Published: (2023)
by: Li, Zhengqi, et al.
Published: (2023)
UpFusion: Novel View Diffusion from Unposed Sparse View Observations
by: Kani, Bharath Raj Nagoor, et al.
Published: (2023)
by: Kani, Bharath Raj Nagoor, et al.
Published: (2023)
Composing People Together: Iterative Pose-Image Generation for Multi-Person Interaction Scenes
by: Peng, Wenxuan, et al.
Published: (2026)
by: Peng, Wenxuan, et al.
Published: (2026)
KFC Foundation grants target hunger initiatives
Published: (2026)
Published: (2026)
Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features
by: Xiangli, Yuanbo, et al.
Published: (2024)
by: Xiangli, Yuanbo, et al.
Published: (2024)
Similar Items
-
LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias
by: Jin, Haian, et al.
Published: (2024) -
Turbo3D: Ultra-fast Text-to-3D Generation
by: Hu, Hanzhe, et al.
Published: (2024) -
Neural Gaffer: Relighting Any Object via Diffusion
by: Jin, Haian, et al.
Published: (2024) -
FlashDepth: Real-time Streaming Video Depth Estimation at 2K Resolution
by: Chou, Gene, et al.
Published: (2025) -
CityRAG: Stepping Into a City via Spatially-Grounded Video Generation
by: Chou, Gene, et al.
Published: (2026)