GenFusion: Closing the Loop between Reconstruction and Generation via Videos
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Sibo, Xu, Congrong, Huang, Binbin, Geiger, Andreas, Chen, Anpei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LaRa: Efficient Large-Baseline Radiance Fields
by: Chen, Anpei, et al.
Published: (2024)
by: Chen, Anpei, et al.
Published: (2024)
GenFusion: Feed-forward Human Performance Capture via Progressive Canonical Space Updates
by: Kwon, Youngjoong, et al.
Published: (2026)
by: Kwon, Youngjoong, et al.
Published: (2026)
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
by: Huang, Haiwen, et al.
Published: (2025)
by: Huang, Haiwen, et al.
Published: (2025)
$R^3$: 3D Reconstruction via Relative Regression
by: Xu, Congrong, et al.
Published: (2026)
by: Xu, Congrong, et al.
Published: (2026)
Animating the Past: Reconstruct Trilobite via Video Generation
by: Wu, Xiaoran, et al.
Published: (2024)
by: Wu, Xiaoran, et al.
Published: (2024)
2D Gaussian Splatting for Geometrically Accurate Radiance Fields
by: Huang, Binbin, et al.
Published: (2024)
by: Huang, Binbin, et al.
Published: (2024)
Modeling Depth Ambiguity: A Mixture-Density Representation for Flying-Point-Free Depth Estimation
by: Bian, Siyuan, et al.
Published: (2026)
by: Bian, Siyuan, et al.
Published: (2026)
NeLF-Pro: Neural Light Field Probes for Multi-Scale Novel View Synthesis
by: You, Zinuo, et al.
Published: (2023)
by: You, Zinuo, et al.
Published: (2023)
Unlocking Complex Visual Generation via Closed-Loop Verified Reasoning
by: Cheng, Hanbo, et al.
Published: (2026)
by: Cheng, Hanbo, et al.
Published: (2026)
TTT3R: 3D Reconstruction as Test-Time Training
by: Chen, Xingyu, et al.
Published: (2025)
by: Chen, Xingyu, et al.
Published: (2025)
ReconViaGen: Towards Accurate Multi-view 3D Object Reconstruction via Generation
by: Chang, Jiahao, et al.
Published: (2025)
by: Chang, Jiahao, et al.
Published: (2025)
DC-VideoGen: Efficient Video Generation with Deep Compression Video Autoencoder
by: Chen, Junyu, et al.
Published: (2025)
by: Chen, Junyu, et al.
Published: (2025)
MedGen: Unlocking Medical Video Generation by Scaling Granularly-annotated Medical Videos
by: Wang, Rongsheng, et al.
Published: (2025)
by: Wang, Rongsheng, et al.
Published: (2025)
Physical Simulator In-the-Loop Video Generation
by: Foo, Lin Geng, et al.
Published: (2026)
by: Foo, Lin Geng, et al.
Published: (2026)
ConeGS: Error-Guided Densification Using Pixel Cones for Improved Reconstruction With Fewer Primitives
by: Baranowski, Bartłomiej, et al.
Published: (2025)
by: Baranowski, Bartłomiej, et al.
Published: (2025)
GameGen-X: Interactive Open-world Game Video Generation
by: Che, Haoxuan, et al.
Published: (2024)
by: Che, Haoxuan, et al.
Published: (2024)
DriveGenVLM: Real-world Video Generation for Vision Language Model based Autonomous Driving
by: Fu, Yongjie, et al.
Published: (2024)
by: Fu, Yongjie, et al.
Published: (2024)
GenDDS: Generating Diverse Driving Video Scenarios with Prompt-to-Video Generative Model
by: Fu, Yongjie, et al.
Published: (2024)
by: Fu, Yongjie, et al.
Published: (2024)
NovaPlan: Zero-Shot Long-Horizon Manipulation via Closed-Loop Video Language Planning
by: Fu, Jiahui, et al.
Published: (2026)
by: Fu, Jiahui, et al.
Published: (2026)
GenShield: Unified Detection and Artifact Correction for AI-Generated Images
by: Xu, Zhipei, et al.
Published: (2026)
by: Xu, Zhipei, et al.
Published: (2026)
RainFusion: Adaptive Video Generation Acceleration via Multi-Dimensional Visual Redundancy
by: Chen, Aiyue, et al.
Published: (2025)
by: Chen, Aiyue, et al.
Published: (2025)
Fail2Drive: Benchmarking Closed-Loop Driving Generalization
by: Gerstenecker, Simon, et al.
Published: (2026)
by: Gerstenecker, Simon, et al.
Published: (2026)
Practical Applications of Advanced Cloud Services and Generative AI Systems in Medical Image Analysis
by: Xu, Jingyu, et al.
Published: (2024)
by: Xu, Jingyu, et al.
Published: (2024)
Taming Hallucinations: Boosting MLLMs' Video Understanding via Counterfactual Video Generation
by: Huang, Zhe, et al.
Published: (2025)
by: Huang, Zhe, et al.
Published: (2025)
RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset
by: Wang, Yongzhong, et al.
Published: (2026)
by: Wang, Yongzhong, et al.
Published: (2026)
Gen4Gen: Generative Data Pipeline for Generative Multi-Concept Composition
by: Yeh, Chun-Hsiao, et al.
Published: (2024)
by: Yeh, Chun-Hsiao, et al.
Published: (2024)
VideoGen-of-Thought: Step-by-step generating multi-shot video with minimal manual intervention
by: Zheng, Mingzhe, et al.
Published: (2024)
by: Zheng, Mingzhe, et al.
Published: (2024)
UniFS: Unified Multi-Contrast MRI Reconstruction via Frequency-Spatial Fusion
by: Li, Jialin, et al.
Published: (2025)
by: Li, Jialin, et al.
Published: (2025)
RealGen: Photorealistic Text-to-Image Generation via Detector-Guided Rewards
by: Ye, Junyan, et al.
Published: (2025)
by: Ye, Junyan, et al.
Published: (2025)
GeoGen: Geometry-Aware Generative Modeling via Signed Distance Functions
by: Esposito, Salvatore, et al.
Published: (2024)
by: Esposito, Salvatore, et al.
Published: (2024)
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
by: Chen, Xinyu, et al.
Published: (2026)
by: Chen, Xinyu, et al.
Published: (2026)
OmniGen: Unified Image Generation
by: Xiao, Shitao, et al.
Published: (2024)
by: Xiao, Shitao, et al.
Published: (2024)
xGen-VideoSyn-1: High-fidelity Text-to-Video Synthesis with Compressed Representations
by: Qin, Can, et al.
Published: (2024)
by: Qin, Can, et al.
Published: (2024)
Nexus-Gen: Unified Image Understanding, Generation, and Editing via Prefilled Autoregression in Shared Embedding Space
by: Zhang, Hong, et al.
Published: (2025)
by: Zhang, Hong, et al.
Published: (2025)
Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation
by: Jeong, Hyeonho, et al.
Published: (2024)
by: Jeong, Hyeonho, et al.
Published: (2024)
BrandFusion: A Multi-Agent Framework for Seamless Brand Integration in Text-to-Video Generation
by: Zhu, Zihao, et al.
Published: (2026)
by: Zhu, Zihao, et al.
Published: (2026)
LLM-Assist: Enhancing Closed-Loop Planning with Language-Based Reasoning
by: Sharan, S P, et al.
Published: (2023)
by: Sharan, S P, et al.
Published: (2023)
TeaserGen: Generating Teasers for Long Documentaries
by: Xu, Weihan, et al.
Published: (2024)
by: Xu, Weihan, et al.
Published: (2024)
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
by: Xu, Xinli, et al.
Published: (2024)
by: Xu, Xinli, et al.
Published: (2024)
Synergistic Global-space Camera and Human Reconstruction from Videos
by: Zhao, Yizhou, et al.
Published: (2024)
by: Zhao, Yizhou, et al.
Published: (2024)
Similar Items
-
LaRa: Efficient Large-Baseline Radiance Fields
by: Chen, Anpei, et al.
Published: (2024) -
GenFusion: Feed-forward Human Performance Capture via Progressive Canonical Space Updates
by: Kwon, Youngjoong, et al.
Published: (2026) -
LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models
by: Huang, Haiwen, et al.
Published: (2025) -
$R^3$: 3D Reconstruction via Relative Regression
by: Xu, Congrong, et al.
Published: (2026) -
Animating the Past: Reconstruct Trilobite via Video Generation
by: Wu, Xiaoran, et al.
Published: (2024)