Ctrl123: Consistent Novel View Synthesis via Closed-Loop Transcription
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Hongxiang, Dai, Xili, Wang, Jianan, Tong, Shengbang, Zhang, Jingyuan, Wang, Weida, Zhang, Lei, Ma, Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Image Clustering via the Principle of Rate Reduction in the Age of Pretrained Models
von: Chu, Tianzhe, et al.
Veröffentlicht: (2023)
von: Chu, Tianzhe, et al.
Veröffentlicht: (2023)
Unposed Sparse Views Room Layout Reconstruction in the Age of Pretrain Model
von: Huang, Yaxuan, et al.
Veröffentlicht: (2025)
von: Huang, Yaxuan, et al.
Veröffentlicht: (2025)
GaussCtrl: Multi-View Consistent Text-Driven 3D Gaussian Splatting Editing
von: Wu, Jing, et al.
Veröffentlicht: (2024)
von: Wu, Jing, et al.
Veröffentlicht: (2024)
EmoCtrl: Controllable Emotional Image Content Generation
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025)
Cascade-Zero123: One Image to Highly Consistent 3D with Self-Prompted Nearby Views
von: Chen, Yabo, et al.
Veröffentlicht: (2023)
von: Chen, Yabo, et al.
Veröffentlicht: (2023)
Recollection from Pensieve: Novel View Synthesis via Learning from Uncalibrated Videos
von: Wang, Ruoyu, et al.
Veröffentlicht: (2025)
von: Wang, Ruoyu, et al.
Veröffentlicht: (2025)
Self-Ensembling Gaussian Splatting for Few-Shot Novel View Synthesis
von: Zhao, Chen, et al.
Veröffentlicht: (2024)
von: Zhao, Chen, et al.
Veröffentlicht: (2024)
NEMTO: Neural Environment Matting for Novel View and Relighting Synthesis of Transparent Objects
von: Wang, Dongqing, et al.
Veröffentlicht: (2023)
von: Wang, Dongqing, et al.
Veröffentlicht: (2023)
Ctrl-VI: Controllable Video Synthesis via Variational Inference
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
von: Duan, Haoyi, et al.
Veröffentlicht: (2025)
ConsistEdit: Highly Consistent and Precise Training-free Visual Editing
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
von: Yin, Zixin, et al.
Veröffentlicht: (2025)
BlobCtrl: Taming Controllable Blob for Element-level Image Editing
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
von: Li, Yaowei, et al.
Veröffentlicht: (2025)
CloseUpShot: Close-up Novel View Synthesis from Sparse-views via Point-conditioned Diffusion Model
von: Zhang, Yuqi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuqi, et al.
Veröffentlicht: (2025)
Enhancing Close-up Novel View Synthesis via Pseudo-labeling
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
von: Xia, Jiatong, et al.
Veröffentlicht: (2025)
Pointmap-Conditioned Diffusion for Consistent Novel View Synthesis
von: Nguyen, Thang-Anh-Quan, et al.
Veröffentlicht: (2025)
von: Nguyen, Thang-Anh-Quan, et al.
Veröffentlicht: (2025)
View-Consistent 3D Editing with Gaussian Splatting
von: Wang, Yuxuan, et al.
Veröffentlicht: (2024)
von: Wang, Yuxuan, et al.
Veröffentlicht: (2024)
Structure Consistent Gaussian Splatting with Matching Prior for Few-shot Novel View Synthesis
von: Peng, Rui, et al.
Veröffentlicht: (2024)
von: Peng, Rui, et al.
Veröffentlicht: (2024)
Consistent Time-of-Flight Depth Denoising via Graph-Informed Geometric Attention
von: Wang, Weida, et al.
Veröffentlicht: (2025)
von: Wang, Weida, et al.
Veröffentlicht: (2025)
Asymmetric Idiosyncrasies in Multimodal Models
von: Tao, Muzi, et al.
Veröffentlicht: (2026)
von: Tao, Muzi, et al.
Veröffentlicht: (2026)
SpikeNVS: Enhancing Novel View Synthesis from Blurry Images via Spike Camera
von: Dai, Gaole, et al.
Veröffentlicht: (2024)
von: Dai, Gaole, et al.
Veröffentlicht: (2024)
Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs
von: Yeh, Chun-Hsiao, et al.
Veröffentlicht: (2025)
von: Yeh, Chun-Hsiao, et al.
Veröffentlicht: (2025)
CMC: Few-shot Novel View Synthesis via Cross-view Multiplane Consistency
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
von: Zhu, Hanxin, et al.
Veröffentlicht: (2024)
Consistent-1-to-3: Consistent Image to 3D View Synthesis via Geometry-aware Diffusion Models
von: Ye, Jianglong, et al.
Veröffentlicht: (2023)
von: Ye, Jianglong, et al.
Veröffentlicht: (2023)
RelaCtrl: Relevance-Guided Efficient Control for Diffusion Transformers
von: Cao, Ke, et al.
Veröffentlicht: (2025)
von: Cao, Ke, et al.
Veröffentlicht: (2025)
Connecting Joint-Embedding Predictive Architecture with Contrastive Self-supervised Learning
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
NerfBaselines: Consistent and Reproducible Evaluation of Novel View Synthesis Methods
von: Kulhanek, Jonas, et al.
Veröffentlicht: (2024)
von: Kulhanek, Jonas, et al.
Veröffentlicht: (2024)
High-Fidelity Novel View Synthesis via Splatting-Guided Diffusion
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
von: Zhang, Xiang, et al.
Veröffentlicht: (2025)
From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models
von: Fang, Irving, et al.
Veröffentlicht: (2025)
von: Fang, Irving, et al.
Veröffentlicht: (2025)
EndoCogniAgent: Closed-Loop Agentic Reasoning with Self-Consistency Validation for Endoscopic Diagnosis
von: Tang, Yi, et al.
Veröffentlicht: (2025)
von: Tang, Yi, et al.
Veröffentlicht: (2025)
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
von: Wang, Yating, et al.
Veröffentlicht: (2024)
von: Wang, Yating, et al.
Veröffentlicht: (2024)
Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward Modeling
von: Zhang, Guiyu, et al.
Veröffentlicht: (2024)
von: Zhang, Guiyu, et al.
Veröffentlicht: (2024)
Diffusion Transformers with Representation Autoencoders
von: Zheng, Boyang, et al.
Veröffentlicht: (2025)
von: Zheng, Boyang, et al.
Veröffentlicht: (2025)
CtrlVDiff: Controllable Video Generation via Unified Multimodal Video Diffusion
von: Xi, Dianbing, et al.
Veröffentlicht: (2025)
von: Xi, Dianbing, et al.
Veröffentlicht: (2025)
WonderFree: Enhancing Novel View Quality and Cross-View Consistency for 3D Scene Exploration
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
von: Ni, Chaojun, et al.
Veröffentlicht: (2025)
BAFNet: Bilateral Attention Fusion Network for Lightweight Semantic Segmentation of Urban Remote Sensing Images
von: Wang, Wentao, et al.
Veröffentlicht: (2024)
von: Wang, Wentao, et al.
Veröffentlicht: (2024)
XScale-NVS: Cross-Scale Novel View Synthesis with Hash Featurized Manifold
von: Wang, Guangyu, et al.
Veröffentlicht: (2024)
von: Wang, Guangyu, et al.
Veröffentlicht: (2024)
MultiDiff: Consistent Novel View Synthesis from a Single Image
von: Müller, Norman, et al.
Veröffentlicht: (2024)
von: Müller, Norman, et al.
Veröffentlicht: (2024)
Learn Your Scales: Towards Scale-Consistent Generative Novel View Synthesis
von: Forghani, Fereshteh, et al.
Veröffentlicht: (2025)
von: Forghani, Fereshteh, et al.
Veröffentlicht: (2025)
WAVE: Warp-Based View Guidance for Consistent Novel View Synthesis Using a Single Image
von: Park, Jiwoo, et al.
Veröffentlicht: (2025)
von: Park, Jiwoo, et al.
Veröffentlicht: (2025)
Diff3DS: Generating View-Consistent 3D Sketch via Differentiable Curve Rendering
von: Zhang, Yibo, et al.
Veröffentlicht: (2024)
von: Zhang, Yibo, et al.
Veröffentlicht: (2024)
XLD: A Cross-Lane Dataset for Benchmarking Novel Driving View Synthesis
von: Li, Hao, et al.
Veröffentlicht: (2024)
von: Li, Hao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Image Clustering via the Principle of Rate Reduction in the Age of Pretrained Models
von: Chu, Tianzhe, et al.
Veröffentlicht: (2023) -
Unposed Sparse Views Room Layout Reconstruction in the Age of Pretrain Model
von: Huang, Yaxuan, et al.
Veröffentlicht: (2025) -
GaussCtrl: Multi-View Consistent Text-Driven 3D Gaussian Splatting Editing
von: Wu, Jing, et al.
Veröffentlicht: (2024) -
EmoCtrl: Controllable Emotional Image Content Generation
von: Yang, Jingyuan, et al.
Veröffentlicht: (2025) -
Cascade-Zero123: One Image to Highly Consistent 3D with Self-Prompted Nearby Views
von: Chen, Yabo, et al.
Veröffentlicht: (2023)