MVCustom: Multi-View Customized Diffusion via Geometric Latent Rendering and Completion
Fuente:
arXiv
Saved in:
| Main Authors: | Shin, Minjung, Cho, Hyunin, Go, Sooyeon, Kim, Jin-Hwa, Uh, Youngjung |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Eye-for-an-eye: Appearance Transfer with Semantic Correspondence in Diffusion Models
by: Go, Sooyeon, et al.
Published: (2024)
by: Go, Sooyeon, et al.
Published: (2024)
ASemConsist: Adaptive Semantic Feature Control for Training-Free Identity-Consistent Generation
by: Kim, Shin Seong, et al.
Published: (2025)
by: Kim, Shin Seong, et al.
Published: (2025)
Semantic Image Synthesis with Unconditional Generator
by: Chae, Jungwoo, et al.
Published: (2024)
by: Chae, Jungwoo, et al.
Published: (2024)
TetraSDF: Precise Mesh Extraction with Multi-resolution Tetrahedral Grid
by: Oh, Seonghun, et al.
Published: (2025)
by: Oh, Seonghun, et al.
Published: (2025)
Geometric Disentanglement of Text Embeddings for Subject-Consistent Text-to-Image Generation using A Single Prompt
by: Li, Shangxun, et al.
Published: (2025)
by: Li, Shangxun, et al.
Published: (2025)
Attribute Based Interpretable Evaluation Metrics for Generative Models
by: Kim, Dongkyun, et al.
Published: (2023)
by: Kim, Dongkyun, et al.
Published: (2023)
Training-free Content Injection using h-space in Diffusion Models
by: Jeong, Jaeseok, et al.
Published: (2023)
by: Jeong, Jaeseok, et al.
Published: (2023)
Syncphony: Synchronized Audio-to-Video Generation with Diffusion Transformers
by: Song, Jibin, et al.
Published: (2025)
by: Song, Jibin, et al.
Published: (2025)
Frequency-Adaptive Sharpness Regularization for Improving 3D Gaussian Splatting Generalization
by: Yun, Youngsik, et al.
Published: (2025)
by: Yun, Youngsik, et al.
Published: (2025)
FlowBlending: Stage-Aware Multi-Model Sampling for Fast and High-Fidelity Video Generation
by: Song, Jibin, et al.
Published: (2025)
by: Song, Jibin, et al.
Published: (2025)
4D Scaffold Gaussian Splatting with Dynamic-Aware Anchor Growing for Efficient and High-Fidelity Dynamic Scene Reconstruction
by: Cho, Woong Oh, et al.
Published: (2024)
by: Cho, Woong Oh, et al.
Published: (2024)
Bridging Implicit and Explicit Geometric Transformation for Single-Image View Synthesis
by: Park, Byeongjun, et al.
Published: (2022)
by: Park, Byeongjun, et al.
Published: (2022)
TCFG: Tangential Damping Classifier-free Guidance
by: Kwon, Mingi, et al.
Published: (2025)
by: Kwon, Mingi, et al.
Published: (2025)
StyleKeeper: Prevent Content Leakage using Negative Visual Query Guidance
by: Jeong, Jaeseok, et al.
Published: (2025)
by: Jeong, Jaeseok, et al.
Published: (2025)
Visual Style Prompting with Swapping Self-Attention
by: Jeong, Jaeseok, et al.
Published: (2024)
by: Jeong, Jaeseok, et al.
Published: (2024)
Rethinking Open-Vocabulary Segmentation of Radiance Fields in 3D Space
by: Lee, Hyunjee, et al.
Published: (2024)
by: Lee, Hyunjee, et al.
Published: (2024)
JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching
by: Kwon, Mingi, et al.
Published: (2025)
by: Kwon, Mingi, et al.
Published: (2025)
Per-Gaussian Embedding-Based Deformation for Deformable 3D Gaussian Splatting
by: Bae, Jeongmin, et al.
Published: (2024)
by: Bae, Jeongmin, et al.
Published: (2024)
Sync-NeRF: Generalizing Dynamic NeRFs to Unsynchronized Videos
by: Kim, Seoha, et al.
Published: (2023)
by: Kim, Seoha, et al.
Published: (2023)
Diffusion Model Patching via Mixture-of-Prompts
by: Ham, Seokil, et al.
Published: (2024)
by: Ham, Seokil, et al.
Published: (2024)
ArtSplat: Feed-Forward Articulated 3D Gaussian Splatting from Sparse Multi-State Uncalibrated Views
by: Lee, Inseo, et al.
Published: (2026)
by: Lee, Inseo, et al.
Published: (2026)
Virtually Being: Customizing Camera-Controllable Video Diffusion Models with Multi-View Performance Captures
by: Xu, Yuancheng, et al.
Published: (2025)
by: Xu, Yuancheng, et al.
Published: (2025)
Denoising Task Routing for Diffusion Models
by: Park, Byeongjun, et al.
Published: (2023)
by: Park, Byeongjun, et al.
Published: (2023)
Compensating Spatiotemporally Inconsistent Observations for Online Dynamic 3D Gaussian Splatting
by: Yun, Youngsik, et al.
Published: (2025)
by: Yun, Youngsik, et al.
Published: (2025)
HARIVO: Harnessing Text-to-Image Models for Video Generation
by: Kwon, Mingi, et al.
Published: (2024)
by: Kwon, Mingi, et al.
Published: (2024)
Balanced conic rectified flow
by: Kim, Shin Seong, et al.
Published: (2025)
by: Kim, Shin Seong, et al.
Published: (2025)
V.I.P. : Iterative Online Preference Distillation for Efficient Video Diffusion Models
by: Kim, Jisoo, et al.
Published: (2025)
by: Kim, Jisoo, et al.
Published: (2025)
Fine-Grained Multi-View Hand Reconstruction Using Inverse Rendering
by: Gan, Qijun, et al.
Published: (2024)
by: Gan, Qijun, et al.
Published: (2024)
Representing 3D Shapes With 64 Latent Vectors for 3D Diffusion Models
by: Cho, In, et al.
Published: (2025)
by: Cho, In, et al.
Published: (2025)
DreamMakeup: Face Makeup Customization using Latent Diffusion Models
by: Park, Geon Yeong, et al.
Published: (2025)
by: Park, Geon Yeong, et al.
Published: (2025)
CAMEO: Correspondence-Attention Alignment for Multi-View Diffusion Models
by: Kwon, Minkyung, et al.
Published: (2025)
by: Kwon, Minkyung, et al.
Published: (2025)
FLoD: Integrating Flexible Level of Detail into 3D Gaussian Splatting for Customizable Rendering
by: Seo, Yunji, et al.
Published: (2024)
by: Seo, Yunji, et al.
Published: (2024)
VideoRFSplat: Direct Scene-Level Text-to-3D Gaussian Splatting Generation with Flexible Pose and Multi-View Joint Modeling
by: Go, Hyojun, et al.
Published: (2025)
by: Go, Hyojun, et al.
Published: (2025)
Polyhedral Complex Derivation from Piecewise Trilinear Networks
by: Kim, Jin-Hwa
Published: (2024)
by: Kim, Jin-Hwa
Published: (2024)
GMapLatent: Geometric Mapping in Latent Space
by: Zeng, Wei, et al.
Published: (2025)
by: Zeng, Wei, et al.
Published: (2025)
Dual Recursive Feedback on Generation and Appearance Latents for Pose-Robust Text-to-Image Diffusion
by: Kim, Jiwon, et al.
Published: (2025)
by: Kim, Jiwon, et al.
Published: (2025)
Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models
by: Liu, Huijie, et al.
Published: (2025)
by: Liu, Huijie, et al.
Published: (2025)
CogME: A Cognition-Inspired Multi-Dimensional Evaluation Metric for Story Understanding
by: Shin, Minjung, et al.
Published: (2021)
by: Shin, Minjung, et al.
Published: (2021)
Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
by: Ahn, Donghoon, et al.
Published: (2024)
by: Ahn, Donghoon, et al.
Published: (2024)
Generic Event Boundary Detection via Denoising Diffusion
by: Hwang, Jaejun, et al.
Published: (2025)
by: Hwang, Jaejun, et al.
Published: (2025)
Similar Items
-
Eye-for-an-eye: Appearance Transfer with Semantic Correspondence in Diffusion Models
by: Go, Sooyeon, et al.
Published: (2024) -
ASemConsist: Adaptive Semantic Feature Control for Training-Free Identity-Consistent Generation
by: Kim, Shin Seong, et al.
Published: (2025) -
Semantic Image Synthesis with Unconditional Generator
by: Chae, Jungwoo, et al.
Published: (2024) -
TetraSDF: Precise Mesh Extraction with Multi-resolution Tetrahedral Grid
by: Oh, Seonghun, et al.
Published: (2025) -
Geometric Disentanglement of Text Embeddings for Subject-Consistent Text-to-Image Generation using A Single Prompt
by: Li, Shangxun, et al.
Published: (2025)