MV-Adapter: Multi-view Consistent Image Generation Made Easy
Fuente:
arXiv
Salvato in:
| Autori principali: | Huang, Zehuan, Guo, Yuan-Chen, Wang, Haoran, Yi, Ran, Ma, Lizhuang, Cao, Yan-Pei, Sheng, Lu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SuperMat: Physically Consistent PBR Material Estimation at Interactive Rates
di: Hong, Yijia, et al.
Pubblicazione: (2024)
di: Hong, Yijia, et al.
Pubblicazione: (2024)
Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting
di: Lu, Guangben, et al.
Pubblicazione: (2024)
di: Lu, Guangben, et al.
Pubblicazione: (2024)
Consistency Models Made Easy
di: Geng, Zhengyang, et al.
Pubblicazione: (2024)
di: Geng, Zhengyang, et al.
Pubblicazione: (2024)
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
di: Wang, Yating, et al.
Pubblicazione: (2024)
di: Wang, Yating, et al.
Pubblicazione: (2024)
High-Resolution Single-Shot Polarimetric Imaging Made Easy
di: Zhou, Shuangfan, et al.
Pubblicazione: (2026)
di: Zhou, Shuangfan, et al.
Pubblicazione: (2026)
UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation
di: Sun, Yang-Tian, et al.
Pubblicazione: (2025)
di: Sun, Yang-Tian, et al.
Pubblicazione: (2025)
AnimaX: Animating the Inanimate in 3D with Joint Video-Pose Diffusion Models
di: Huang, Zehuan, et al.
Pubblicazione: (2025)
di: Huang, Zehuan, et al.
Pubblicazione: (2025)
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
di: Liu, Fengqi, et al.
Pubblicazione: (2024)
From Parts to Whole: A Unified Reference Framework for Controllable Human Image Generation
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
di: Huang, Zehuan, et al.
Pubblicazione: (2024)
Repurposing 3D Generative Model for Autoregressive Layout Generation
di: Feng, Haoran, et al.
Pubblicazione: (2026)
di: Feng, Haoran, et al.
Pubblicazione: (2026)
Personalize Anything for Free with Diffusion Transformer
di: Feng, Haoran, et al.
Pubblicazione: (2025)
di: Feng, Haoran, et al.
Pubblicazione: (2025)
Ouroboros3D: Image-to-3D Generation via 3D-aware Recursive Diffusion
di: Wen, Hao, et al.
Pubblicazione: (2024)
di: Wen, Hao, et al.
Pubblicazione: (2024)
AdR-Gaussian: Accelerating Gaussian Splatting with Adaptive Radius
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
di: Wang, Xinzhe, et al.
Pubblicazione: (2024)
VideoMV: Consistent Multi-View Generation Based on Large Video Generative Model
di: Zuo, Qi, et al.
Pubblicazione: (2024)
di: Zuo, Qi, et al.
Pubblicazione: (2024)
MV-ROPE: Multi-view Constraints for Robust Category-level Object Pose and Size Estimation
di: Yang, Jiaqi, et al.
Pubblicazione: (2023)
di: Yang, Jiaqi, et al.
Pubblicazione: (2023)
CryoFastAR: Fast Cryo-EM Ab Initio Reconstruction Made Easy
di: Zhang, Jiakai, et al.
Pubblicazione: (2025)
di: Zhang, Jiakai, et al.
Pubblicazione: (2025)
Stereo World Model: Camera-Guided Stereo Video Generation
di: Sun, Yang-Tian, et al.
Pubblicazione: (2026)
di: Sun, Yang-Tian, et al.
Pubblicazione: (2026)
P3P Made Easy
di: Lee, Seong Hun, et al.
Pubblicazione: (2025)
di: Lee, Seong Hun, et al.
Pubblicazione: (2025)
DragMesh: Interactive 3D Generation Made Easy
di: Zhang, Tianshan, et al.
Pubblicazione: (2025)
di: Zhang, Tianshan, et al.
Pubblicazione: (2025)
Collaborative Face Experts Fusion in Video Generation: Boosting Identity Consistency Across Large Face Poses
di: Wang, Yuji, et al.
Pubblicazione: (2025)
di: Wang, Yuji, et al.
Pubblicazione: (2025)
I2V-Adapter: A General Image-to-Video Adapter for Diffusion Models
di: Guo, Xun, et al.
Pubblicazione: (2023)
di: Guo, Xun, et al.
Pubblicazione: (2023)
UNIC-Adapter: Unified Image-instruction Adapter with Multi-modal Transformer for Image Generation
di: Duan, Lunhao, et al.
Pubblicazione: (2024)
di: Duan, Lunhao, et al.
Pubblicazione: (2024)
VoxHammer: Training-Free Precise and Coherent 3D Editing in Native 3D Space
di: Li, Lin, et al.
Pubblicazione: (2025)
di: Li, Lin, et al.
Pubblicazione: (2025)
DrivingForward: Feed-forward 3D Gaussian Splatting for Driving Scene Reconstruction from Flexible Surround-view Input
di: Tian, Qijian, et al.
Pubblicazione: (2024)
di: Tian, Qijian, et al.
Pubblicazione: (2024)
PoseAnything: Universal Pose-guided Video Generation with Part-aware Temporal Coherence
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
di: Wang, Ruiyan, et al.
Pubblicazione: (2025)
IAR2: Improving Autoregressive Visual Generation with Semantic-Detail Associated Token Prediction
di: Yi, Ran, et al.
Pubblicazione: (2025)
di: Yi, Ran, et al.
Pubblicazione: (2025)
AnyDepth: Depth Estimation Made Easy
di: Ren, Zeyu, et al.
Pubblicazione: (2026)
di: Ren, Zeyu, et al.
Pubblicazione: (2026)
Dual-End Consistency Model
di: Dong, Linwei, et al.
Pubblicazione: (2026)
di: Dong, Linwei, et al.
Pubblicazione: (2026)
UniMLVG: Unified Framework for Multi-view Long Video Generation with Comprehensive Control Capabilities for Autonomous Driving
di: Chen, Rui, et al.
Pubblicazione: (2024)
di: Chen, Rui, et al.
Pubblicazione: (2024)
Continuous Piecewise-Affine Based Motion Model for Image Animation
di: Wang, Hexiang, et al.
Pubblicazione: (2024)
di: Wang, Hexiang, et al.
Pubblicazione: (2024)
PMNI: Pose-free Multi-view Normal Integration for Reflective and Textureless Surface Reconstruction
di: Pei, Mingzhi, et al.
Pubblicazione: (2025)
di: Pei, Mingzhi, et al.
Pubblicazione: (2025)
InstanceV: Instance-Level Video Generation
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
MV-SAM: Multi-view Promptable Segmentation using Pointmap Guidance
di: Jeong, Yoonwoo, et al.
Pubblicazione: (2026)
di: Jeong, Yoonwoo, et al.
Pubblicazione: (2026)
Direct3D-S2: Gigascale 3D Generation Made Easy with Spatial Sparse Attention
di: Wu, Shuang, et al.
Pubblicazione: (2025)
di: Wu, Shuang, et al.
Pubblicazione: (2025)
BEV-VAE: Multi-view Image Generation with Spatial Consistency for Autonomous Driving
di: Chen, Zeming, et al.
Pubblicazione: (2025)
di: Chen, Zeming, et al.
Pubblicazione: (2025)
MV-S2V: Multi-View Subject-Consistent Video Generation
di: Song, Ziyang, et al.
Pubblicazione: (2026)
di: Song, Ziyang, et al.
Pubblicazione: (2026)
MV-Swin-T: Mammogram Classification with Multi-view Swin Transformer
di: Sarker, Sushmita, et al.
Pubblicazione: (2024)
di: Sarker, Sushmita, et al.
Pubblicazione: (2024)
ID-Sculpt: ID-aware 3D Head Generation from Single In-the-wild Portrait Image
di: Hao, Jinkun, et al.
Pubblicazione: (2024)
di: Hao, Jinkun, et al.
Pubblicazione: (2024)
EpiDiff: Enhancing Multi-View Synthesis via Localized Epipolar-Constrained Diffusion
di: Huang, Zehuan, et al.
Pubblicazione: (2023)
di: Huang, Zehuan, et al.
Pubblicazione: (2023)
Documenti analoghi
-
SuperMat: Physically Consistent PBR Material Estimation at Interactive Rates
di: Hong, Yijia, et al.
Pubblicazione: (2024) -
Pinco: Position-induced Consistent Adapter for Diffusion Transformer in Foreground-conditioned Inpainting
di: Lu, Guangben, et al.
Pubblicazione: (2024) -
Consistency Models Made Easy
di: Geng, Zhengyang, et al.
Pubblicazione: (2024) -
MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation
di: Huang, Zehuan, et al.
Pubblicazione: (2024) -
Reconstructing Topology-Consistent Face Mesh by Volume Rendering from Multi-View Images
di: Wang, Yating, et al.
Pubblicazione: (2024)