Improving Virtual Try-On with Garment-focused Diffusion Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wan, Siqi, Li, Yehao, Chen, Jingwen, Pan, Yingwei, Yao, Ting, Cao, Yang, Mei, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
von: Wan, Siqi, et al.
Veröffentlicht: (2025)
Improving Text-guided Object Inpainting with Semantic Pre-inpainting
von: Chen, Yifu, et al.
Veröffentlicht: (2024)
von: Chen, Yifu, et al.
Veröffentlicht: (2024)
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs
von: Yao, Ting, et al.
Veröffentlicht: (2024)
von: Yao, Ting, et al.
Veröffentlicht: (2024)
Boosting Diffusion Models with Moving Average Sampling in Frequency Domain
von: Qian, Yurui, et al.
Veröffentlicht: (2024)
von: Qian, Yurui, et al.
Veröffentlicht: (2024)
SD-DiT: Unleashing the Power of Self-supervised Discrimination in Diffusion Transformer
von: Zhu, Rui, et al.
Veröffentlicht: (2024)
von: Zhu, Rui, et al.
Veröffentlicht: (2024)
Pursuing Temporal-Consistent Video Virtual Try-On via Dynamic Pose Interaction
von: Li, Dong, et al.
Veröffentlicht: (2025)
von: Li, Dong, et al.
Veröffentlicht: (2025)
Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
von: Zheng, Guangting, et al.
Veröffentlicht: (2025)
Visual Autoregressive Modeling for Instruction-Guided Image Editing
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
von: Mao, Qingyang, et al.
Veröffentlicht: (2025)
TRIP: Temporal Residual Learning with Image Noise Prior for Image-to-Video Diffusion Models
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2024)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2024)
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
von: Luo, Jianjie, et al.
Veröffentlicht: (2024)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
von: Chen, Yang, et al.
Veröffentlicht: (2024)
von: Chen, Yang, et al.
Veröffentlicht: (2024)
Hi3D: Pursuing High-Resolution Image-to-3D Generation with Video Diffusion Models
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
Creatively Upscaling Images with Global-Regional Priors
von: Qian, Yurui, et al.
Veröffentlicht: (2025)
von: Qian, Yurui, et al.
Veröffentlicht: (2025)
HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer
von: Cai, Qi, et al.
Veröffentlicht: (2025)
von: Cai, Qi, et al.
Veröffentlicht: (2025)
DreamMesh: Jointly Manipulating and Texturing Triangle Meshes for Text-to-3D Generation
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
von: Yang, Haibo, et al.
Veröffentlicht: (2024)
MotionPro: A Precise Motion Controller for Image-to-Video Generation
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer
von: Cai, Qi, et al.
Veröffentlicht: (2026)
von: Cai, Qi, et al.
Veröffentlicht: (2026)
Learning Spatial Adaptation and Temporal Coherence in Diffusion Models for Video Super-Resolution
von: Chen, Zhikai, et al.
Veröffentlicht: (2024)
von: Chen, Zhikai, et al.
Veröffentlicht: (2024)
FreeEnhance: Tuning-Free Image Enhancement via Content-Consistent Noising-and-Denoising Process
von: Luo, Yang, et al.
Veröffentlicht: (2024)
von: Luo, Yang, et al.
Veröffentlicht: (2024)
Region-Constraint In-Context Generation for Instructional Video Editing
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Zhongwei, et al.
Veröffentlicht: (2025)
DreamVAR: Taming Reinforced Visual Autoregressive Model for High-Fidelity Subject-Driven Image Generation
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
von: Jiang, Xin, et al.
Veröffentlicht: (2026)
DiffuseST: Unleashing the Capability of the Diffusion Model for Style Transfer
von: Hu, Ying, et al.
Veröffentlicht: (2024)
von: Hu, Ying, et al.
Veröffentlicht: (2024)
A Tri-Dynamic Preprocessing Framework for UGC Video Compression
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
von: Zhao, Fei, et al.
Veröffentlicht: (2025)
Language-Guided Diffusion Model for Visual Grounding
von: Chen, Sijia, et al.
Veröffentlicht: (2023)
von: Chen, Sijia, et al.
Veröffentlicht: (2023)
M&M VTO: Multi-Garment Virtual Try-On and Editing
von: Zhu, Luyang, et al.
Veröffentlicht: (2024)
von: Zhu, Luyang, et al.
Veröffentlicht: (2024)
MagicTryOn: Harnessing Diffusion Transformer for Garment-Preserving Video Virtual Try-on
von: Li, Guangyuan, et al.
Veröffentlicht: (2025)
von: Li, Guangyuan, et al.
Veröffentlicht: (2025)
GGAvatar: Reconstructing Garment-Separated 3D Gaussian Splatting Avatars from Monocular Video
von: Chen, Jingxuan
Veröffentlicht: (2024)
von: Chen, Jingxuan
Veröffentlicht: (2024)
DIVE: Inverting Conditional Diffusion Models for Discriminative Tasks
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
von: Li, Yinqi, et al.
Veröffentlicht: (2025)
FD2Talk: Towards Generalized Talking Head Generation with Facial Decoupled Diffusion Model
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
von: Yao, Ziyu, et al.
Veröffentlicht: (2024)
BiTDiff: Fine-Grained 3D Conducting Motion Generation via BiMamba-Transformer Diffusion
von: Jia, Tianzhi, et al.
Veröffentlicht: (2026)
von: Jia, Tianzhi, et al.
Veröffentlicht: (2026)
2D or 3D: Who Governs Salience in VLA Models? -- Tri-Stage Token Pruning Framework with Modality Salience Awareness
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
von: Zheng, Zihao, et al.
Veröffentlicht: (2026)
Palmprint De-Identification Using Diffusion Model for High-Quality and Diverse Synthesis
von: Yan, Licheng, et al.
Veröffentlicht: (2025)
von: Yan, Licheng, et al.
Veröffentlicht: (2025)
JIGMARK: A Black-Box Approach for Enhancing Image Watermarks against Diffusion Model Edits
von: Pan, Minzhou, et al.
Veröffentlicht: (2024)
von: Pan, Minzhou, et al.
Veröffentlicht: (2024)
A Novel Approach to Industrial Defect Generation through Blended Latent Diffusion Model with Online Adaptation
von: Li, Hanxi, et al.
Veröffentlicht: (2024)
von: Li, Hanxi, et al.
Veröffentlicht: (2024)
Improving Multi-modal Large Language Model through Boosting Vision Capabilities
von: Sun, Yanpeng, et al.
Veröffentlicht: (2024)
von: Sun, Yanpeng, et al.
Veröffentlicht: (2024)
DreamJourney: Perpetual View Generation with Video Diffusion Models
von: Pan, Bo, et al.
Veröffentlicht: (2025)
von: Pan, Bo, et al.
Veröffentlicht: (2025)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
von: Liu, Luping, et al.
Veröffentlicht: (2024)
von: Liu, Luping, et al.
Veröffentlicht: (2024)
Error Analyses of Auto-Regressive Video Diffusion Models: A Unified Framework
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
Exploring Phrase-Level Grounding with Text-to-Image Diffusion Model
von: Yang, Danni, et al.
Veröffentlicht: (2024)
von: Yang, Danni, et al.
Veröffentlicht: (2024)
ChatVTG: Video Temporal Grounding via Chat with Video Dialogue Large Language Models
von: Qu, Mengxue, et al.
Veröffentlicht: (2024)
von: Qu, Mengxue, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Incorporating Visual Correspondence into Diffusion Model for Virtual Try-On
von: Wan, Siqi, et al.
Veröffentlicht: (2025) -
Improving Text-guided Object Inpainting with Semantic Pre-inpainting
von: Chen, Yifu, et al.
Veröffentlicht: (2024) -
HIRI-ViT: Scaling Vision Transformer with High Resolution Inputs
von: Yao, Ting, et al.
Veröffentlicht: (2024) -
Boosting Diffusion Models with Moving Average Sampling in Frequency Domain
von: Qian, Yurui, et al.
Veröffentlicht: (2024) -
SD-DiT: Unleashing the Power of Self-supervised Discrimination in Diffusion Transformer
von: Zhu, Rui, et al.
Veröffentlicht: (2024)