Voost: A Unified and Scalable Diffusion Transformer for Bidirectional Virtual Try-On and Try-Off
Fuente:
arXiv
Guardado en:
| Autores principales: | Lee, Seungyong, Kwak, Jeong-gi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Masked Extended Attention for Zero-Shot Virtual Try-On In The Wild
por: Orzech, Nadav, et al.
Publicado: (2024)
por: Orzech, Nadav, et al.
Publicado: (2024)
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers
por: Dahary, Omer, et al.
Publicado: (2026)
por: Dahary, Omer, et al.
Publicado: (2026)
Frankenstein: Generating Semantic-Compositional 3D Scenes in One Tri-Plane
por: Yan, Han, et al.
Publicado: (2024)
por: Yan, Han, et al.
Publicado: (2024)
Street TryOn: Learning In-the-Wild Virtual Try-On from Unpaired Person Images
por: Cui, Aiyu, et al.
Publicado: (2023)
por: Cui, Aiyu, et al.
Publicado: (2023)
BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation
por: Wu, Zhennan, et al.
Publicado: (2024)
por: Wu, Zhennan, et al.
Publicado: (2024)
Hierarchical Hybrid Sliced Wasserstein: A Scalable Metric for Heterogeneous Joint Distributions
por: Nguyen, Khai, et al.
Publicado: (2024)
por: Nguyen, Khai, et al.
Publicado: (2024)
End-to-End Training for Unified Tokenization and Latent Denoising
por: Duggal, Shivam, et al.
Publicado: (2026)
por: Duggal, Shivam, et al.
Publicado: (2026)
Gaussian Splashing: Unified Particles for Versatile Motion Synthesis and Rendering
por: Feng, Yutao, et al.
Publicado: (2024)
por: Feng, Yutao, et al.
Publicado: (2024)
DiffusionBrowser: Interactive Diffusion Previews via Multi-Branch Decoders
por: Hong, Susung, et al.
Publicado: (2025)
por: Hong, Susung, et al.
Publicado: (2025)
GeneOH Diffusion: Towards Generalizable Hand-Object Interaction Denoising via Denoising Diffusion
por: Liu, Xueyi, et al.
Publicado: (2024)
por: Liu, Xueyi, et al.
Publicado: (2024)
Navigating with Annealing Guidance Scale in Diffusion Space
por: Yehezkel, Shai, et al.
Publicado: (2025)
por: Yehezkel, Shai, et al.
Publicado: (2025)
Neural Isometries: Taming Transformations for Equivariant ML
por: Mitchel, Thomas W., et al.
Publicado: (2024)
por: Mitchel, Thomas W., et al.
Publicado: (2024)
Controlling Text-to-Image Diffusion by Orthogonal Finetuning
por: Qiu, Zeju, et al.
Publicado: (2023)
por: Qiu, Zeju, et al.
Publicado: (2023)
Infinite-Resolution Integral Noise Warping for Diffusion Models
por: Deng, Yitong, et al.
Publicado: (2024)
por: Deng, Yitong, et al.
Publicado: (2024)
ComboStoc: Combinatorial Stochasticity for Diffusion Generative Models
por: Xu, Rui, et al.
Publicado: (2024)
por: Xu, Rui, et al.
Publicado: (2024)
TryOffDiff: Virtual-Try-Off via High-Fidelity Garment Reconstruction using Diffusion Models
por: Velioglu, Riza, et al.
Publicado: (2024)
por: Velioglu, Riza, et al.
Publicado: (2024)
TAUE: Training-free Noise Transplant and Cultivation Diffusion Model
por: Nagai, Daichi, et al.
Publicado: (2025)
por: Nagai, Daichi, et al.
Publicado: (2025)
PASTA: Controllable Part-Aware Shape Generation with Autoregressive Transformers
por: Li, Songlin, et al.
Publicado: (2024)
por: Li, Songlin, et al.
Publicado: (2024)
Diffusion Self-Distillation for Zero-Shot Customized Image Generation
por: Cai, Shengqu, et al.
Publicado: (2024)
por: Cai, Shengqu, et al.
Publicado: (2024)
Hollowed Net for On-Device Personalization of Text-to-Image Diffusion Models
por: Cho, Wonguk, et al.
Publicado: (2024)
por: Cho, Wonguk, et al.
Publicado: (2024)
Size-Variable Virtual Try-On with Physical Clothes Size
por: Yamashita, Yohei, et al.
Publicado: (2024)
por: Yamashita, Yohei, et al.
Publicado: (2024)
LoMOE: Localized Multi-Object Editing via Multi-Diffusion
por: Chakrabarty, Goirik, et al.
Publicado: (2024)
por: Chakrabarty, Goirik, et al.
Publicado: (2024)
Detection-Driven Object Count Optimization for Text-to-Image Diffusion Models
por: Zafar, Oz, et al.
Publicado: (2024)
por: Zafar, Oz, et al.
Publicado: (2024)
Add-it: Training-Free Object Insertion in Images With Pretrained Diffusion Models
por: Tewel, Yoad, et al.
Publicado: (2024)
por: Tewel, Yoad, et al.
Publicado: (2024)
Layout-Corrector: Alleviating Layout Sticking Phenomenon in Discrete Diffusion Model
por: Iwai, Shoma, et al.
Publicado: (2024)
por: Iwai, Shoma, et al.
Publicado: (2024)
SALAD: Skeleton-aware Latent Diffusion for Text-driven Motion Generation and Editing
por: Hong, Seokhyeon, et al.
Publicado: (2025)
por: Hong, Seokhyeon, et al.
Publicado: (2025)
InfiniteDiffusion: Bridging Learned Fidelity and Procedural Utility for Open-World Terrain Generation
por: Goslin, Alexander
Publicado: (2025)
por: Goslin, Alexander
Publicado: (2025)
ArtiFixer: Enhancing and Extending 3D Reconstruction with Auto-Regressive Diffusion Models
por: de Lutio, Riccardo, et al.
Publicado: (2026)
por: de Lutio, Riccardo, et al.
Publicado: (2026)
RealmDreamer: Text-Driven 3D Scene Generation with Inpainting and Depth Diffusion
por: Shriram, Jaidev, et al.
Publicado: (2024)
por: Shriram, Jaidev, et al.
Publicado: (2024)
ArchComplete: Autoregressive 3D Architectural Design Generation with Hierarchical Diffusion-Based Upsampling
por: Rasoulzadeh, S., et al.
Publicado: (2024)
por: Rasoulzadeh, S., et al.
Publicado: (2024)
COLLAGE: Collaborative Human-Agent Interaction Generation using Hierarchical Latent Diffusion and Language Models
por: Daiya, Divyanshu, et al.
Publicado: (2024)
por: Daiya, Divyanshu, et al.
Publicado: (2024)
DiffH2O: Diffusion-Based Synthesis of Hand-Object Interactions from Textual Descriptions
por: Christen, Sammy, et al.
Publicado: (2024)
por: Christen, Sammy, et al.
Publicado: (2024)
DiT-VTON: Diffusion Transformer Framework for Unified Multi-Category Virtual Try-On and Virtual Try-All with Integrated Image Editing
por: Li, Qi, et al.
Publicado: (2025)
por: Li, Qi, et al.
Publicado: (2025)
M&M VTO: Multi-Garment Virtual Try-On and Editing
por: Zhu, Luyang, et al.
Publicado: (2024)
por: Zhu, Luyang, et al.
Publicado: (2024)
FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On
por: Karras, Johanna, et al.
Publicado: (2026)
por: Karras, Johanna, et al.
Publicado: (2026)
DreamCatalyst: Fast and High-Quality 3D Editing via Controlling Editability and Identity Preservation
por: Kim, Jiwook, et al.
Publicado: (2024)
por: Kim, Jiwook, et al.
Publicado: (2024)
Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers
por: Zheng, Shuhong, et al.
Publicado: (2026)
por: Zheng, Shuhong, et al.
Publicado: (2026)
ELMO: Enhanced Real-time LiDAR Motion Capture through Upsampling
por: Jang, Deok-Kyeong, et al.
Publicado: (2024)
por: Jang, Deok-Kyeong, et al.
Publicado: (2024)
RigidFormer: Learning Rigid Dynamics using Transformers
por: Dou, Zhiyang, et al.
Publicado: (2026)
por: Dou, Zhiyang, et al.
Publicado: (2026)
SAiD: Speech-driven Blendshape Facial Animation with Diffusion
por: Park, Inkyu, et al.
Publicado: (2023)
por: Park, Inkyu, et al.
Publicado: (2023)
Ejemplares similares
-
Masked Extended Attention for Zero-Shot Virtual Try-On In The Wild
por: Orzech, Nadav, et al.
Publicado: (2024) -
On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers
por: Dahary, Omer, et al.
Publicado: (2026) -
Frankenstein: Generating Semantic-Compositional 3D Scenes in One Tri-Plane
por: Yan, Han, et al.
Publicado: (2024) -
Street TryOn: Learning In-the-Wild Virtual Try-On from Unpaired Person Images
por: Cui, Aiyu, et al.
Publicado: (2023) -
BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation
por: Wu, Zhennan, et al.
Publicado: (2024)